Among many possible
deliverables of our project, we have in our working agenda at least
the following items :
i) Upgrading
of our two existing "syllabaries" :
Since its appearance
some 60 years ago, S.
L. Wong's classic A Chinese Syllabary Pronounced according to the
Dialect of Canton ( 黃錫凌《粵音韻彙》)
has become an indispensable reference work for any responsible Chinese
teacher in Hong Kong. However, the rather simplistic layout (initially
even without any index) prevented
this work from becoming more popularised. In the late seventies, the
Syllabary was overhauled with an indexing system added. Around
1995, when information technology became a catchword, we gradually came
up with the idea to have the Syllabary reprocessed electronically,
making its hidden wealth more accessible to users over the network.
This plan became realised when the de facto editor of the current
printed version of the Syllabary agreed to become one of our
collaborators. We used client/server, PERL scripts, javascripts and
multimedia techniques to generate the first on-line version of the Syllabary
on Internet. - Then in 1997, we constructed a second, much enlarged
syllabary A
Chinese Talking Syllabary of the Cantonese Dialect: An Electronic Repository
( 粵語音韻集成電子版 ).
This time we have been able to use some data provided by the Linguistic
Society of Hong Kong. Besides a much expanded database (from 8,000 entries
to 12,000), we have added further options and functions to the syllabary.
With the self-made search engine, the sound database, the many indices
and query options provided, and the character level linkages to other
character dictionaries all over the Internet, the Syllabary has injected
vital force in school language education as well as in higher level
research. - On the basis of what we have done, we are now planning to
incorporate more lexical information into our current syllabary databases
so as to transform the syllabary into an online talking lexicon in its
own right.
ii) Online
edition of Lin
Yu-Tang's Chinese-English Dictionary of Modern Usage ( 林語堂《當代漢英詞典》網上版
)
So many years after
its first publication back in 1972 by the Chinese University Press,
the impact of Lin
Yu-Tang's Chinese-English Dictionary of Modern Usage
is indisputably fading out. However, in regard of academic content and
practical value, the dictionary remains in many ways uncontested and
unsurpassed. Now with the permission of the Chinese University Press,
we plan to reprocess the dictionary material into a network-accessible
lexical tool suitable for today's use. Instead of just mechanically
inputting or copy-editing the dictionary text for public access, what
we plan to do will involve advance information engineering procedures
which will add new features to the dictionary on the one hand and allow
hidden potentials of the dictionary to be exhausted on the other. To
name a few examples, the following features, not thinkable in a paper
version of the dictionary, will be available by a simple mouse click:
option of searching for an English instead of for a Chinese expression;
searching for Chinese expressions in successive (順序搜索)as well as in
reverse order (逆序搜索) searching words or lexical structures belonging
to a particular meaning category (eg. figurative, Buddhistic, mythological,
chemical, colloquial, derogatory etc.); choices between full text search
or headword search; optional display of Wade-Giles or Pinyin
transcription of Putonghua pronunciation; hyperlinkage of the
headwords to our other lexical resources, and last but not least, all
possible Boolean search operations etc. With all these features added,
I am sure that the intrinsic value of Lin's dictionary will not just
be revived, but will be brought to even greater heights. As the first
and only one online lexical tool of this type in the foreseeable future,
the dictionary will certainly play an important role in the promotion
of biliteracy in Hong Kong. - To ensure smooth and systematic processing
of our work, we have already written a program (in PERL) tailored exclusively
for the editing of Lin's dictionary. This electronic workbench will
also provide basis for the preparation of the second edition of the
dictionary which is now being considered by the CU Press.
iii) An
Integrated Database for Chinese Language Usage:
As is widely known,
written symbols, spoken words, and meaning ( 形、音、義 ) are the three major
elements governing the use of language. As far as language didactics
is concerned, the teaching of the Chinese language (and all her dialects)
faces one peculiar problem which does not constitute an issue in the
teaching of Indo-European languages. In Indo-European languages, writing,
speech sound, and meaning exhibit a predominantly rectilinear relationship,
with writing automatically mapped to the speech, speech to our intended
meaning. As a result of this, the reading of a text is a relatively
easy job for an Indo-European child, as long as he/she has mastered
the basic pronunciation skills. But the situation in the teaching and
learning of the Chinese language is a completely different one. Unlike
the case in Western languages, writing, speech sound, and meaning in
Chinese exhibit a triangular relationship. Given a written sign
(character) in Chinese, there is no easy guideline telling the learner
how it should be pronounced., and given a character together with the
correct pronunciation, we do not necessarily know its meaning. For a
Chinese child to become literate, thousands of characters have to be
learned in respect of both their written and sound form as well as their
semantic content. Without detriment to the many structural advantages
of the Chinese language, which have arouse much research interest nowadays,
it is precisely this peculiarly triangular, i.e. non-rectilinear relationship
which has contributed to the higher rate of illiteracy among the Chinese
population. - It is out of this concern that an "integrated database
for Chinese language usage" could be of help in the teaching and
learning of the Chinese language. The idea behind such a database is
the use of modern IT and multimedia to "tighten up" the otherwise
loose relationship between writing, speech sound and meaning (i.e. orthography,
phonetics, and semantics) of the Chinese characters. The database will
be implemented stage by stage by integrating the computational results
of the above named plans
[as described in i) and ii)] with
further modules appended. Modules to be added to the final database
will include the production of Putonghua sound files, text-to-speech
generation for both Cantonese and Putonghua, information on etymological
usage, genealogy of script forms, an archive of related educational
images, further linkage options to external resources etc. With all
these components in place, the resultant database will be of great educational
value not only to local users, but also to a huge population of users
around the globe. Given the current popularity of our web services,
I can foresee that our Integrated Database will very soon become a major
attraction of the internet community, with impact comparable to internationally
renowned sites such as the Perseus Project (Greek) of Tufts or
the WordNet (English) of Princeton.
iv) Topic-specific
Web pages for language education in the Hong Kong Context:
Besides more ambitious
plans as given above, we will from time to time construct web pages
related to specific topics of language education in the Hong Kong context.
One possible example of jobs of this kind is a repository of commonly
made mistakes in pronunciation and compositions, Chinese as well as
English. In this connection, we can invite input from school teachers
and benefit from their real life teaching experience. Another possible
example is a report on the Cantonese-English bilingual language behaviour
which is typical of some children growing up in Hong Kong. This kind
of observation might shed light on questions such as language differentiation,
comparison between monolingual and bilingual development, possibility
of delay and degree of balance etc.