CEV (CEDICT viewer) v0.5 readme file
This is an application for using CEDICT-style formatted files as chinese<->english dictionaries.
Usage is rather simple and I think you will get a point by yourself. Just will say some tips:
To use any CEDICT file:
You must be sure that it's in UCS-2 encoding and with .ced extension. For example you have cedict.txt file. Rename it to cedict.ced and place in application's folder. Run cev.exe. Then select Dict ->Create index for CEDICT file from program's main menu. After index creation - fill it's name (it will be a dic's title) and then restart application so it can use this index.
NOTE:
a) you can check (or uncheck) Dict ->Also add english/russian words to index option to make that index searchable for english/russian in addition to default chinese. Only words that contains more than 3 eng/rus characters will be added to index.
b) you can check (or uncheck) Dict ->Also add pinyin to index option to make that index searchable for pinyin syllables. Program will skip numbers in pinyin indexing (so that "ma4" and "ma1" will be same). Search results will depend on pinyin string in CEDICT file. In instance - if you have "mo4 sheng1" string in pinyin field in CEDICT file - then to find it - you must enter "mo sheng" exactly and if you will enter "mo4 sheng" or "mo sheng1" or even "mo sheng" (more than one space character inside) - then search will fail. And if you have "mo4sheng1" pinyin string in CEDICT file - then you must enter "mosheng" to find this entry.
c) CEDICT file can be more complex than original format. Example of string is here: word word2 word3 ... wordX [transcription] /translation1 | examples1/ translation2 | examples2/ ... / where 'word' is a chinese word, 'word2', 'word3' - it's another forms (like traditional, simplified and whatever you want, but it will make index only for 1st of them), 'transcription' - a pinyin string, also can be with tone marks instead of numbers, 'translation' and 'examples' can also contain some tags: [br] - new line, [is]Text[ie] - italic green, [bs]Text[be] - bold.
You can use Unihan database for searching info about characters:
To do so - you must download an Unihan database text file (unihan.txt) from Unicode (www.unicode.org). Place it in program's folder, then select Dict -> Create index for Unihan. After index creation - restart program. (Also be sure to have View -> Show Unihan info checked). Then it will show info for first chinese character in input.
You can create or modify your own CEDICT-styled dictionaries:
Use Dict -> Edit user dictionary file option. But before do this - ensure that you have userdict.txt file (also just a text file in UCS-2) in program's folder. This file will be your user dictionary.
Searching for words in entire text:
Use Settings -> Translate text (instead of word) - if this option unchecked - then program will try to translate only one word that was entered in the Input field. If you check this option - then program will try to translate all possible words from entire phrase. This feature will not work when you entered non-chinese character somewhere in the input field.
Searching both traditional and simplified words:
If you check Settings -> Search both chinese forms - then program will automatically search for both simplified and traditional forms of chinese characters regardless of entered or CEDICT file's form. In instance if you will enter a simplified character, and in CEDICT file you have a traditional one - it will be found and vice versa. Just I am not that sure will it always work correctly. So - please write to me if you will find wrongly mappings of simplified to traditional characters.
Please - visit www.mandarintools.com/cedict page for information about CEDICT project.
No warranty for this software. Use it only on your own risk!
/// From Russia with love ... and sorry for my awful english here ;)
COPYRIGHT AND LICENSE INFORMATION:
WARNING: following license statement applies ONLY to files of CEV project: CEV.EXE, EMPTY.HTML and ... this README.HTML.
1. Permissions are granted to use, modify and distribute above-mentioned CEV project files for any (even commercial) purposes.
RUSSIAN: (а теперь на великом и могучем ...)
Это небольшой софт для использования CEDICT-файлов в качестве словарей (как правило китайско-английских). Собственно поэтому вся небольшая инфа (вверху) и идет на английском (несмотря на то что я его не знаю ... хотя я и в русском не особо ...).
Использование очень простое (как чьи-то тапки), поясню только несколько стремных моментов:
1. Скачиваем словарную базу в формате CEDICT (в кодировке UCS-2). Вроде на моем сайте была одна такая - можно ее например. Файл должен быть с расширением .ced
2. Скармливаем базу программе. Кормежка осуществляется вот так: Копируем файлик в каталог программы. Выбираем Dict -> Create index for CEDICT file ... некоторое время (в зависимости от размера файла и скорости тачки) можно курить. Иногда программа ругается на ошибки в файле CEDICT (я там ввел кое-какие ограничения на размер слова, размер перевода и т.д.) - можно не обращать внимания (просто такие слова будут выброшены из индекса). Когда индексный файл будет создан - можно перезагрузить софтину чтобы она его "подхватила".
Обратите внимание на пункт меню Dict -> Also add english/russian words to index - если его выбрать - то в индекс будут добавлены английские и русские слова (размером больше 3-х символов). То есть можно будет искать не только по китайским словам, но и по англо-русским ... Также перед созданием индекса можно поставить чек на Also add pinyin to index чтобы прога добавила туда пиньинь, ну и соответственно потом по нему можно будет что-либо найти. Правда есть некоторые особенности. Например если в файле CEDICT написан такой пиньинь [mo4 sheng1] - тогда его можно будет найти написав "mo sheng", то есть циферки или тоновые значки в индекс не пойдут. А если в файле будет [mo4sheng1] или [mosheng] - тогда его можно отыскать набрав "mosheng".
3. Программа также может питаться файлами от Юникода (Unicode), а именно юниханьской базой данных (Unihan database). Скачивается такой файл (unihan.txt) с юникодного сайта. (www.unicode.org). Потом процедура кормежки аналогичная описанной выше, только надо использовать пункт меню Create index for Unihan. Также перезапускаем софт и вперед. Только главное не забыть поставить чек на View -> Show Unihan info.
4. Можно самому сделать (или подредактировать) файлик формата CEDICT, выбрав Dict -> Edit user dictionary file.
5. Если выбран пункт Translate text (instead of word) - тогда программа сделает попытку перевести все слова идущие подряд в введенном тексте, иначе будет переводить только одно слово.
6. Если выбран пункт Search both chinese forms - тогда софтине будет по барабану вводите ли Вы упрощенные или сложные иероглифы, а также все равно какие будут использоваться в CEDICT файле - она все их попытается отыскать ... Возможны глюки при выборе этого пункта поскольку не было времени (да и лень собственно) проверять файл перекодировки традиц.<->упрощ. который я откопал для таких целей.
7. Можно проверить CEDICT файл на наличие повторяющихся слов. Жмем на Dict -> Duplicates in CEDICT file. При выбранном пункте Also compare pinyin слова будут считаться совпадающими если у них также в точности совпадает пиньинь. Открываем файл (Open file), ждем некоторое время пока программа его прочитает (для отмены можно нажать Cancel). Ну и далее либо отмечаем в таблице ненужные записи, либо их редактируем (двойной клик) и жмем Save changes. Программа запишет новый файл в соответствии с изменениями. Файл будет иметь такое же имя, но с расширением .new
КОПИРАЙТ:
(!) Применяется только к файлам данной программы! То есть к cev.exe, empty.html, setup.ini и к данному readme.html.
1. Эту программу можно копировать, распространять, изменять в любых (даже коммерческих) целях. Информировать автора о вышеперечисленных действиях необязательно (да и не нужно).
ГАРАНТИИ:
Никаких гарантий на правильность работы нет! Используйте данный софт только на свой страх и риск!
С уважением ко всем изучающим китайский язык ...
... пишите: chamine@chamine.ru