menu
Tatoeba
language
Registrar-se Entrar
language Português (Brasil)
menu
Tatoeba

chevron_right Registrar-se

chevron_right Entrar

Navegar

chevron_right Mostrar frase aleatória

chevron_right Navegar por idioma

chevron_right Navegar por lista

chevron_right Navegar por etiqueta

chevron_right Navegar por áudio

Comunidade

chevron_right Mural

chevron_right Lista de todos os membros

chevron_right Idiomas dos membros

chevron_right Falantes nativos

search
clear
swap_horiz
search
JimBreen JimBreen 21 de março de 2010 21 de março de 2010 06:11:23 UTC flag Report link Link permanente

Traditional and Simplified Chinese

I saw the comment about converting hanzi on-the-fly. Be very cautious about that, as there are many cases where it simply doesn't work. Proper Traditional<->Simplified conversion needs to work at the lexeme level and in some cases needs some context for disambiguation.

Jack Halpern wrote a very good paper about this about 10 years ago:
http://www.cjk.org/cjk/c2c/c2cbasis.htm

PS: how do I make a comment on another posting?

{{vm.hiddenReplies[377] ? 'expand_more' : 'expand_less'}} Ocultar respostas Mostrar respostas
JimBreen JimBreen 21 de março de 2010 21 de março de 2010 06:40:28 UTC flag Report link Link permanente

OK, I worked out how to do a follow-on. I'd clicked "reply" but it hadn't worked. Now it does.

sysko sysko 21 de março de 2010 21 de março de 2010 11:10:11 UTC flag Report link Link permanente

the traditional to simplified chinese is not made at "character by character" level, but try to decompose the sentence (you can see how the sentence has been segmented by looking to pinyin)
As I've said I'm in conctact with the guy who develop it, so don't hesitate to report any bad segmentations, I will report to him