Wall (7.482 threads)
Савети
Before asking a question, make sure to read the FAQ.
We aim to maintain a healthy atmosphere for civilized discussions. Please read our rules against bad behavior.
AlanF_US
пре 8 сати
Synonyms
јуче
Thanuir
пре 3 дана
soweli
пре 3 дана
gillux
пре 5 дана
AlanF_US
пре 6 дана
maaster
пре 6 дана
AlanF_US
пре 6 дана
maaster
пре 9 дана
Qaztat
пре 12 дана
Hi everyone! I'd like to share a small project that uses Tatoeba.
Synopedia (https://www.synopedia.com) is a free thesaurus that groups synonyms by meaning, with antonyms and short definitions, in English, French, Spanish, Italian, German and Catalan. On the English word pages, the example sentences come from Tatoeba (CC BY 2.0 FR), credited on every page and on the sources page: https://www.synopedia.com/sources
Example: https://www.synopedia.com/synonym/bright
Thank you to all the contributors, the sentences really help show how a word is used in context. If a wiki editor thinks it fits, I'd be happy to see it added to the "Projects using Tatoeba" page (I couldn't get a wiki account). Feedback welcome too!
I added it to the wiki page. Glad to see Tatoeba credited correctly on the site.
I see that you have trend lines in the style of Google Ngram Viewer that show the rise or fall in usage of the various synonyms over time. However, it's not possible to compare the frequency across the synonyms as it would be on a Google Ngram Viewer page that showed them all together. The ability to go directly to Google Ngram Viewer (if that's allowed by its license) would also be useful because it would let the user see many more examples of the words in context.
I linked two sentences incorrectly (#4694466 and #65981) because I was going too fast and misread a word ("pini" as "pali"). I'll be more careful in the future but for now can someone with permissions unlink the two? Thanks :-)
Done. Please check it was the correct sentence.
Hello, Tatoeba team. A few months ago, I shared sentences in Tawallammat Tamajaq with you—translated into English and Arabic—across two files. Could you please let me know why those sentences haven't been added yet?
Because the amount of work required to do that is huge, and there are plenty of other things to do already, and Tatoeba has very, very few resources. I understand that you would like the sentence pairs to be added, but at the moment the most useful thing you can do is to be patient, or to get involved and do the work yourself, or to pay somebody to do the work, or to donate money to Tatoeba.
Could someone unblock the sentences with the tag "unblock"?
Sentences are blocked for a variety of reasons, among them:
(1) they violate a copyright
(2) their content violates community guidelines
(3) they are posted by someone who has written a large number of sentences that belong to categories 1 or 2 and that therefore we can't trust
(4) they are translation of sentences in the other three categories
These characteristics don't change simply with the passage of time.
As I've said before, I really hope we're not so short of ideas for sentences that we need to turn to such problematic sources.
I added some tags "unblock" to some blocked sentences.
They're simply sentences and are O.K.
(I suppose–I'm not sure–the sentences tegged by other ones are also O.K.)
For the same reason that you're not sure whether sentences tagged by other people are acceptable, we can't be sure whether the sentences you tagged are acceptable, and it's not feasible for us to look into this further.
Садржај ове реченице није у складу са нашим правилима и зато је сакривен. Реченицу виде само администратори и њен аутор.
Minemçä, saytta inde cömlälärne tatar telenä kirill häm latin grafikasında avtomat räweştä tärcemä itü öçen quldan kertelgän küp cömlälär bar. Bu citärlek. Monnan tış, yullar, ısullar, kodlar häm başqa mäğlümat birelä.
Минемчә, сайтта инде җөмләләрне татар теленә кирилл һәм латин графикасында автомат рәвештә тәрҗемә итү өчен кулдан кертелгән күп җөмләләр бар. Бу җитәрлек. Моннан тыш, юллар, ысуллар, кодлар һәм башка мәгълүмат бирелә.
In my opinion, in order to automate the translation of sentences into Tatar in two versions - Cyrillic and Latin, there are already many sentences entered manually on the site. That's quite enough. In addition, paths, methods, codes, and more are indicated.
💯
Monda, asta, barısı da inde çäynäp beterelgän kebek toyıla.
Монда, аста, барысы да инде чәйнәп бетерелгән кебек тоела.
Әгәр соравыгыз татар телендә кириллицадан латиницага автомат рәвештә күчерү турында булса, моны башкару өчен берничә ысул бар:
Онлайн конвертерлар: Төрле веб-сайтлар татар телендәге текстны кириллицадан латиницага күчерү мөмкинлеге бирә. Мисал өчен, google transliterate инструментлары яки махсус татарча онлайн транслит сайтлары кулланырга мөмкин.
Автоматизация өчен скриптлар: Python кебек телләрдә махсус китапханәләр кулланып, текстны берничә кагыйдә буенча кириллицадан латиницага автомат күчереп була. Мәсәлән, таблица-карта кулланып һәр хәрефне латинча аналогына алыштыру.
MS Word яки Google Docs макросы: Әгәр күп күләмле документлар белән эшлисез икән, макрос язып, һәр хәрефне автомат рәвештә кириллицадан латиницага алыштырырга була.
Тел исәпкә алып, транслитерацияне башкарганда татар фонетикасын саклау мөһим, чөнки турыдан-туры алыштыру кайвакыт дөрес итеп укуны тәэмин итмәскә мөмкин.
Әгәр теләсәгез, мин сезгә тулысынча автоматлаштырылган Python скрипт үрнәген китерә алам, ул кириллицадагы татар текстын латиницага күчерәчәк. Сез шуны кулланып, теләсә кайсы текстны берничә секунд эчендә әйләндерә аласыз.
Бу мисалда скрипт татар телендә кириллицада язылган текстны укый, барлык сүзләрне тәкъдим итә, һәр сүзнең озынлыгын саный һәм нәтиҗәне автомат рәвештә файлга саклый.
# Python 3
# Тулы автоматлаштырылган татар текстын эшкәртү скрипты
import os
# Текстны уку функциясе
def read_tatar_text(file_path):
with open(file_path, 'r', encoding='utf-8') as f:
text = f.read()
return text
# Сүзләрне аеру һәм саннарын санау
def process_text(text):
# Барлык символларны түбән регистрга күчерү
text = text.lower()
# Сүзләрнең исемлеген булдыру
words = []
current_word = ''
for ch in text:
if ch.isalpha() or ch == 'ә' or ch == 'җ' or ch == 'ү' or ch == 'ң' or ch == 'ө' or ch == 'һ' or ch == 'ӓ' or ch == 'ӱ' or ch == 'ҫ':
current_word += ch
else:
if current_word:
words.append(current_word)
current_word = ''
# Ахыргы сүзне өстәргә онытмагыз
if current_word:
words.append(current_word)
# Ә һәр сүзнең озынлыгын санау
word_lengths = {word: len(word) for word in words}
return words, word_lengths
# Нәтиҗәне саклау
def save_results(words, word_lengths, output_file):
with open(output_file, 'w', encoding='utf-8') as f:
f.write("Сүзләр исемлеге:
")
f.write(', '.join(words) + '
')
f.write("Сүз озынлыклары:
")
for word, length in word_lengths.items():
f.write(f"{word}: {length}
")
# Башкару
if __name__ == "__main__":
input_file = 'tatar_text.txt' # Кириллицадагы татар текстын монда куегыз
output_file = 'processed_tatar_text.txt'
if not os.path.exists(input_file):
print(f"Файл '{input_file}' табылмады. Зинһар өчен текстны урнаштырыгыз.")
else:
text = read_tatar_text(input_file)
words, word_lengths = process_text(text)
save_results(words, word_lengths, output_file)
print(f"Эшкәртү тәмамланды! Нәтиҗә '{output_file}' файлыгында сакланды.")
Куллану үрнәге:
tatar_text.txt исемле файлга татар телендәге текстны кертәсез.
Скриптны эшләтеп җибәрәсез:
python process_tatar_text.py
Нәтиҗә processed_tatar_text.txt файлында саклана: сүзләрнең исемлеге һәм аларның озынлыклары.
Бу скрипт тулы автоматлаштырылган рәвештә файл уку, текстны эшкәртү һәм нәтиҗәләрне саклау функцияләрен башкара. Кириллицадагы барлык татар хәрефләрен дә таный һәм сүзләрне дөрес аера.
Çit illärdäge tatarlarğa, kirill älifbasınnan tış, latinçası da kiräk.
(Чит илләрдәге татарларга, кирилл әлифбасыннан тыш, латинчасы да кирәк.)
Интернеттан:
Чит илләрдә яшәүче татарлар өчен латин әлифбасы (Zamanälif яки башка вариантлар) бик мөһим, чөнки:Техник уңайлылык: Күпчелек чит ил санаклары һәм телефоннарында кирилл клавиатурасы булмаска мөмкин. Татар латин әлифбасы исә стандарт инглиз яки төрек клавиатурасына җиңел җайлаша.Телне өйрәнү: Диспорада туып-үскән яшьләр өчен латин хәрефләре танышрак, шуңа күрә алар телне латинча укып тизрәк үзләштерә.Төрки телләр мохите: Төркиядә яки башка телне латинчага күчергән төрки илләрдә яшәүче татарларга бу формат кабул итү өчен күпкә якынрак.Татоэба кебек халыкара платформаларда җөмләләрне ике вариантта да (кирилл һәм латин) бирү — телне дөньякүләм дәрәҗәдә саклау һәм тарату өчен бик файдалы адым.