TAUS Language Translation Data | Parallel translation for Colloquial English into various languages for Machine Learning
收藏Datarade2024-04-19 收录
下载链接:
https://datarade.ai/data-products/taus-parallel-text-colloquial-domain-english-low-resource-see-description-taus
下载链接
链接失效反馈官方服务:
资源简介:
The corpus is a great fit for training chat bots or social media content, and will give the conversation with your local audience a friendly, casual tone. From product user reviews and blog post comments to everyday business small talk, your MT engine will be able to handle even the most creative user voices. This corpus contains over 1 million words, and a total vocabulary of more than 37000 different words. Need more data? In the following months, TAUS will release more equally sized corpora for the same domain and language combinations, with a significant increase of vocabulary. English - Hindi English - Urdu English - Tamil English - Nepali English - Turkish English - Pashto English - Sorani English - Bengali English - Burmese English - Assamese English - Telugu English - Sinhalese English - Dari English - Punjabi (Pakistan) English - Punjabi (India) English - Lao English - Kurmanji (lat) English - Kurmanji (arab) Other languages are available on demand.
提供机构:
TAUS



