Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

And ICU uses data from CLDR, which is mentioned in the blog. Here, there are 380 xml files: https://github.com/unicode-org/cldr/tree/main/common/transfo...

Yes, ICU is ubiquitous. But, some NLP projects use various other libraries, such as uroman (just for romanization - to Latin script).



Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: