2 results found Sort:

9
209
apache-2.0
6
一个 Golang 实现的相对智能、无需规则维护的通用新闻网站数据提取工具库。含域名探测、网页编码语种识别、网页链接分类提取、网页新闻要素抽取以及新闻正文抽取等组件。
Created 2022-07-15
265 commits to main branch, last one 3 months ago
✨ Split text by languages (e.g. 你喜欢看アニメ吗 -> 你喜欢看 | アニメ | 吗) for NLP tasks (e.g. parse, TTS). Powered by fasttext and langua
Created 2024-06-28
117 commits to main branch, last one about a month ago