Awesome Go
Section: Tokenizers · This is a Go implementation of jieba which a Chinese word splitting algorithm.
Entry
Appears in 5 awesome lists
This is a Go implementation of jieba which a Chinese word splitting algorithm.
Section: Tokenizers · This is a Go implementation of jieba which a Chinese word splitting algorithm.
Section: 分词器 · star:2616 这是一个Go实现的jieba,这是一个中文分词算法。
Section: Libraries · Go implementation of the jieba Chinese word segmentation algorithm.
Section: 搜索推荐 · "结巴"中文分词的 Go 语言版本
Section: Other · "结巴"中文分词的Golang版本
(formerly known as pytorch-transformers and pytorch-pretrained-bert) provides state-of-the-art general-purpose architectures (BERT, GPT-2, RoBERTa, XLM, DistilBert, XLNet, CTRL...) for Natural Language Understanding (NLU) and Natural Language Generation (NLG) with over 32+ pretrained models in…
Open-source AI orchestration framework for building context-engineered, production-ready LLM applications. Design modular pipelines and agent workflows with explicit control over retrieval, routing, memory, and generation. Built for scalable agents, RAG, multimodal applications, semantic search,…
A curated list of resources dedicated to Natural Language Processing and text processing for Ruby.
Industrial-strength natural language processing with 75+ languages, transformer pipelines, and production-grade NER, parsing, and text classification.
Hugging Face's tokenizers for modern NLP pipelines (original implementation) with bindings for Python.
Python framework for adversarial attacks, data augmentation, and model training in NLP. Augment datasets to increase model robustness and generate adversarial examples. MIT licensed.
The largest hub of ready-to-use NLP datasets for ML models with fast, easy-to-use and efficient data manipulation tools.