Awesome Go
Section: Tokenizers · Go efficient text segmentation; support english, chinese, japanese and other.
Entry
Appears in 4 awesome lists
Go efficient multilingual NLP and text segmentation; support English, Chinese, Japanese and others.
Section: Tokenizers · Go efficient text segmentation; support english, chinese, japanese and other.
Section: 分词器 · star:2777 高效的文本分割;支持英语、汉语、日语等。
Section: 搜索推荐 · Go 语言分词
Section: Other · Go efficient multilingual NLP and text segmentation; support English, Chinese, Japanese and others.
Deprecated: Use the official Elasticsearch client for Go at https://github.com/elastic/go-elasticsearch
This is a Go implementation of jieba which a Chinese word splitting algorithm.
An offline recommender system backend based on collaborative filtering written in Go.
The official Go client for Elasticsearch
star:463 Sentence tokenizer: converts text into a list of sentences.
A tokenizer based on the dictionary and Bigram language models for Golang. (Now only support chinese segmentation)
This is a GO implementation of MMSEG which a Chinese word splitting algorithm.
Go package for n-gram based text categorization, with support for utf-8 and raw text.