Fasttext Open Source Projects
Browse 90 Fasttext open source projects, ranked by GitHub stars. Find the most popular Fasttext tools and libraries.
facebookresearch/fastText
Library for fast text representation and classification.
Metrics details
| Stars | 26,549 |
piskvorky/gensim
Topic Modelling for Humans
Metrics details
| Stars | 16,464 |
brightmart/text_classification
all kinds of text classification models and more with deep learning
Metrics details
| Stars | 7,941 |
DA-southampton/NLP_ability
总结梳理自然语言处理工程师(NLP)需要积累的各方面知识,包括面试题,各种基础知识,工程能力等等,提升核心竞争力
Metrics details
| Stars | 7,511 |
bentrevett/pytorch-sentiment-analysis
Tutorials on getting started with PyTorch and TorchText for sentiment analysis.
Metrics details
| Stars | 4,610 |
facebookresearch/MUSE
A library for Multilingual Unsupervised or Supervised word Embeddings
Metrics details
| Stars | 3,248 |
songyingxin/NLPer-Interview
该仓库主要记录 NLP 算法工程师相关的面试题
Metrics details
| Stars | 2,762 |
duoergun0729/nlp
兜哥出品 <一本开源的NLP入门书籍>
Metrics details
| Stars | 2,490 |
shibing624/python-tutorial
Python实用教程,包括:Python基础,Python高级特性,面向对象编程,多线程,数据库,数据科学,Flask,爬虫开发教程。
Metrics details
| Stars | 2,455 |
Kyubyong/wordvectors
Pre-trained word vectors of 30+ languages
Metrics details
| Stars | 2,234 |
Tencent/NeuralNLP-NeuralClassifier
An Open-source Neural Hierarchical Multi-label Text Classification Toolkit
Metrics details
| Stars | 1,921 |
yongzhuo/Keras-TextClassification
中文长文本分类、短句子分类、多标签分类、两句子相似度(Chinese Text Classification of Keras NLP, multi-label classify, or sentence classify, long or short),字词句向量嵌入层(embeddings)和网络层(graph)构建基类,FastText,TextCNN,CharCNN,TextRNN, RCNN, DCNN, DPCNN, VDCNN, CRNN, Bert, Xlnet, Albert, Attention, DeepMoji, HAN, 胶囊网络-CapsuleNet, Transformer-encode, Seq2seq, SWEM, LEAM, TextGCN
Metrics details
| Stars | 1,812 |
plasticityai/magnitude
A fast, efficient universal vector embedding utility package.
Metrics details
| Stars | 1,665 |
msgi/nlp-journey
Documents, papers and codes related to Natural Language Processing, including Topic Model, Word Embedding, Named Entity Recognition, Text Classificatin, Text Generation, Text Similarity, Machine Translation),etc.
Metrics details
| Stars | 1,629 |
wangyuGithub01/Machine_Learning_Resources
:fish::fish::fish: 机器学习面试复习资源
Metrics details
| Stars | 1,239 |
chenyuntc/PyTorchText
1st Place Solution for Zhihu Machine Learning Challenge . Implementation of various text-classification models.(知乎看山杯第一名解决方案)
Metrics details
| Stars | 1,057 |
iphysresearch/TOP250movie_douban
TOP250豆瓣电影短评:Scrapy 爬虫+数据清理/分析+构建中文文本情感分析模型
Metrics details
| Stars | 1,020 |
brightmart/bert_language_understanding
Pre-training of Deep Bidirectional Transformers for Language Understanding: pre-train TextCNN
Metrics details
| Stars | 965 |
miguelgfierro/sciblog_support
Support content for my blog
Metrics details
| Stars | 864 |
curiosity-ai/catalyst
🚀 Catalyst is a C# Natural Language Processing library built for speed. Inspired by spaCy's design, it brings pre-trained models, out-of-the box support for training word and document embeddings, and flexible entity recognition models.
Metrics details
| Stars | 854 |
ShawnyXiao/TextClassification-Keras
Text classification models implemented in Keras, including: FastText, TextCNN, TextRNN, TextBiRNN, TextAttBiRNN, HAN, RCNN, RCNNVariant, etc.
Metrics details
| Stars | 812 |
MaartenGr/PolyFuzz
Fuzzy string matching, grouping, and evaluation.
Metrics details
| Stars | 800 |
mayabot/mynlp
一个生产级、高性能、模块化、可扩展的中文NLP工具包。(中文分词、平均感知机、fastText、拼音、新词发现、分词纠错、BM25、人名识别、命名实体、自定义词典)
Metrics details
| Stars | 688 |
zlsdu/Word-Embedding
Word2vec, Fasttext, Glove, Elmo, Bert, Flair pre-train Word Embedding
Metrics details
| Stars | 653 |
dengxiuqi/WeiboSentiment
基于各种机器学习和深度学习的中文微博情感分析
Metrics details
| Stars | 635 |
oborchers/Fast_Sentence_Embeddings
Compute Sentence Embeddings Fast!
Metrics details
| Stars | 625 |
ncbi-nlp/BioSentVec
BioWordVec & BioSentVec: pre-trained embeddings for biomedical words and sentences
Metrics details
| Stars | 615 |
RandolphVI/Multi-Label-Text-Classification
About Muti-Label Text Classification Based on Neural Network.
Metrics details
| Stars | 561 |
loujie0822/Pre-trained-Models
预训练语言模型综述
Metrics details
| Stars | 546 |
neuml/codequestion
🔎 Semantic search for developers
Metrics details
| Stars | 541 |
AnubhavGupta3377/Text-Classification-Models-Pytorch
Implementation of State-of-the-art Text Classification Models in Pytorch
Metrics details
| Stars | 491 |
ThoughtRiver/lmdb-embeddings
Fast word vectors with little memory usage in Python
Metrics details
| Stars | 416 |
kermitt2/delft
a Deep Learning Framework for Text https://delft.readthedocs.io/
Metrics details
| Stars | 415 |
JackHCC/Chinese-Text-Classification-PyTorch
中文文本分类任务,基于PyTorch实现(TextCNN,TextRNN,FastText,TextRCNN,BiLSTM_Attention, DPCNN, Transformer,Bert,ERNIE),开箱即用!
Metrics details
| Stars | 401 |
dccuchile/spanish-word-embeddings
Spanish word embeddings computed with different methods and from different corpora
Metrics details
| Stars | 364 |
explosion/floret
🌸 fastText + Bloom embeddings for compact, full-coverage vectors with spaCy
Metrics details
| Stars | 343 |
chakki-works/chakin
Simple downloader for pre-trained word vectors
Metrics details
| Stars | 334 |
facebookresearch/DME
Dynamic Meta-Embeddings for Improved Sentence Representations
Metrics details
| Stars | 333 |
bureaucratic-labs/dostoevsky
Sentiment analysis library for russian language
Metrics details
| Stars | 322 |
sagorbrur/bnlp
BNLP is a natural language processing toolkit for Bengali Language.
Metrics details
| Stars | 309 |
apcode/tensorflow_fasttext
Simple embedding based text classifier inspired by fastText, implemented in tensorflow
Metrics details
| Stars | 302 |
brightmart/ai_law
all kinds of baseline models for long text classificaiton( text categorization)
Metrics details
| Stars | 292 |
vngrs-ai/vnlp
State-of-the-art, lightweight NLP tools for Turkish language. Developed by VNGRS.
Metrics details
| Stars | 288 |
ChenglongChen/tensorflow-XNN
4th Place Solution for Mercari Price Suggestion Competition on Kaggle using DeepFM variant.
Metrics details
| Stars | 285 |
lujiaying/MovieTaster-Open
A practical movie recommend project based on Item2vec.
Metrics details
| Stars | 279 |
dalinvip/cw2vec
cw2vec: Learning Chinese Word Embeddings with Stroke n-gram Information
Metrics details
| Stars | 274 |
PavelOstyakov/toxic
Toxic Comment Classification Challenge
Metrics details
| Stars | 266 |
vinhkhuc/JFastText
Java interface for fastText
Metrics details
| Stars | 248 |
Magic-Bubble/Zhihu
知乎看山杯 第二名 解决方案
Metrics details
| Stars | 241 |
lettergram/sentence-classification
Sentence Classifications with Neural Networks
Metrics details
| Stars | 237 |
shaypal5/skift
scikit-learn wrappers for Python fastText.
Metrics details
| Stars | 234 |
ChenglongChen/tensorflow-DSMM
Tensorflow implementations of various Deep Semantic Matching Models (DSMM).
Metrics details
| Stars | 230 |
vrasneur/pyfasttext
Yet another Python binding for fastText
Metrics details
| Stars | 224 |
vunb/vntk
Vietnamese NLP Toolkit for Node
Metrics details
| Stars | 219 |
natasha/navec
Compact high quality word embeddings for Russian language
Metrics details
| Stars | 218 |
moneyDboat/data_grand
2018达观杯文本智能处理挑战赛 Top10解决方案(10/3830)
Metrics details
| Stars | 214 |
amansrivastava17/embedding-as-service
One-Stop Solution to encode sentence to fixed length vectors from various embedding techniques
Metrics details
| Stars | 210 |
icoxfog417/fastTextJapaneseTutorial
Tutorial to train fastText with Japanese corpus
Metrics details
| Stars | 205 |
loretoparisi/fasttext.js
FastText for Node.js
Metrics details
| Stars | 200 |
giacbrd/ShallowLearn
An experiment about re-implementing supervised learning models based on shallow neural network approaches (e.g. fastText) with some additional exclusive features and nice API. Written in Python and fully compatible with Scikit-learn.
Metrics details
| Stars | 198 |
