word2vec-slim
Word model reducer
Slims down a large pre-trained word2vec model to reduce size and improve loading time
word2vec Google News model slimmed down to 300k English words
213 stars
11 watching
37 forks
Language: Python
last commit: about 9 years agoRelated projects:
| Repository | Description | Stars |
|---|---|---|
| A repository hosting pre-trained word vector model (3 million 300-dimension English word vectors) from the Google News corpus. | 519 | |
| A framework to learn word embeddings using lexical dictionaries | 115 | |
| This is a word embedding model trained on Stack Overflow posts for use in natural language processing tasks. | 40 | |
| A PyTorch implementation of the skip-gram model for learning word embeddings. | 188 | |
| A Scala implementation of the word2vec model representation. | 11 | |
| A package for reading and analyzing word embeddings from the word2vec format in Go. | 56 | |
| An algorithm to reduce word embeddings to a specified size while maintaining performance | 129 | |
| A tool for learning vector representations of words and entities from Wikipedia text data. | 946 | |
| An implementation of a self-supervised speech representation model using PyTorch and disentangled speaker embeddings | 471 | |
| An implementation of a word embedding model that uses character n-grams and achieves state-of-the-art results in multiple NLP tasks | 803 | |
| A library for training and evaluating a type of word embedding model that extends the original Word2Vec algorithm | 20 | |
| A C implementation of a stemming algorithm to reduce words to their base form | 39 | |
| This implementation provides a way to represent words as multivariate Gaussian distributions, allowing scalable word embeddings. | 190 | |
| A fast Python library for reading and writing seismic data formats | 494 | |
| Provides pre-trained word vectors for multiple languages to facilitate NLP tasks | 2,216 |