CrossLingualContextualEmb

Language embedding aligner

Enables alignment of word embeddings across multiple languages to facilitate cross-lingual text analysis and machine learning tasks

Cross-Lingual Alignment of Contextual Word Embeddings

GitHub

99 stars
8 watching
9 forks
Language: Python
last commit: over 6 years ago
Linked from 1 awesome list

allennlpbertcontextual-embeddingscrosslingualelmonlppytorchwordembeddingszeroshot-learning

Backlinks from these awesome lists:

Related projects:

RepositoryDescriptionStars
hit-scir/elmoformanylangsProvides pre-trained ELMo representations for multiple languages to improve NLP tasks.1,462
botcenter/spanishwordembeddingsThis project generates Spanish word embeddings using fastText on large corpora.9
babylonhealth/fasttext_multilingualA repository providing aligned multilingual word vectors for 78 languages using the SVD method.1,197
pedrada88/crossembeddings-twitterPre-trained word embeddings from Twitter for natural language processing tasks14
pku-yuangroup/languagebindExtending pretraining models to handle multiple modalities by aligning language and video representations751
tanloong/interlaced.nvimA plugin for aligning bilingual parallel texts by re-positioning text and applying highlighting.7
lowresourcelanguages/champollionA toolkit providing ready-to-use parallel text sentence alignment tools for multiple language pairs.18
ukplab/arxiv2018-xling-sentence-embeddingsReproducible research on cross-lingual sentence embeddings using power mean word embeddings186
machinalis/yalignAutomates the process of extracting parallel sentences from comparable corpora to aid in statistical machine translation127
clab/fast_alignA fast and simple unsupervised word aligner for generating parallel corpus alignments.740
bheinzerling/bpembA collection of pre-trained subword embeddings in 275 languages, useful for natural language processing tasks.1,189
harsh19/spineTransforms existing word embeddings into more interpretable ones by applying a novel extension of k-sparse autoencoder with stricter sparsity constraints52
cmesher/inuktitutalignerdataScripts for aligning laboratory speech production data in Inuktitut3
guitarbum722/alignAn application and library for aligning text with flexible formatting options.84
malllabiisc/wordgcnA deep learning model that generates word embeddings by predicting words based on their dependency context291