multifit

Fine-tuning

Reproduces results from a paper on efficient multi-lingual language model fine-tuning using a rewritten framework on top of the fastai library

The code to reproduce results from paper "MultiFiT: Efficient Multi-lingual Language Model Fine-tuning" https://arxiv.org/abs/1909.04761

GitHub

284 stars
17 watching
56 forks
Language: Jupyter Notebook
last commit: over 6 years ago
fastaimultiple-languagesnlpulmfit

Related projects:

RepositoryDescriptionStars
roboflow/maestroA tool to streamline fine-tuning of multimodal models for vision-language tasks1,415
codefuse-ai/mftcoderA framework for fine-tuning large language models with multiple tasks to improve their accuracy and efficiency647
eleutherai/polyglotLarge language models designed to perform well in multiple languages and address performance issues with current multilingual models.476
openai/finetune-transformer-lmThis project provides code and model for improving language understanding through generative pre-training using a transformer-based architecture.2,167
openai/lm-human-preferencesTraining methods and tools for fine-tuning language models using human preferences1,240
ymcui/macbertImproves pre-trained Chinese language models by incorporating a correction task to alleviate inconsistency issues with downstream tasks646
jerry1993-tech/cornucopia-llama-fin-chineseA Chinese finance-focused large language model fine-tuning framework596
ys-zong/vlguardImproves safety and helpfulness of large language models by fine-tuning them using safety-critical tasks47
wenkehuang/rethinkflImproves federated learning performance by incorporating domain knowledge and regularization to adapt models across diverse domains93
pku-yuangroup/languagebindExtending pretraining models to handle multiple modalities by aligning language and video representations751
babylonhealth/fasttext_multilingualA repository providing aligned multilingual word vectors for 78 languages using the SVD method.1,197
git-cloner/llama2-lora-fine-tuningFine-tuning the LLaMA 2 chat model using DeepSpeed and Lora for improved performance on a large dataset.171
jshilong/gpt4roiTraining and deploying large language models on computer vision tasks using region-of-interest inputs517
multimodal-art-projection/omnibenchEvaluates and benchmarks multimodal language models' ability to process visual, acoustic, and textual inputs simultaneously.15
microsoft/archaiAutomates the search for optimal neural network configurations in deep learning applications468