FLAN

Language model tuner

A repository providing tools and datasets to fine-tune language models for specific tasks

GitHub

1k stars
32 watching
156 forks
Language: Python
last commit: almost 2 years ago
Linked from 1 awesome list


Backlinks from these awesome lists:

Related projects:

RepositoryDescriptionStars
openai/finetune-transformer-lmThis project provides code and model for improving language understanding through generative pre-training using a transformer-based architecture.2,167
openai/lm-human-preferencesTraining methods and tools for fine-tuning language models using human preferences1,240
leks-forever/nllb-tuningThis is an experimental project for fine-tuning the NLB language model with a specific dataset and evaluating its performance on translation tasks.7
thunlp/ernieA toolkit for fine-tuning pre-trained language models with knowledge graph representations to improve performance on entity typing and relation classification tasks.1,413
google-deepmind/recurrentgemmaAn implementation of a fast and efficient language model architecture613
elanmart/psmmAn implementation of a neural network model for character-level language modeling.50
csuhan/onellmA framework for training and fine-tuning multimodal language models on various data types601
ibm-granite/granite-3.0-language-modelsA collection of lightweight state-of-the-art language models designed to support multilinguality, coding, and reasoning tasks on constrained resources.232
spandan-madan/pytorch_fine_tuning_tutorialProvides guidance on fine-tuning pre-trained models for image classification tasks using PyTorch.279
ieit-yuan/yuan2.0-m32A high-performance language model designed to excel in tasks like natural language understanding, mathematical computation, and code generation182
deepset-ai/farmAn open-source framework for adapting representation models to various tasks and industries1,743
felixgithub2017/mmcuMeasures the understanding of massive multitask Chinese datasets using large language models87
facebookresearch/compilergymA reinforcement learning environment library for compiler optimization tasks917
google-research/relay-policy-learningEnvironments and data for training reinforcement learning agents in a kitchen simulator108
nvidia/sentiment-discoveryLarge-scale unsupervised language modeling for robust sentiment classification and related NLP tasks1,061