OneLLM

Language model trainer

A framework for training and fine-tuning multimodal language models on various data types

[CVPR 2024] OneLLM: One Framework to Align All Modalities with Language

GitHub

601 stars
11 watching
33 forks
Language: Python
last commit: almost 2 years ago
Linked from 1 awesome list


Backlinks from these awesome lists:

Related projects:

RepositoryDescriptionStars
microsoft/mpnetDevelops a method for pre-training language understanding models by combining masked and permuted techniques, and provides code for implementation and fine-tuning.288
yunwentechnology/unilmThis project provides pre-trained models and tools for natural language understanding (NLU) and generation (NLG) tasks in Chinese.439
bobazooba/xllmA tool for training and fine-tuning large language models using advanced techniques387
elanmart/psmmAn implementation of a neural network model for character-level language modeling.50
openai/finetune-transformer-lmThis project provides code and model for improving language understanding through generative pre-training using a transformer-based architecture.2,167
vhellendoorn/code-lmsA guide to using pre-trained large language models in source code analysis and generation1,789
bilibili/index-1.9bA lightweight, multilingual language model with a long context length920
llava-vl/llava-plus-codebaseA platform for training and deploying large language and vision models that can use tools to perform tasks717
bytedance/lynx-llmA framework for training GPT4-style language models with multimodal inputs using large datasets and pre-trained models231
openai/lm-human-preferencesTraining methods and tools for fine-tuning language models using human preferences1,240
brightmart/xlnet_zhTrains a large Chinese language model on massive data and provides a pre-trained model for downstream tasks230
yiren-jian/blitextDevelops and trains models for vision-language learning with decoupled language pre-training24
pleisto/yuren-baichuan-7bA multi-modal large language model that integrates natural language and visual capabilities with fine-tuning for various tasks73
lyuchenyang/macaw-llmA multi-modal language model that integrates image, video, audio, and text data to improve language understanding and generation1,568
luogen1996/lavinAn open-source implementation of a vision-language instructed large language model513