multimodal_rerprogramming

Model reprogramming

Cross-modal Adversarial Reprogramming enables retraining of image models on text classification tasks

Multimodal adversarial rerprogramming

GitHub

11 stars
5 watching
1 forks
Language: Jupyter Notebook
last commit: almost 5 years ago

Related projects:

RepositoryDescriptionStars
paarthneekhara/rnn_adversarial_reprogrammingRepurposes pre-trained neural networks for new classification tasks through adversarial reprogramming of their inputs.6
prinsphield/adversarial_reprogrammingThis project enables reprogramming of pre-trained neural networks to work on new tasks by fine-tuning them on smaller datasets.33
multimodal-art-projection/omnibenchEvaluates and benchmarks multimodal language models' ability to process visual, acoustic, and textual inputs simultaneously.15
dodohow1011/speechadvreprogramDeveloping low-resource speech command recognition systems using adversarial reprogramming and transfer learning18
yunyuntsai/black-box-adversarial-reprogrammingAn approach to adapt machine learning models using scarce data and limited resources by modifying their internal workings without changing the model's original architecture or training data.37
pku-yuangroup/languagebindExtending pretraining models to handle multiple modalities by aligning language and video representations751
huckiyang/voice2series-reprogrammingAn approach to reprogramming acoustic models for time series classification using differential mel-spectrograms and adversarial training70
yerevann/warpAn approach to transfer learning for NLP tasks using adversarial reprogramming and word-level task-specific embeddings.83
openbmb/viscpmA family of large multimodal models supporting multimodal conversational capabilities and text-to-image generation in multiple languages1,098
subho406/omninetAn implementation of a unified architecture for multi-modal multi-task learning using PyTorch.515
ailab-cvc/seedAn implementation of a multimodal language model with capabilities for comprehension and generation585
xverse-ai/xverse-v-13bA large multimodal model for visual question answering, trained on a dataset of 2.1B image-text pairs and 8.2M instruction sequences.78
l0sg/relational-rnn-pytorchAn implementation of DeepMind's Relational Recurrent Neural Networks (Santoro et al. 2018) in PyTorch for word language modeling245
lyuchenyang/macaw-llmA multi-modal language model that integrates image, video, audio, and text data to improve language understanding and generation1,568
pleisto/yuren-baichuan-7bA multi-modal large language model that integrates natural language and visual capabilities with fine-tuning for various tasks73