FoolyourVLLMs
by ys-zong
[ICML 2024] Fool Your (Vision and) Language Model With Embarrassingly Simple Permutations
AI summary
Attack framework
An attack framework to manipulate the output of large language models and vision-language models
- stars
- 14
- forks
- 2
- watching
- 1
Similar projects
Found by comparing what the projects do, not just their names.
Attack framework
An adversarial attack framework on large vision-language models
Safety fine-tuner
Improves safety and helpfulness of large language models by fine-tuning them using safety-critical tasks
Learning benchmark
A benchmarking suite for multimodal in-context learning models
Federated Learning Attack
A framework for attacking federated learning systems with adaptive backdoor attacks
Backdoor detector
Detecting backdoors in language models to prevent malicious AI usage
yuxie11/r2d2157
Vision-Language Framework
A framework for large-scale cross-modal benchmarks and vision-language tasks in Chinese
Attack defense
A defense mechanism against model poisoning attacks in federated learning
Gradient attack tool
A tool to demonstrate and analyze attacks on private data in machine learning models using gradients
zjunlp/knowlm1.3K
Knowledge model framework
A framework for training and utilizing large language models with knowledge augmentation capabilities
Malware attacker tool
An open-source reinforcement learning framework to generate adversarial examples for malware classification models.
Model inversion attack
This implementation allows an attacker to directly obtain user data from federated learning gradient updates by modifying the shared model architecture.
Model validator
Analyzing and mitigating object hallucination in large vision-language models to improve their accuracy and reliability.
Backdoor defense framework
A framework for defending against backdoor attacks in federated learning systems
Image captioner
An end-to-end image captioning system that uses large multi-modal models and provides tools for training, inference, and demo usage.
Adversarial examples generator
A tool for generating adversarial examples to attack text classification and inference models