lm_risk_cards

Model risk assessment toolkit

A set of tools and guidelines for assessing the security vulnerabilities of language models in AI applications

Risks and targets for assessing LLMs & LLM vulnerabilities

GitHub

28 stars
6 watching
7 forks
Language: Python
last commit: over 2 years ago
Linked from 1 awesome list

llmllm-securityred-teamingsecurityvulnerability

Backlinks from these awesome lists:

Related projects:

RepositoryDescriptionStars
safellama/plexiglassA toolkit to detect and protect against vulnerabilities in Large Language Models.122
howiehwong/trustllmA toolkit for assessing trustworthiness in large language models491
vhellendoorn/code-lmsA guide to using pre-trained large language models in source code analysis and generation1,789
protectai/llm-guardA security toolkit designed to protect interactions with large language models from various threats and vulnerabilities.1,296
ucsc-vlaa/vllm-safety-benchmarkA benchmark for evaluating the safety and robustness of vision language models against adversarial attacks.72
melih-unsal/demogptA comprehensive toolset for building Large Language Model (LLM) based applications1,733
aiplanethub/beyondllmAn open-source toolkit for building and evaluating large language models267
deadbits/vigil-llmA security scanner for Large Language Model prompts to detect potential threats and vulnerabilities326
ethz-spylab/rlhf_trojan_competitionDetecting backdoors in language models to prevent malicious AI usage109
mpaepper/llm_agentsBuilds agents controlled by large language models (LLMs) to perform tasks with tool-based components940
lzw-lzw/remoteglmDevelops a multimodal large-scale model for analyzing remote sensing images in scene analysis tasks108
davidmigloz/langchain_dartProvides a set of tools and components to simplify the integration of Large Language Models into Dart/Flutter applications441
13o-bbr-bbq/machine_learning_securityAn open-source project that explores the intersection of machine learning and security to develop tools for detecting vulnerabilities in web applications.1,987
mlgroupjlu/llm-eval-surveyA repository of papers and resources for evaluating large language models.1,450