llmperf

LLM benchmarking tool

A tool for evaluating the performance of large language model APIs

LLMPerf is a library for validating and benchmarking LLMs

GitHub

678 stars
9 watching
115 forks
Language: Python
last commit: almost 2 years ago
Linked from 1 awesome list


Backlinks from these awesome lists:

Related projects:

RepositoryDescriptionStars
qcri/llmebenchA benchmarking framework for large language models81
damo-nlp-sg/m3examA benchmark for evaluating large language models in multiple languages and formats93
wgryc/phasellmA framework for managing and testing large language models to evaluate their performance and optimize user experiences.451
ajndkr/lanarkyA Python web framework specifically designed to build LLM microservices with built-in support for FastAPI and streaming capabilities.978
relari-ai/continuous-evalProvides a comprehensive framework for evaluating Large Language Model (LLM) applications and pipelines with customizable metrics455
ai-hypercomputer/maxtextA high-performance LLM written in Python/Jax for training and inference on Google Cloud TPUs and GPUs.1,557
luogen1996/lavinAn open-source implementation of a vision-language instructed large language model513
r2d4/openlmLibrary that provides a unified API to interact with various Large Language Models (LLMs)367
mlcommons/inferenceMeasures the performance of deep learning models in various deployment scenarios.1,256
ray-project/rayA unified framework for scaling AI and Python applications by providing a distributed runtime and a set of libraries for machine learning and other compute tasks.34,412
aifeg/benchlmmAn open-source benchmarking framework for evaluating cross-style visual capability of large multimodal models84
dreadnode/riggingA framework for leveraging language models in production code216
multimodal-art-projection/omnibenchEvaluates and benchmarks multimodal language models' ability to process visual, acoustic, and textual inputs simultaneously.15
aiplanethub/beyondllmAn open-source toolkit for building and evaluating large language models267
internlm/lagentA lightweight framework for building agent-based applications using LLMs and transformer architectures1,924