opik

LM testing platform

A platform for evaluating and testing large language models (LLMs) during development and production.

Open-source end-to-end LLM Development Platform

GitHub

3k stars
38 watching
158 forks
Language: Java
last commit: almost 2 years ago
Linked from 9 awesome lists


Backlinks from these awesome lists:

Related projects:

RepositoryDescriptionStars
poyro/poyroAn extension of Vitest for testing LLM applications using local language models31
innogames/ltcA tool for managing load tests and analyzing performance results200
lcm-proj/lcmA set of libraries and tools for efficient message passing and data marshalling in real-time systems.1,011
sunlemuria/opengptandbeyondAn effort to develop and compare large language models beyond OpenGPT105
johnsnowlabs/langtestA tool for testing and evaluating large language models with a focus on AI safety and model assessment.506
luogen1996/lavinAn open-source implementation of a vision-language instructed large language model513
norman/telescopeA test library for Lua that supports declarative testing with nested contexts and code coverage reports.161
qcri/llmebenchA benchmarking framework for large language models81
openolat/openolatA web-based e-learning platform with features like assessment, content management, and learning resources, built using Java.337
llm-ui-kit/llm-uiA React library designed to work with Large Language Models (LLMs) by providing features such as syntax removal, custom component addition, and rendering at a native frame rate.425
ailab-cvc/seed-benchA benchmark for evaluating large language models' ability to process multimodal input322
melih-unsal/demogptA comprehensive toolset for building Large Language Model (LLM) based applications1,733
talkdai/dialogAn application framework to simplify the deployment and testing of large language models (LLMs) for natural language processing tasks.380
davidmigloz/langchain_dartProvides a set of tools and components to simplify the integration of Large Language Models into Dart/Flutter applications441
mlabonne/llm-autoevalA tool to automate the evaluation of large language models in Google Colab using various benchmarks and custom parameters.566