llama-stack
by meta-llama
Composable building blocks to build Llama Apps
AI summary
AI toolkit
Provides pre-packaged building blocks for generative AI applications with standardized APIs and service-oriented design.
- stars
- 5.2K
- forks
- 659
- watching
- 138
Similar projects
Found by comparing what the projects do, not just their names.
meta-llama/llama56.8K
LLAMA framework
A collection of tools and utilities for deploying, fine-tuning, and utilizing large language models.
meta-llama/llama327.5K
Language model library
Provides pre-trained and instruction-tuned Llama 3 language models and tools for loading and running inference
LLM toolkit
Provides tools and examples for fine-tuning the Meta Llama model and building applications with it
Language Model
An implementation of a large language model using the nanoGPT architecture
meta-llama/codellama16.1K
Code generator
Provides inference code and tools for fine-tuning large language models, specifically designed for code generation tasks
Language Model Runtime
An efficient C#/.NET library for running Large Language Models (LLMs) on local devices
LLaMA Tuner
A tool for efficiently fine-tuning large language models across multiple architectures and methods.
Data augmentation tool
A data framework for augmenting Large Language Models (LLMs) with private data
ggerganov/llama.cpp69.2K
LLM inference software
Enables LLM inference with minimal setup and high performance on various hardware platforms
LLM app builder
A framework for building enterprise LLM-based applications using small, specialized models
LLM toolkit
An open-source toolkit for pretraining and fine-tuning large language models
LLM evaluator
A framework for evaluating large language models
LLM integrator
A data framework for integrating large language models into applications with custom data
microsoft/lmops3.7K
LLM booster
A research initiative focused on developing fundamental technology to improve the performance and efficiency of large language models.
Instruction-following model tuner
An implementation of a method for fine-tuning language models to follow instructions with high efficiency and accuracy