AI Frameworks
Libraries for LLM-powered apps
24 tools in this category
LangChain
Python • MIT
Chain-based LLM framework — prompts, chains, agents, memory
LlamaIndex
Python • MIT
Data framework for LLMs — RAG, agents, data connectors
vLLM
Python • Apache-2.0
High-throughput LLM serving — PagedAttention, continuous batching
Text Generation Inference
Python/Rust • Apache-2.0
HuggingFace production serving — optimized for inference
LMDeploy
Python • Apache-2.0
Efficient LLM serving — compression, quantization, high throughput
Transformers
Python • Apache-2.0
HuggingFace unified model framework — NLP, vision, audio, multimodal
Diffusers
Python • Apache-2.0
State-of-the-art diffusion models for image, video, and audio generation
DSPy
Python • MIT
Programming—not prompting—language models. Declarative specs compiled to optimized prompts.
Unsloth
Python • Apache-2.0
Fine-tune LLMs 2-5x faster with 70% less VRAM — consumer GPU friendly
Axolotl
Python • Apache-2.0
YAML-configurable fine-tuning for LoRA, QLoRA, and full fine-tuning
SGLang
Python • Apache-2.0
High-performance LLM serving with structured output — JSON mode, regex constraints
BentoML
Python • Apache-2.0
Unified ML model serving — any model type, production-ready
Milvus
Go/Python • Apache-2.0
Distributed vector database — cloud-native, billion-scale
Qdrant
Rust • Apache-2.0
High-performance vector database — billion-scale similarity search
Weaviate
Go • BSD-3
Multimodal vector database — GraphQL, REST, hybrid search
Chroma
Python/Rust • Apache-2.0
Embedded vector database — simplest path to production RAG
RAGFlow
Python • Apache-2.0
Open-source RAG engine — deep document understanding
Docling
Python • MIT
Document conversion — PDF, Word, Excel, PowerPoint to markdown
Marker
Python • Apache-2.0
PDF to markdown — fast, accurate, open-source
Unstructured
Python • Apache-2.0
Document preprocessing — any format to LLM-ready chunks
TensorRT-LLM
Python/C++ • Apache-2.0
NVIDIA optimized LLM inference — maximum speed on NVIDIA GPUs
Chroma
Python/Rust • Apache-2.0
Embedded vector database — simplest path to production RAG
Qdrant (alt)
Rust • Apache-2.0
High-performance vector database — billion-scale similarity search
Weaviate (alt)
Go • BSD-3
Multimodal vector database — GraphQL, REST, hybrid search