FORGED GOODSsmall, specific, verified digital tools

HomeLocal-AI Stack Directory: 40 Self-Hosted LLM & Vector-DB Tools, Verified Specs › Apache-2.0

License Apache-2.0: 19 tools compared (category, min RAM GPU, offline capable)

19 of 40 rows from Local-AI Stack Directory: 40 Self-Hosted LLM & Vector-DB Tools, Verified Specs — every row carries the source it was checked against. Click a name for its notes.

tool namecategorymin RAM GPUoffline capablematuritysource URL
vLLMLLM runtime16GB+ GPU VRAM recommendedyesmature, very activesource
text-generation-inferenceLLM runtimeGPU required, 16GB+ VRAMyesmature, activesource
FastChatLLM runtime/frameworkGPU recommended, 16GB VRAMyesmature, moderate activitysource
MLC-LLMLLM runtime4GB+ RAM, mobile/GPU supportyesmature, activesource
llamafileLLM runtime4GB+ RAM, CPU-only okyesactivesource
h2oGPTLLM runtime/RAG16GB RAM, GPU optionalyesactivesource
PrivateGPTRAG framework8GB RAM, GPU optionalyesactivesource
HaystackRAG frameworkdepends on backend modelyes (with local models)mature, activesource
XinferenceLLM runtime/servingdepends on model sizeyesactivesource
OpenLLMLLM servingGPU recommendedyesactivesource
SGLangLLM runtimeGPU required, 16GB+ VRAMyesactivesource
Text Embeddings InferenceEmbedding serverGPU optional, 4GB+ RAMyesactivesource
ChromaVector DB2GB+ RAM, no GPU neededyesmature, activesource
QdrantVector DB2GB+ RAM, no GPU neededyesmature, very activesource
MilvusVector DB8GB+ RAM recommendedyesmature, very activesource
VespaVector DB/search engine4GB+ RAM, scalableyesmature, activesource
MarqoVector search engine4GB+ RAM, GPU optionalyesactivesource
ValdVector DBscalable, k8s-basedyesactivesource
LanceDBVector DB2GB+ RAM, no GPU neededyesactivesource
The full verified table: Local-AI Stack Directory: 40 Self-Hosted LLM & Vector-DB Tools, Verified Specs — 40 rows, CSV + JSON, every row with a checked source. €14   See the product
More slices of this table