← Back to Explore
Local and Self-Hosted AI🚀 production

vLLM

vllm-project/vllm
GitHub ↗

High-throughput LLM serving engine with PagedAttention for production-grade local inference.

LanguagePython
SourceREADME.md (Line 814)
#Python#Local

Related Resources in Local and Self-Hosted AI

Self-hosted personal AI assistant on your own VPS that lives in Telegram, with persistent long-term memory, voice, and Claude-powered reasoning.

TypeScript#Telegram
Website ↗