Skills
- 34.4K
AirLLM
lyogavin · Tool / CLI
Runs 70B models on a single 4 GB GPU without quantization, by keeping only one layer on the GPU at a time.
DevOpsAny agent / LLMActive. Last activity Sep 15, 2026 - 22.4K
DwarfStar (ds4)
antirez · Tool / CLI
antirez's focused engine for running DeepSeek V4 Flash and a few other large open models locally on Metal, CUDA and ROCm.
DevOpsAny agent / LLMActive. Last activity Sep 14, 2026 - 29.8K
Modular Platform (MAX and Mojo)
Modular · Tool / CLI
Modular's open platform for AI development and deployment: the MAX framework, the Mojo language and an OpenAI-compatible server.
DevOpsAny agent / LLMActive. Last activity Sep 15, 2026 - 36K
SGLang
SGLang project · Tool / CLI
A high-performance serving framework for LLMs and multimodal models, from a single GPU to large distributed clusters.
DevOpsAny agent / LLMActive. Last activity Sep 15, 2026 - 91.8K
vLLM
vLLM project · Tool / CLI
A high-throughput, memory-efficient engine for serving LLMs, with PagedAttention, continuous batching and prefix caching.
DevOpsAny agent / LLMActive. Last activity Sep 15, 2026