Vllm
A high-throughput and memory-efficient inference and serving engine for LLMs
Aktiv gepflegte Filament-Plugins, Laravel-Pakete und Open-Source-Starter-Kits. Alle MIT, mit Releases für v3, v4 und v5, sofern zutreffend.
A high-throughput and memory-efficient inference and serving engine for LLMs
Unified AI router with 356 providers, RTK+Caveman compression, auto fallback, MCP/A2A, desktop, PWA, and OpenAI-compatible APIs.
Extensible local model router for Codex and Cursor
See exactly what your coding agent sends to the model — a local logging reverse-proxy + web dashboard for Claude Code, Codex, DeepSeek-TUI, Reasonix, and Kimi. Just run `ccglass`.
Model-neutral agent desktop/runtime for private, enterprise, and OpenAI-compatible / Anthropic-compatible model API. Tested on DeepSeek, Qwen, Kimi, GLM models. Supports MCP, Skills,and local code search.
Alle Pakete haben CI, Pest-Tests und offene Issues mit dem Tag good first issue.
Neue Version verfügbar.