Vllm
A high-throughput and memory-efficient inference and serving engine for LLMs
Plugins Filament, packages Laravel et starter kits open source activement maintenus. Tous en MIT, avec des releases pour v3, v4 et v5 selon les cas.
A high-throughput and memory-efficient inference and serving engine for LLMs
Unified AI router with 356 providers, RTK+Caveman compression, auto fallback, MCP/A2A, desktop, PWA, and OpenAI-compatible APIs.
Extensible local model router for Codex and Cursor
See exactly what your coding agent sends to the model — a local logging reverse-proxy + web dashboard for Claude Code, Codex, DeepSeek-TUI, Reasonix, and Kimi. Just run `ccglass`.
Model-neutral agent desktop/runtime for private, enterprise, and OpenAI-compatible / Anthropic-compatible model API. Tested on DeepSeek, Qwen, Kimi, GLM models. Supports MCP, Skills,and local code search.
Tous les packages ont une CI, des tests Pest et des issues ouvertes taguées good first issue.
Nouvelle version disponible.