← Back to the wire

GitHub - Gysho/LLMrPro: LLM Router PRO: Your machines run the models. The cloud only fills the gaps. Call every model through one OpenAI format. Self-hosted, and yours to own

AnnouncementProductJul 20, 2026

LLMrPro is a self-hosted LLM router exposing one OpenAI-compatible API that routes requests to user-owned machines running local inference engines including LM Studio, Ollama, vLLM, llama.cpp, and MLX. When local workers cannot serve a request, it falls back to configured cloud providers such as OpenAI, Anthropic, Google, or Azure OpenAI. The system is single-tenant, with encrypted provider keys and per-source rate limiting. A balancer and desktop agent comprise the two installable components.

Receipt № 7861 source · awaiting confirmation ◐

Evidence

1source· awaiting independent confirmation

No score is assigned. Sources and their independence are shown in the citation chain below.

Citation chain · 1 source

01medPRIMARY
OpenAICompanyAnthropicCompanyvLLMCompanyOllamaCompanyllama.cppModelMLXModelGoogleCompanyAzure OpenAICompanyLM StudioCompany
Canonical: https://github.com/Gysho/LLMrPro