Thanks to visit codestin.com
Credit goes to github.com

Skip to content
#

locomo

Here are 37 public repositories matching this topic...

Drop-in FastAPI proxy for llama.cpp, Ollama, vLLM and similar backends. Automatically prunes, summarizes, and extracts the most relevant context from large inputs using advanced strategies (readagent, rlm, embeddings) so your local models answer accurately without cloud services.

  • Updated Sep 4, 2026
  • Python

Add this topic to your repo

To associate your repository with the locomo topic, visit your repo's landing page and select "manage topics."

Learn more