Achieve state of the art inference performance with modern accelerators on Kubernetes
-
Updated
Sep 13, 2026 - Shell
Achieve state of the art inference performance with modern accelerators on Kubernetes
Production-ready Python library for multi-provider LLM orchestration
A production-grade Python library that intelligently coordinates multiple AI models to solve complex problems. Install with: pip install ai-council-orchestrator
Enterprise LLM API gateway for multi-model routing and unified auth.
A high-performance, tiered routing engine for Large Language Model (LLM) notifications. Built with AWS CDK and TypeScript to optimize latency, manage model costs, and reduce alert fatigue through intelligent classification.
Hardware-adaptive local LLM & cloud cascading gateway. Sub-5ms intelligent routing, 3-token lookahead failover, asymmetric verification, 1-click IDE config, and MCP for $0 token cost.
AI-powered customer support automation workflow built with n8n
Production-grade autonomous AI agent with multi-layer optimization: pattern matching, RAG personalization, and intelligent caching. Built from scratch in Python with cost tracking and automatic failure recovery. Demonstrates real-world agent engineering without frameworks.
🧠 Intelligent RAG with smart query routing - Choose the right search strategy automatically (FastSearch <1s, DeepResearch ~10s, WebSearch 2-5s)
AIGateWay-Universal - Global first production-grade, semantic-driven, fully compatible open-source AI capability intelligent routing and global scheduling engine
APIPOOL — Intelligent API marketplace for AI agents. 4-pillar scoring: self-learning, predictive, anomaly detection, contextual understanding.
Building intelligent communication for disconnected environments. Abhimanyu is an AI-assisted Delay-Tolerant Networking (DTN) system that enables reliable, autonomous data transmission across high-latency and intermittently connected networks.
AetherWeaver — an intelligent AI gateway & orchestrator for the Serverless Edge, with weighted intent routing across LangChain experts and end-to-end streaming.
Intelligent LLM router that reduces AI API costs by up to 60% through smart model selection and caching. FastAPI service with multi-provider support (Gemini, Claude, OpenRouter) and Claude Desktop MCP integration.
An autonomous orchestration platform that revolutionizes customer support by transforming unstructured inquiries into categorized, prioritized, and actionable tickets using Agentic AI.
多引擎智能路由 - 大模型API统一接入方案 | 智能调度 | 故障切换 | 成本优化
🚀 Automate your project builds with the Supreme Auto-Build Protocol, seamlessly integrating AI models and deploying to GitHub with a professional touch.
Aegis - OpenClaw智能优化插件。提供模型选择建议、Prompt优化、成本统计和质量评估功能。
To associate your repository with the intelligent-routing topic, visit your repo's landing page and select "manage topics."