MCPFast / Tools / Local AI proxy engine with semantic caching and intelligent routing

GitHubTool★★★★☆

Local AI proxy engine with semantic caching and intelligent routing

A local AI/LLM proxy in Rust/Axum featuring semantic caching, MCP injection, cost budgeting, and intelligent model routing.

View on GitHub

Kotro Proxy Engine: Local AI/LLM Proxy for Developers

The Kotro Proxy Engine is a high-performance, local AI/LLM proxy built in Rust using the Axum web framework. It's designed to streamline the development and deployment of AI-powered applications by providing a centralized, intelligent layer for interacting with various AI models. This tool offers advanced features like semantic caching, cost budgeting, and intelligent model routing, enabling developers to optimize performance, control expenses, and manage complex AI infrastructure efficiently.

What it Does

Kotro Proxy Engine acts as an intermediary between your applications and AI models. It intercepts requests, applies intelligent logic, and forwards them to the most suitable model. Key functionalities include:

Key Features

Developed with developers in mind, Kotro Proxy Engine offers a robust set of features for building and managing AI applications:

Who it's For

The Kotro Proxy Engine is an essential tool for: