Skip to content
Open source

Otoroshi extension · Open source

Otoroshi LLM Extension

The open-source AI gateway for Otoroshi.

Connect, set up, secure and seamlessly manage LLMs through a universal, OpenAI-compatible API. 50+ providers, guardrails, semantic cache, budgets, MCP and AI Studio.

Stars
19
Language
Scala
License
Apache-2.0
Last update
today

Features

What's inside

Unified, OpenAI-compatible API

One interface for 50+ providers, including sovereign European ones (Mistral, Scaleway, OVHcloud, Cloud Temple) and local models with Ollama.

Load balancing, fallbacks, retries

Distribute workloads across providers and rescue failing requests automatically.

Simple & semantic cache

Speed up repeated queries, improve response times and reduce costs.

Quotas, budgets & costs

Token quotas per consumer, budgets, cost tracking and ecological impact reporting.

Guardrails

Validate prompts and responses against PII and secrets leakage, prompt injection, toxicity, gibberish and more.

MCP, agents & workflows

MCP connectors, virtual MCP servers and registry, tool calling, persistent memories and agentic workflows.

Multi-modal

Audio (TTS, STT, translation), images, video, embeddings and vector stores through the same gateway.

Key vault & fine-grained authorizations

Keep provider keys in Otoroshi vaults and constrain model usage by identity, API key or any request detail.

AI Studio

A self-service AI console on top of the gateway, with workspaces, playground, logs and budgets.

  • OpenAI
  • Anthropic
  • Mistral AI
  • Gemini
  • Ollama
  • Azure AI
  • Groq
  • DeepSeek
  • Scaleway
  • OVHcloud
  • Cohere
  • Hugging Face
  • Cloudflare AI
  • ElevenLabs
  • + many more, local models too

AI Studio

One API
for every model.

AI Studio is an OpenRouter-like console served by Otoroshi itself: workspaces, bring-your-own-key providers, API keys, a chat playground, activity, logs and budgets — all stored as plain Otoroshi entities, manageable through the admin API, Kubernetes or GitOps.

  • Workspaces
  • API keys & BYOK
  • Chat playground
  • Activity & logs
ai-studio · Overview
AI Studio overview

Guardrails

Safe prompts.
Safe answers.

Guardrails validate prompts and responses before anything reaches a model or a user: personal and sensitive information, secrets, prompt injection, toxic or irrelevant content. Chain them per route, per model or per workspace.

ai-studio · Guardrails
AI Studio guardrails configuration

AI Studio Enterprise

Share AI Studio with your whole organization

Your company login, members and roles in every workspace, secrets kept on the server and an audit trail: the same studio, as a standalone application for every team.

Discover AI Studio Enterprise

More open-source projects

Run it in productionwithout running Otoroshi.

Our extensions are available on Otoroshi Managed, fully operated by the people that wrote them. Or get professional support for your own clusters.