EveryDev.ai
Subscribe
Home
Developers

3,289+ AI companies

  • Radar
  • Trending
AI Tools by Topic
  • AI Coding Assistants
  • Agent Frameworks
  • MCP Servers
  • AI Prompt Tools
  • Vibe Coding Tools
  • AI Design Tools
  • AI Database Tools
  • AI Website Builders
  • AI Testing Tools
  • LLM Evaluations
Follow Us
  • X / Twitter
  • LinkedIn
  • Reddit
  • Discord
  • Threads
  • Bluesky
  • Mastodon
  • YouTube
  • GitHub
  • Instagram
Get Started
  • About
  • Editorial Standards
  • Corrections & Disclosures
  • Community Guidelines
  • Advertise
  • Contact Us
  • Newsletter
  • Submit a Tool
  • Start a Discussion
  • Write A Blog
  • Share A Build
  • Terms of Service
  • Privacy Policy
Explore with AI
  • ChatGPT
  • Gemini
  • Claude
  • Grok
  • Perplexity
Agent Experience
  • llms.txt
Theme
With AI, Everyone is a Dev. EveryDev.ai © 2026
    1. Home
    2. Developers
    3. Kottos AI, Inc.

    Kottos AI, Inc.

    Kottos AI builds exchange-grade infrastructure for LLM inference: it records request metadata such as cost, cache usage, timing and errors, then uses that evidence to provide intelligent routing across model providers and venues. Its open-source llmbridge gateway supplies the low-latency translation and proxy layer.

    Visit Website

    At a Glance

    1Tool Listed
    4Products
    8Capabilities
    Discussions
    Austin, TexasHeadquarters
    2026Est.
    Focus Areas
    AI Infrastructure
    LLM Orchestration
    API Integration Platforms
    Connect
    Latest News
    llmbridge v0.59.1: Bedrock credential handling fixSep 16, 2026
    llmbridge v0.59.0 and v0.58.0 add request sequencing and optional prefix hashing metadataSep 14, 2026
    Markets
    • AI application developers and teams operating LLM inference at scale
    • Enterprises requiring managed or customer-hosted inference infrastructure
    • Latency-sensitive AI applications
    • Teams using multiple model providers or open-weight model venues
    • +2 more

    AI Tools by Kottos AI, Inc.

    (1)
    View llmbridge
    llmbridge tool icon

    llmbridge

    Open Source C++ LLM Gateway

    AI InfrastructureLLM OrchestrationAPI Integrations

    Discussions

    No discussions yet

    Be the first to start a discussion about Kottos AI, Inc.

    Latest News

    09/16/2026

    llmbridge v0.59.1: Bedrock credential handling fix

    github.com
    09/14/2026

    llmbridge v0.59.0 and v0.58.0 add request sequencing and optional prefix hashing metadata

    github.com
    09/12/2026

    llmbridge v0.57.0 adds the upstream venue request ID to request records

    github.com
    09/11/2026

    llmbridge v0.56.0 adds recording of OpenAI cache-write counts

    github.com

    Products & Services

    4
    Kottos AI intelligent routing

    Commercial routing intelligence that observes served requests, provider quotes, cache behavior, price, timing and venue health, predicts expected cost/performance, and routes requests according to customer constraints.

    Kottos hosted gateway (private beta)

    Managed OpenAI-compatible endpoint that records venue, cost, timing and failures, provides routing, observability and team-management capabilities, and supports BYOK credentials passed per request.

    llmbridge

    Apache-2.0 open-source C++ sub-millisecond LLM gateway. It accepts OpenAI-compatible requests, translates to provider dialects such as Anthropic, Gemini and Cohere, proxies responses, supports streaming SSE, tool calling, prompt-cache forwarding, timing headers, TLS and credential passthrough.

    Claude Code trial and public dashboard

    No-signup trial that places Kottos between Claude Code and Anthropic and exposes request metadata such as model, venue, time to first token, cache reads, cost and status in a shared public dashboard; prompt and response text are not stored.

    Market Position

    Kottos positions llmbridge against general-purpose LLM gateways such as LiteLLM, Bifrost, Portkey, GoModel and similar proxies by emphasizing C++ implementation, microsecond translation overhead, single-core efficiency and high concurrency. It differentiates the commercial Kottos layer from a basic gateway through an inference tape, live provider price/latency data, cache-aware expected-cost routing, verification of expected versus actual cost, observability and managed deployment. The company explicitly describes LiteLLM, Bifrost and Helicone as adjacent gateway/observability work and published benchmark comparisons with several of them.

    Leadership

    Founders

    LA

    Lluís Antoni Jiménez Rugama

    Founder; PhD in applied mathematics with nine years in high-frequency and electronic trading. His LinkedIn profile identifies him as Kottos AI's founder and places him in Austin; he also maintains the kottos-ai GitHub organization.

    Executive Team

    LA

    Lluís Antoni Jiménez Rugama

    Founder

    PhD in applied mathematics and nine years in high-frequency and electronic trading; listed by LinkedIn as Kottos AI founder and described by Kottos as the person who founded and built the company.

    Founding Story

    The founder describes an analogy between LLM inference and pre-TRACE fixed-income markets: providers quote prices, but users need an auditable record of what was actually delivered, including effective cost, latency and failures. Kottos AI was started to build an inference 'tape' and use it to improve routing decisions; llmbridge was built as the exchange-grade, microsecond-overhead gateway underneath it.

    Business Model

    Revenue Model

    The commercial layer is managed/private-beta hosted infrastructure and enterprise or customer-hosted routing intelligence, observability and integrations. The open-source llmbridge core is Apache 2.0; hosted use is currently private beta and BYOK, with customers pointing an OpenAI-compatible client at Kottos and Kottos managing the gateway and routing.

    Pricing Tiers

    Claude Code trial
    No signup; no Kottos charge

    Shared public trial account/dashboard; provider quotas apply and the trial page says nothing is billed to Kottos.

    Hosted gateway private beta

    Managed infrastructure, routing, monitoring and founder support; prospective customers apply by email.

    Customer-hosted / Enterprise

    Customer runs the gateway; Kottos supplies routing intelligence and receives routing/usage metadata.

    Private company; no IPO plans or public listing were found.

    Target Markets

    Industries & Segments
    • AI application developers and teams operating LLM inference at scale
    • Enterprises requiring managed or customer-hosted inference infrastructure
    • Latency-sensitive AI applications
    • Teams using multiple model providers or open-weight model venues
    • Organizations needing production observability, cost attribution and SSO/enterprise controls
    • Design partners evaluating broad provider comparisons
    Use Cases
    • Production LLM inference routing and provider selection
    • Reducing inference cost while preserving prompt-cache benefits
    • Latency-sensitive agent loops
    • Voice agents and high-concurrency streaming responses
    • Trading agents and other workloads where request-path latency is critical
    • Cost attribution, monitoring and observability by user, team and project

    Quick Facts

    Headquarters
    Austin, Texas, United States
    Founded
    2026
    Entity Type
    Inc.

    History & Milestones

    2026-05-26

    Kottos AI, Inc. was reported as a Delaware domestic corporation filed on May 26, 2026.

    2026-08-28

    Kottos AI filed the KOTTOS AI trademark application (serial 50079098) for downloadable software for routing and transmitting API requests and prompts.

    2026-09-04

    Kottos published its external ENTERPILOT gateway benchmark comparison and llmbridge changelog version 0.55.0; the benchmark reported 0.07 ms non-streaming added latency and 40,677 peak requests/second for llmbridge in its test setup.

    2026-09-14

    llmbridge changelog added RequestFacts::prefix_hash and RequestFacts::seq in versions 0.58.0 and 0.59.0.

    2026-09-16

    llmbridge changelog reached version 0.59.1 with a Bedrock credential error-handling fix; the repository showed 293 commits and an active latest commit on the same date.

    Key Capabilities

    8
    Inference tape recording request metadata without storing prompt or response text
    Cost, cache-read/cache-write, timing, status and error observability
    Intelligent routing based on expected cost, cache behavior, latency, provider quotes and venue health
    OpenAI-compatible API with on-the-fly provider dialect translation
    Sub-millisecond gateway overhead, with published p99 below 1 ms at 1,000 RPS on one core
    Streaming SSE and tool-calling translation between OpenAI and Anthropic

    Integrations & Partnerships

    Platform Integrations

    • OpenAI-compatible client/API
    • Anthropic API and Claude Code
    • Google Gemini
    • Cohere
    • AWS Bedrock
    • Together
    • Fireworks
    • Groq

    Key Partnerships

    ENTERPILOT / Jakub A. Wąsek's AI gateway reproducible benchmark, used as an external benchmark harness
    Design partners for broader provider comparisons and routing coverage (described by Kottos as in development)

    Connect

    Website
    kottos.ai
    GitHub
    kottos-ai
    X / Twitter
    KottosAI
    LinkedIn
    lluisantonijimenezrugama

    AI Topics

    3

    Kottos AI, Inc. focuses on these topics:

    AI Infrastructure(1)
    LLM Orchestration(1)
    API Integration Platforms(1)
    Back to all developersSuggest an edit