EveryDev.ai
Subscribe
Home
Developers

3,756+ AI companies

  • Radar
  • Trending
AI Tools by Topic
  • AI Coding Assistants
  • Agent Frameworks
  • MCP Servers
  • AI Prompt Tools
  • Vibe Coding Tools
  • AI Design Tools
  • AI Database Tools
  • AI Website Builders
  • AI Testing Tools
  • LLM Evaluations
Follow Us
  • X / Twitter
  • LinkedIn
  • Reddit
  • Discord
  • Threads
  • Bluesky
  • Mastodon
  • YouTube
  • GitHub
  • Instagram
Get Started
  • Users
  • Rate Tools
  • About
  • Editorial Standards
  • Corrections & Disclosures
  • Community Guidelines
  • Advertise
  • Contact Us
  • Newsletter
  • Submit a Tool
  • Start a Discussion
  • Write A Blog
  • Share A Build
  • Terms of Service
  • Privacy Policy
Explore with AI
  • ChatGPT
  • Gemini
  • Claude
  • Grok
  • Perplexity
Agent Experience
  • llms.txt
Theme
With AI, Everyone is a Dev. EveryDev.ai © 2026
    1. Home
    2. Developers
    3. Baseten

    Baseten

    Baseten builds an inference cloud and infrastructure stack for bringing AI products to production. It provides fast model runtimes, multi-cloud GPU capacity, deployment tooling, training, observability, and support for mission-critical inference.

    Visit Website

    At a Glance

    1Tool Listed
    7Products
    10Capabilities
    Discussions
    San Francisco, CaliforniaHeadquarters
    2019Est.
    300Employees
    $2BRaised
    Focus Areas
    AI Infrastructure
    Model Management
    Cloud Computing Platforms
    Latest News
    Agentic inference optimization: 50–90% faster enginesOct 2, 2026
    Baseten partners with OpenAI to serve open models through Codex and the Responses APISep 29, 2026
    Markets
    • AI-native startups
    • Enterprise software companies
    • Model labs and model creators
    • Healthcare and clinical AI
    • +4 more

    AI Tools by Baseten

    (1)
    View Baseten
    Baseten tool icon

    Baseten

    AI Model Inference Platform

    AI InfrastructureModel ManagementCloud Platforms

    Discussions

    No discussions yet

    Be the first to start a discussion about Baseten

    Latest News

    10/02/2026

    Agentic inference optimization: 50–90% faster engines

    baseten.co
    09/29/2026

    Baseten partners with OpenAI to serve open models through Codex and the Responses API

    baseten.co
    09/28/2026

    Baseten joins NVIDIA OpenShell/Open Agent Safety Platform effort and supports Blaxel Carbon sandboxes

    baseten.co
    09/25/2026

    Sheila Vashee joins Baseten as Chief Marketing Officer

    baseten.co

    Products & Services

    7
    Dedicated Inference / Dedicated Deployments

    Deploy open-source, custom, and fine-tuned models on dedicated GPUs with configurable hardware, autoscaling, environments, release lifecycle, multi-cloud capacity, and performance engineering.

    Model APIs

    Hosted, pre-optimized open-model APIs with OpenAI-compatible and Anthropic Messages-compatible interfaces, designed for rapid prototyping and production workloads and billed by token.

    Training

    Train or fine-tune models using Loops or Training Jobs on dedicated GPU clusters, then deploy resulting checkpoints on the same inference stack.

    Loops SDK

    Training and post-training tooling for supervised fine-tuning and reinforcement learning, including workflows from LangSmith traces.

    Market Position

    Baseten positions itself as an inference-first, production-grade alternative to generic cloud GPU infrastructure and API-only model hosts: it combines optimized runtimes and performance research with dedicated, multi-cloud, hybrid, and enterprise deployment control. Commonly cited alternatives include Modal, Replicate, RunPod, Together AI, Fireworks AI, Anyscale, and AWS SageMaker; Baseten differentiates on high-performance production inference, customization, and hands-on forward-deployed engineering.

    Leadership

    Founders

    TS

    Tuhin Srivastava

    CEO and co-founder; previously worked on machine learning and co-founded Sutro Health.

    AH

    Amir Haghighat

    CTO and co-founder; previously Head of Engineering at Gumroad and worked in engineering at Clover Health.

    P(

    Philip (Phil) Howes

    Co-founder and Chief Scientist; holds a PhD in mathematics from the University of Sydney and previously co-founded Sutro Health.

    PG

    Pankaj Gupta

    Co-founder and model-performance leader; previously a software engineer at Uber.

    Executive Team

    TS

    Tuhin Srivastava

    CEO and Co-Founder

    Co-founded Baseten in 2019 and previously worked in machine learning and co-founded Sutro Health.

    AH

    Amir Haghighat

    CTO and Co-Founder

    Previously Head of Engineering at Gumroad and an engineer at Clover Health.

    Board of Directors

    JS
    Jay Simons
    Board member; General Partner at BOND

    Founding Story

    The founders started Baseten after repeatedly seeing strong ML models get stuck in deployment hell: productionization took weeks, infrastructure was fragile, and training, serving, scaling, and hardware orchestration were disconnected. They set out to build the integrated platform they wanted to use themselves so builders could bring AI into products and operate it reliably at scale.

    Business Model

    Revenue Model

    Usage-based infrastructure and inference: Model APIs are billed per million input and output tokens; dedicated deployments and training are billed for GPU/CPU time, generally per minute. Enterprise customers can buy dedicated support and cloud, self-hosted, or hybrid deployments.

    Pricing Tiers

    Model APIs
    Per million input and output tokens

    Prices vary by hosted model; the public catalog lists model-specific token rates.

    Dedicated Deployments
    Per-minute compute pricing

    Price varies by selected CPU/GPU hardware and deployment configuration; scale-to-zero can avoid idle GPU charges.

    Training
    Per-minute compute pricing

    Price varies by hardware and training job configuration.

    Enterprise
    Custom

    Dedicated support on Slack and Zoom plus enterprise hosting, networking, security, and capacity options.

    Private company; no IPO announcement or public filing identified.

    Target Markets

    Industries & Segments
    • AI-native startups
    • Enterprise software companies
    • Model labs and model creators
    • Healthcare and clinical AI
    • Legal AI
    • Developer tools and coding agents
    Use Cases
    • High-scale, low-latency LLM inference
    • Agentic coding and compound AI applications
    • Speech transcription, diarization, and text-to-speech
    • Image and video generation
    • Embeddings, reranking, and classification
    • Fine-tuning and reinforcement-learning post-training
    Notable Customers
    • Abridge
    • Cursor
    • Lovable
    • Notion

    Quick Facts

    Headquarters
    San Francisco, California, United States
    Founded
    2019
    Entity Type
    BaseTen Labs, Inc.
    Employees
    300
    Total Funding
    More than $2 billion
    Investors
    Altimeter Capital, Conviction Partners
    Office Locations
    San Francisco
    New York City
    Toronto
    Montreal
    +1 more

    Funding History

    SeedAmount not reported in the cited sources
    2022
    Series AAmount not reported in the cited sources
    2022
    Series BAmount not reported in the cited sources
    2024

    History & Milestones

    January 2026

    Announced a $300M Series E at a $5B valuation, led by IVP and CapitalG with participation from NVIDIA and other existing investors.

    June 22, 2026

    Announced a $1.5B Series F at a $13B valuation, led by Altimeter Capital, Conviction Partners, and Spark Capital, co-led by Sands Capital and Wellington Management.

    September 2026

    Announced partnerships with OpenAI for open-model inference through Codex and the Responses API, and with NVIDIA OpenShell/Blaxel on secure agentic infrastructure.

    February 19, 2025

    Announced a $75M Series C, co-led by IVP and Spark Capital, to build the inference platform for mission-critical AI workloads.

    September 5, 2025

    Announced a $150M Series D led by BOND; Jay Simons joined the board, and Conviction and CapitalG joined the investor group.

    Key Capabilities

    10
    Fast inference runtimes and model-performance research
    TensorRT-LLM, distributed MoE/KV-aware routing, and embedding/reranking/classification engines
    Multi-cloud and multi-region GPU capacity management
    Scale-to-zero, autoscaling, fast cold starts, and 99.99% uptime positioning
    Cloud, self-hosted, hybrid, private-networking, and region/data-residency options
    Logs, metrics, traces, secrets, environment promotion, and integrations with Datadog, Prometheus, Grafana, and New Relic

    Integrations & Partnerships

    Platform Integrations

    • OpenAI SDK and OpenAI-compatible API
    • Anthropic Messages API (beta)
    • Baseten CLI and Python model classes
    • Hugging Face models
    • TensorRT-LLM, vLLM, and SGLang/custom Docker serving
    • Datadog, Prometheus, Grafana, and New Relic observability exports
    • LangSmith/LangChain
    • OpenAI Codex and Responses API

    Key Partnerships

    OpenAI: open models available through the OpenAI B2B Marketplace, Codex, and Responses API
    NVIDIA: launch partner for OpenShell and member of the Open Secure AI Alliance
    Blaxel: collaboration on secure agentic infrastructure and Carbon sandboxes

    Connect

    Website
    baseten.co

    AI Topics

    3

    Baseten focuses on these topics:

    AI Infrastructure(1)
    Model Management(1)
    Cloud Computing Platforms(1)
    Back to all developersSuggest an edit