1endpoint
1endpoint is a unified AI inference gateway that gives developers one API for accessing multiple AI model providers and formats. Its stated focus is compatible integrations, lower per-token costs, prompt caching, and spend tracking.
At a Glance
- Individual developers
- Development teams
- AI-agent and coding-agent users
- Teams building AI applications and automation
- +1 more
AI Tools by 1endpoint
(1)1endpoint
Unified AI API Gateway for LLMs
Discussions
No discussions yet
Be the first to start a discussion about 1endpoint
Latest News
Products & Services
A single API root (https://1endpoint.dev/api/v1) for routing requests to enabled models. It supports OpenAI-compatible Chat Completions, Responses, and Embeddings routes plus Anthropic-compatible Messages and token counting; users change the model ID rather than rewriting provider integrations.
Market Position
1endpoint positions itself as a low-cost, compatibility-first alternative to integrating separately with official model-provider APIs: one base URL, multiple model families, exact-model routing, and token-level pricing. Its differentiation claims center on prompt caching, transparent component rates, and switching models without rewriting client integrations.
Leadership
Founders
DustinPham12 / dustin2
The builder who introduced 1endpoint on Hacker News and VOZ.vn; the posts describe building it with a friend/team after experiencing high AI subscription and API costs while coding and running agents. Public sources identify the person by these handles rather than a verified legal name or prior employer.
Founding Story
The founder said that frequent AI coding and agent use made subscriptions and direct API bills expensive, while unused subscription capacity was wasteful. He and a team built 1endpoint to consolidate multiple model providers behind one API, reduce integration work, and optimize costs through caching and lower rates.
Business Model
Revenue Model
Usage-based API billing: customers preload credits and are charged for model token consumption. Input, cached input, and output are priced independently; the homepage states that 1,000 credits equal $1 and that there is no blended platform fee.
Pricing Tiers
No monthly subscription is required; users pay for actual token usage. Public model rates shown on the homepage start at $0.0300 per 1 million input tokens for GLM 5.3 Flash, with cached and output rates varying by model.
Target Markets
- Individual developers
- Development teams
- AI-agent and coding-agent users
- Teams building AI applications and automation
- Users seeking alternatives to direct, higher-cost official model APIs
- Coding agents and developer tools
- AI application development using multiple providers
- Automation workloads
- Long, multi-turn conversations where prompt caching lowers repeated-context costs
- Teams comparing model cost/performance or switching models without code redeploys