Pricing

Simple, usage-based pricing.

Pay per token on the fastest model on Earth. Runs inside AWS and GCP, in your region - one API call, nothing to deploy.

Most popular Pay as you go
$0.2/ M input tokens
$0.7/ M output tokens

Full-speed Celeris-1. Pay for what you use.

  • +Usage-based billing
  • +Production rate limits
  • +Streaming, OpenAI-compatible API
  • +Email support
Get started
Enterprise
Custom

For teams running Celeris on the critical path at scale.

  • +Volume pricing & dedicated clusters
  • +Custom rate limits
  • +SSO · SOC 2 · VPC deployment
  • +Solutions engineering, dedicated support & SLA
Contact sales

Prices in USD. Billed per million tokens; input and output metered separately.

Frequently asked questions

What is Celeris-1?

OPEN

Celeris-1 is the first model from Celeris: a general purpose language model that delivers near-GPT-5 level intelligence with up to 15x faster response times. It scores 75.9% on MMLU-Pro at a p50 latency of 158ms.

How can it be this fast without giving up intelligence?

OPEN

Celeris-1 is built from the ground up for low latency rather than adapted for it after the fact. We measure intelligence and speed on the same footing and publish the methodology, so you can verify the tradeoff yourself: 75.9% on MMLU-Pro at a p50 latency of 158ms.

Do you benchmark against other models?

OPEN

Yes, and we publish everything. One harness, identical prompts and scoring across GPT-5, GPT-5-mini, Gemini 2.5 Flash, and Celeris-1. Every claim on this site links back to that data.

Where does low latency actually matter?

OPEN

Voice, agents, and real-time applications. Anywhere waiting for a model breaks the experience, Celeris-1 is built to remove the wait.

Do all plans run at the same speed?

OPEN

Yes. Every plan runs the same full-speed Celeris-1: same weights, same latency, same quality. Speed is the product, not a premium tier, so what you test is exactly what you ship.

How does billing work?

OPEN

Per million tokens, input and output metered separately. No reserved capacity, no idle GPU charges, no surprises at the end of the month.

Can we run it inside our own cloud?

OPEN

Enterprise plans can deploy in your VPC, in your region, so requests never leave your cloud's network. Talk to us about compliance and data residency requirements.

How do I get access?

OPEN

Celeris-1 is available starting today. Sign up, get an API key, and start building.

Start in minutes.
Scale when you're ready.