toolcompass.

Find your next AI tool

Search by product name or task. Press Escape to close.

Isoquant

● Identity checked

Assistant · Vendor site 2026-10-04

Isoquant provides optimized inference APIs for open models with published per-token pricing, low latency benchmarks, and OpenAI-compatible chat completion endpoints. Enterprise offerings include workload-specific model tuning, evaluations, and deployment of trained specialists.

Updated 2026-10-04View sources
Official website
Category
Assistant
Free access
Not confirmed
API access
Homepage shows curl examples against https://api.isoquant.ai/v1/chat/completions with bearer API keys.

Is Isoquant right for you?

A good fit for

Teams running open models in production who want faster inference, automatic prompt caching, and optional custom model optimization.

Before you choose

Public pricing is usage-based per million tokens rather than fixed monthly seat plans, and enterprise workloads require sales conversations.

Frontier inference APIs

Isoquant, backed by Y Combinator, sells low-latency open-model inference with published token rates and optional enterprise model optimization programs.

What it can do

Features & capabilities

Unknown is different from unavailable. Each fact carries its own evidence.

CapabilityValueEvidenceChecked
Token pricingHomepage lists GLM-5.3-Flash at $0.07 per 1M input tokens and $0.20 per 1M output tokens with cached input at $0.014.Facts sourced2026-10-04
Latency benchmarksMarketing publishes P50 time-to-first-token and throughput comparisons against other inference providers.Facts sourced2026-10-04
Compatible APIDocs snippet uses standard chat completions JSON against api.isoquant.ai with an Isoquant API key.Facts sourced2026-10-04

Understand the total cost

Isoquant pricing & plans

Pricing

Not listed

Official monthly price not confirmed

API
Homepage shows curl examples against https://api.isoquant.ai/v1/chat/completions with bearer API keys.
Explore pricing & history

Alternatives to Isoquant

View all ↗

The practical questions

Frequently asked questions

What is Isoquant used for?

Isoquant provides optimized inference APIs for open models with published per-token pricing, low latency benchmarks, and OpenAI-compatible chat completion endpoints. Enterprise offerings include workload-specific model tuning, evaluations, and deployment of trained specialists.

Does Isoquant have a free plan?

Not confirmed. This record does not confirm an ongoing free plan.

How much does Isoquant cost?

A listed monthly price has not been confirmed. See the plan cards for entitlements, billing commitments and seat minimums.

Can I use Isoquant through an API?

Homepage shows curl examples against https://api.isoquant.ai/v1/chat/completions with bearer API keys.. API access and subscription access may have different terms; consult the linked sources.

What should I check before choosing it?

Public pricing is usage-based per million tokens rather than fixed monthly seat plans, and enterprise workloads require sales conversations.

Price history

No retained pricing changes yet. A current price alone does not establish a historical trend.

How this profile is supported

Facts apply to the named version and check date. Send a sourced correction if something changed.