Pricing
Not listed
Official monthly price not confirmed
- API
- Homepage describes running benchmarks and environments through a CLI or API with exportable JSON, CSV, or HTML reports.
Research · Vendor site 2026-10-04
Kimpton provides evaluation infrastructure for models and agents, including private environments, behavioral benchmarks, and Koliseum-style live evaluations. Teams run benchmarks via CLI or API, compare checkpoints on long-horizon and multi-agent tasks, and export structured reports.
ML teams evaluating frontier models and agents who need private environments and reproducible behavioral benchmarks.
Custom private benchmarks and enterprise evaluations require contacting the team rather than checkout pricing.
Kimpton records agent trajectories inside environments so teams compare model checkpoints on realistic, long-running tasks instead of static prompts alone.
What it can do
Unknown is different from unavailable. Each fact carries its own evidence.
| Capability | Value | Evidence | Checked |
|---|---|---|---|
| Eval infrastructure | Site positions Kimpton as evaluation infrastructure with live environments and behavioral benchmarks. | Facts sourced | 2026-10-04 |
| CLI or API | Marketing states teams can run evaluations through a CLI or API and export scores. | Facts sourced | 2026-10-04 |
| Long-horizon tasks | Homepage highlights measuring planning, memory, and recovery across extended action sequences. | Facts sourced | 2026-10-04 |
Understand the total cost
Pricing
Not listed
Official monthly price not confirmed
FutureSearch provides AI forecasting and research agents with a public track record, letting users ask probability or numeric questions about future outcomes. Plans combine free starting credit with paid tiers that raise concurrent researcher limits and discount usage rates.
Explore toolHumata is a PDF and document AI that lets teams upload files, ask questions across their knowledge base, summarize long papers, compare documents, and embed answers on webpages. It targets researchers and professionals who need cited Q&A over private files with team permissions on higher tiers.
Explore toolsupermemory provides memory and continual-learning infrastructure for AI agents, including storage, retrieval, connectors, and team billing through a credit-based console. Developers use it to keep long-lived context across chats, documents, and production workloads.
Explore toolThe practical questions
Kimpton provides evaluation infrastructure for models and agents, including private environments, behavioral benchmarks, and Koliseum-style live evaluations. Teams run benchmarks via CLI or API, compare checkpoints on long-horizon and multi-agent tasks, and export structured reports.
Not confirmed. This record does not confirm an ongoing free plan.
A listed monthly price has not been confirmed. See the plan cards for entitlements, billing commitments and seat minimums.
Homepage describes running benchmarks and environments through a CLI or API with exportable JSON, CSV, or HTML reports.. API access and subscription access may have different terms; consult the linked sources.
Custom private benchmarks and enterprise evaluations require contacting the team rather than checkout pricing.
No retained pricing changes yet. A current price alone does not establish a historical trend.
Reviewed vendor source
Read original source ↗Facts apply to the named version and check date. Send a sourced correction if something changed.