Free
Free
Research site and open-source GitHub repositories are published without commercial plan pricing.
- Access
- Research site and open-source GitHub repositories are published without commercial plan pricing.
Research · Vendor site 2026-10-04
LLMEval is a Fudan NLP Lab research series building comprehensive LLM evaluation frameworks across academic disciplines, medical AI, and large generative question sets. It publishes papers, leaderboards, and open-source datasets such as LLMEval-Logic with solver-verified answers.
NLP researchers benchmarking LLMs for fairness, robustness, and domain-specific reasoning tasks.
Offerings are academic benchmarks and papers rather than a hosted paid evaluation SaaS.
LLMEval aggregates rigorous evaluation frameworks, leaderboards, and open datasets spanning disciplines, medical AI, and logic reasoning benchmarks.
What it can do
Unknown is different from unavailable. Each fact carries its own evidence.
| Capability | Value | Evidence | Checked |
|---|---|---|---|
| Benchmark scale | Homepage cites 59 LLMs benchmarked and 220K questions in LLMEval-Fair alongside multiple peer-reviewed papers. | Facts sourced | 2026-10-04 |
| LLMEval-Logic release | Featured research announces open-source LLMEval-Logic with Z3-verified answers and adversarial hardening. | Facts sourced | 2026-10-04 |
Understand the total cost
Free
Free
Research site and open-source GitHub repositories are published without commercial plan pricing.
Consensus is an AI academic search engine over peer-reviewed literature. It supports paper search plus Pro and Deep modes that synthesize evidence, study snapshots, and collections, with optional API and MCP access on paid tiers for students, clinicians, and researchers.
Explore toolEnago Read (formerly RAxter) is an AI literature-review workspace for researchers. It summarizes papers section by section, surfaces key insights, finds related literature from a large database, and adds a Copilot chat mode for deeper paper Q&A and collaboration.
Explore toolHumata is a PDF and document AI that lets teams upload files, ask questions across their knowledge base, summarize long papers, compare documents, and embed answers on webpages. It targets researchers and professionals who need cited Q&A over private files with team permissions on higher tiers.
Explore toolThe practical questions
LLMEval is a Fudan NLP Lab research series building comprehensive LLM evaluation frameworks across academic disciplines, medical AI, and large generative question sets. It publishes papers, leaderboards, and open-source datasets such as LLMEval-Logic with solver-verified answers.
Research site and open-source GitHub repositories are published without commercial plan pricing.. This record lists ongoing free access; check the plan limits before starting.
No paid monthly price is listed; this record treats the product as free to start. See the plan cards for entitlements, billing commitments and seat minimums.
Not confirmed. API access and subscription access may have different terms; consult the linked sources.
Offerings are academic benchmarks and papers rather than a hosted paid evaluation SaaS.
No retained pricing changes yet. A current price alone does not establish a historical trend.
Reviewed vendor source
Read original source ↗Facts apply to the named version and check date. Send a sourced correction if something changed.