← Back to the directory

Research Analysis / RESEARCH PROFILE

ForecastBench

A benchmark platform for evaluating language-model forecasts against human forecasts on real-world questions.

AI unclearTool
Visit website

Source check
Oct 8, 2026

OUR TAKE / EDITORIAL OPINION

ForecastBench asks the question a forecast demo cannot answer

Evaluation infrastructure deserves its own place in an AI directory.

ForecastBench is a benchmark for evaluating language-model forecasts against human forecasts on real-world questions. It is a benchmark, not a forecasting agent or an agent integration. Our view is that this distinction is the reason to pay attention. A persuasive example can show what a system produces; an evaluation framework creates a place to ask how much confidence that example deserves.

We would consider ForecastBench for researchers, developers and buyers trying to assess evidence about forecasting systems. The tradeoff is standardized comparison versus the specific conditions of a proposed application. A result should be interpreted within its question set and evaluation procedure, rather than silently extended to a different market, time horizon or operational workflow.

For evaluation, choose the exact benchmark result relevant to the model under consideration and inspect the documented setup before comparing scores. Record which parts of the intended use are represented and which are absent. Ask whether the conclusion would change under a narrower, application-specific set of questions. We have not independently reproduced benchmark results. Our criterion is whether the benchmark helps a reader make a more bounded claim about a model; it should make product judgments more careful, not provide a universal seal of forecasting competence.

AI-assisted editorial analysis based on cited documentation; not a hands-on product test.

Sources for this opinion: Source 1 · Source 2

Official product artwork from ForecastBench.
Official product artwork from ForecastBench. Image source

The overview

Helps researchers compare model forecasting results and evaluation approaches using a dedicated benchmark.

Company
Not yet verified
Founded
Not yet verified
Product launch
Not yet verified
Headquarters
Not yet verified
Founding team
Not yet verified
Product status
Not yet verified

Company founding and product launch are different events. Dates show only the precision supported by a source.

What is documented about AI?

ForecastBench is an automated benchmark that evaluates LLM and human forecasts; it is not an agent integration or forecasting agent.

“No AI documented” means the checked sources do not establish an AI feature. It does not prove that the product uses no AI.

Pricing & access

Pricing Not Yet Verified

Officially checked pricing was not established.

Platforms: web

  • forecasting
  • benchmark
  • llm-evaluation
  • research

The source record2

Official product documentation reviewed. No hands-on product testing was performed.

  1. Checked Oct 8, 2026
  2. Checked Oct 8, 2026