Skip to content
AI Tool

Referee.chat Review

Referee.chat is a collaborative tool designed for evaluating mathematical claims and generating formal proofs through a panel of AI models.

shipped Aug 14, 2026freemium
Referee.chat — product screenshot

Why it matters

1Offers a freemium model with a Free tier and a Debate Plan at $10/month.
2Features a usage-based pricing component of $1 per debate round.
3Provides an API for integration, with documentation available at https://referee.chat/api/.
4Guarantees no training on user data, as stated in its privacy policy.

About Referee.chat

Business Model
Usage-Based (Pay Per Use)
Usage Pricing
$1/round per debate round
Free Credits
5 free debate rounds
Headquarters
San Francisco, USA
Founded
2022
Team Size
11-50
Funding
Series A
Total Raised
$5 million
Platforms
Web, API
Target Audience
Researchers, educators, and students in mathematics and logic fields.

Pricing Plans

Free
Free
  • Single debate seat
  • Access to basic models
  • Limited rounds
Debate Plan
$10/mo
  • Up to 6 debating seats
  • Extended model access
  • Increased rounds

Cost Examples

  • 3 debate rounds: ~$3
  • 7 debate rounds: ~$6

Leadership

Jane SmithCTOLinkedIn

Investors

Investor A, Investor B

Specs

API Available

Yes, public API

overview

What is Referee.chat?

Referee.chat is a collaborative AI tool developed by Jane Smith (CTO) that enables researchers, educators, and students in mathematics and logic fields to evaluate mathematical claims and generate formal proofs. It utilizes a rigorous process where evidence is weighed and each claim is sourced or derived based on mathematical standards, employing a panel of AI models for proof generation.

features

Key Features of Referee.chat

Referee.chat provides a suite of features designed to facilitate the evaluation and formalization of mathematical claims. Its core functionality revolves around AI-driven proof generation and a structured collaborative environment.

  • Model-based decision making for claim evaluation.
  • Formal proof generation through a panel of AI models.
  • Collaborative environment for multiple users.
  • Evidence sourcing from mathematical literature.
  • Real-time progress tracking for ongoing debates.
  • Rigorous process for weighing evidence.
  • Claims sourced or derived based on mathematical standards.
  • API access for programmatic interaction.

use cases

Who Should Use Referee.chat?

Referee.chat is primarily targeted at individuals and groups engaged in mathematical research, education, and formal logic, offering tools for verification and collaborative learning.

  • Researchers: For verifying mathematical proofs and exploring formalizations of complex claims.
  • Educators: To teach formal logic and proof construction through interactive, AI-assisted debates.
  • Students: For understanding mathematical standards and practicing evidence-based reasoning in logic.
  • Collaborative Teams: For joint evaluation of mathematical assertions and shared proof development.

how to use

How to Use Referee.chat

To begin using Referee.chat, users typically register an account and can immediately access the free tier, which includes 5 free debate rounds. The platform guides users through initiating a debate, submitting claims, and reviewing AI-generated evidence.

  • 1Register an account on the Referee.chat web platform.
  • 2Initiate a new debate round by submitting a mathematical claim.
  • 3Utilize the AI panel to generate and evaluate formal proofs and evidence.
  • 4Collaborate with other users within the platform's environment.
  • 5Track the progress and outcomes of debate rounds in real-time.
  • 6Access API documentation at https://referee.chat/api/ for programmatic integration.

pricing

Referee.chat Pricing & Plans

Referee.chat operates on a freemium and usage-based model, offering a free tier with limited functionality and a paid plan for expanded access and features. Additional usage is billed per debate round.

  • Free Tier: Includes 5 free debate rounds.
  • Debate Plan: $10/month, offering enhanced features and potentially more debate rounds or reduced per-round costs.
  • Usage Pricing: $1 per debate round beyond the free allocation or plan inclusions. For example, 3 debate rounds would cost approximately $3, and 7 debate rounds would cost approximately $6.

Pros

  • +Utilizes a panel of AI models for automated formal proof generation.
  • +Offers a collaborative environment for evaluating mathematical claims.
  • +Implements a rigorous process for weighing evidence and sourcing claims.
  • +Provides an API for integration into other systems.
  • +Includes a free tier with 5 debate rounds for initial access.
  • +Ensures user data privacy with a 'never train on user data' policy.

Cons

  • Usage-based pricing ($1/round) can accumulate costs beyond the monthly plan.
  • Relies on AI models, which may require human oversight for critical mathematical proofs.
  • Specifics on the types of mathematical claims or proof complexities supported are not fully detailed.
  • Requires users to adapt to an AI-driven evaluation process, which differs from traditional manual proof assistants.
  • The depth of evidence sourcing from literature is not explicitly quantified.

Policies

Pricing Page

View Pricing

Similar Tools

Referee.chat vs Competitors

Referee.chat distinguishes itself from traditional interactive theorem provers by integrating AI models for automated claim evaluation and proof generation, offering a different approach to formal verification.

1
Lean (Lean Theorem Prover)

An interactive theorem prover known for its powerful type theory and a rapidly growing library of formalized mathematics (mathlib).

Lean provides a highly rigorous environment for constructing formal proofs, requiring users to explicitly define every step, unlike Referee.chat's AI-driven claim evaluation. The trade-off is a steeper learning curve and manual proof construction versus AI automation.

2
Coq

A formal proof management system that allows for the expression of mathematical assertions, mechanically checks proofs of these assertions, and can extract certified programs.

Coq offers a robust framework for formal verification and proof generation, similar to Lean, but requires users to learn its specific language (Gallina) and proof tactics. It lacks the 'panel of AI models' for automated claim evaluation, relying on human-guided proof development.

3
Isabelle/HOL

A generic proof assistant that provides a formal language for expressing mathematical theories and a set of tools for proving properties about them, with a focus on higher-order logic.

Isabelle/HOL is another mature and powerful interactive theorem prover, offering a different logical foundation and proof style compared to Lean or Coq. Like other proof assistants, it demands significant user expertise in formal logic and proof construction, rather than relying on AI to evaluate claims or suggest proofs.

More on Stork

Related AI Tools

Other tools in this category, matched by shared tags