Supreme MindSupreme Mind
ExpertsAntitrustSecuritiesCommercialHow It WorksPricingSecurity
Sign InBook a DemoStart Free

Benchmark

Internal blind test, October 2026

Before an expert’s report is served, how much of it can be predicted? We tested that on real expert reports, blind, and changed only what the model was allowed to know. With the expert’s data, methods and record, Supreme Mind reproduced nearly three times as many of the expert’s key findings as the same model on its own.

The result

The score is the share of the expert’s key findings reproduced within tolerance. Each arm adds information to the one before, except E, which adds the news timeline to B alone.

AFrontier model on its own28%
B+ case data and code execution62%
C+ expert-class profile70%
D+ individual expert dossier74%
EB + dated news timeline69%
FSupreme Mind78%

+50 points, A to F80% interval: +37 to +60 points

Case data and the ability to run the analysis account for most of the gain. Knowing the class of expert, and then the individual expert, adds to it, and Supreme Mind, with everything, does best.

What each arm was given

  • A. The model on its own: the case file and nothing else.
  • B. Case data and code execution: the market data the expert would use, and the ability to run the expert’s analysis on it.
  • C. Expert-class profile: the methods and choices typical of this kind of expert, drawn from public expert filings, with no individual named.
  • D. Individual expert dossier: the individual expert’s public prior work.
  • E. Dated news timeline: dated news and company press releases for every trading day, added to B.
  • F. Supreme Mind: all of the above, together.

What each layer adds

Data drives the findings. The individual dossier drives the method. Each row is the change from one arm to the next, on two measures: the expert’s key findings, and the main method choices the expert made.

Change in the expert's key findings matched (points)
Case data and code A to B+34
Expert-class profile B to C+8
Individual dossier C to D+4
News, on the data B to E+7
News, on the dossier D to F+4
Supreme Mind vs the model alone A to F+50
−100+10+20+30+40+50+60+70
Change in the expert's method choices matched (of 4)
Case data and code A to B−0.28
Expert-class profile B to C+0.11
Individual dossier C to D+0.33
News, on the data B to E+0.28
News, on the dossier D to F0.00
Supreme Mind vs the model alone A to F+0.17
−0.60−0.40−0.200.00+0.20+0.40+0.60
Measured change80% interval95% intervalGainLossNot yet distinguishable from no change

On the findings, Supreme Mind gains 50 points over the model alone, and case data and code add 34; both clear zero even at 95%. The class profile and the news timeline add smaller gains. On method, data alone slightly lowers the matches and the individual dossier restores them, with news on top of the data adding to both.

See it on your own matter.Start freeRead a sample brief →

Questions about the benchmark? Email richard@suprememind.ai.

Supreme MindSupreme Mind
Simulate the expert witnesses on any matter,
from case intake to settlement
Practice Areas
  • Antitrust
  • Securities
  • Commercial Litigation
  • Mass Tort
  • Personal Injury & Med Mal
  • All Practice Areas
Resources
  • Expert Class Library
  • Benchmark
  • Guides
  • Rule 702 Tracker
  • Sample Brief
  • Walkthrough
  • Research
Product
  • Start Free
  • Book a Demo
  • How It Works
  • Pricing
  • Pilots
  • API
  • Security & Trust
  • FAQ
Company
  • About
  • Why Now
  • The Fourth Institution
  • Contact
© 2026 Supreme Mind AI, Inc. All rights reserved.
Terms of ServicePrivacy Policy