FrontierThe story, in brief

GPT and Claude failed Bridgewater's finance tests because the right answers were never public

GPT and Claude just got beaten at their own game. A fine-tuned open model outperformed them on Bridgewater's finance benchmarks—at a fraction of the cost.

Illustration of a transparent lens revealing connected networks across layers of paper.
Exploring the next frontier of AI research.AI illustration by KeyNews
The KeyNews take

Why it matters

Closed-model dominance is cracking in specialized domains. When evaluation data is domain-specific and unavailable publicly, fine-tuned open models can outperform frontier models—and that changes ROI calculus for enterprises.

The key facts

5 to know
  1. Fine-tuned open-weight model outperformed GPT and Claude on Bridgewater's financial document evaluation

  2. Performance achieved at a fraction of the cost of frontier models

  3. Evaluation based on domain-specific financial data not in public training sets

  4. Bridgewater and Thinking Machines Lab conducted the analysis

  5. Published July 3, 2026

Go to the source

The Decoderthe-decoder.com

Publisher excerpt: The hedge fund Bridgewater and Thinking Machines Lab report that a finely tuned open-weight model outperforms the most powerful AI models in the evaluation of financial documents, at a fraction of the cost. The figures come from their own analysis.
Read original report
Back to today's editionMore frontier news

Keep reading

Related stories

More from Frontier