Skip to main content
wideriver_

$ watch the frontier --daily --verified

Which model should you actually reach for? Verified scores from official releases and papers, news from seven primary sources — readings for people who build, tinker, and experiment on their own. No hot takes, no aggregator soup.

From the Desk

the Ledger

496
Models measured
2347
Articles tracked
Claude Opus 4.7
Ledger leader
7
Tonight's wire
  1. 1
    Claude Opus 4.7
    Anthropic · 5 benchmarks
    92
  2. 2
    GPT-5.4 Pro
    OpenAI · 4 benchmarks
    77
  3. 3
    GPT-5.5
    OpenAI · 6 benchmarks
    69
  4. 4
    Gemini 3.1 Pro Preview
    Google · 6 benchmarks
    68
  5. 5
    Claude Opus 4.6
    Anthropic · 5 benchmarks
    33

Method: mid-rank percentile within each benchmark's population · ties by number of eligible benchmarks

Full ledger →

the wire

7 items on tonight's wire

Full wire →

The Evening Briefing, in your inbox by 7.

One quiet email a day — the verified movers and the wire that matters — read before your first coffee finishes.

We'll send you a confirmation email. No spam, unsubscribe anytime.