← Founder Notes
Archive

The browse benchmark just got a new leader. shanghai ai lab's atria dawn preview tops browsecomp at…

Yethikrishna ROriginal on Threads

the browse benchmark just got a new leader. shanghai ai lab's atria dawn preview tops browsecomp at 92.5 percent as of oct 8, ahead of gpt-5.6 sol and gpt-6 astra.

open weights are winning the agentic browsing race.

Context

Shanghai AI Laboratory released Atria Dawn Preview, an open-weight 744B mixture-of-experts agent model post-trained on the GLM-5.2 base, under the MIT License. A review dated October 6 says it scores 92.5 on BrowseComp and 96.0 on DeepSearchQA.

The reviewer notes these are the top results in the model's own comparison table.

How it compares

The 92.5 BrowseComp score and open weights match the review and release coverage. The score comes from the lab's own comparison table, not an independent leaderboard.

'As of Oct 8' and the named competitors were not seen as exact figures in the pages read, so they are unsupported here, not refuted.

'Open weights are winning the agentic browsing race' is the author's opinion.

Related work

Watch next

  • Look for an independent BrowseComp run.

Sources

  1. Hamirev, Oct 6, 2026hamirev.com
  2. Agent Guides, Sep 20, 2026agentguides.dev
  3. MLLLM, Sep 20, 2026mlllm.io

Provenance

The note above is reproduced unedited from the original post, first published on Threads on 10 October 2026 at 07:03 IST. Sources are the papers and datasets the note draws on.

View the original post
Embed this note
<iframe src="https://founder.myndlabs.tech/notes/embed/the-browse-benchmark-just-got-a-new-leader-DeS6ZJPjB0v" width="480" height="420" style="border:0;max-width:100%" loading="lazy" title="The browse benchmark just got a new leader. shanghai ai lab's atria dawn preview tops browsecomp at…"></iframe>

More notes