← Founder Notes
Archive

The cheapest model just won the routing vote. langchain's open recipe for picking models per task…

Yethikrishna ROriginal on Threads

the cheapest model just won the routing vote. langchain's open recipe for picking models per task cut openswe's median thread cost 64 percent without hurting merge rates, per october reporting.

the smart part of the agent is now the dispatch.

Context

Verified event date: LangChain published How to Build a Model Router in the Harness on Oct 1, 2026. It says that in its experiments, model routing cut median cost per coding task by 64 percent compared with the baseline, with no noticeable change in quality, using a router built into the harness of Open SWE, its open-source coding agent.

Falkster's summary says the router picks one of three model tiers on the first message, and that it was tested in an A/B on LangChain's own engineers.

How it compares

The 64 percent median cost cut and Open SWE match LangChain's post. LangChain says 'no noticeable change in quality'; 'merge rates' as the quality measure is not seen in the excerpts read, so unsupported here, not refuted. The post says 'per october reporting', which fits Oct 1. 'The dispatch is the smart part' is the author's.

'the smart part of the agent is now the dispatch' is the author's opinion.

Related work

Watch next

  • Read LangChain's post for how it measured quality.

Sources

  1. LangChain, Oct 1, 2026langchain.com
  2. dev48, Oct 1, 2026dev48.org
  3. Falkster, Label before you routefalkster.com

Provenance

The note above is reproduced unedited from the original post, first published on Threads on 11 October 2026 at 06:16 IST. Sources are the papers and datasets the note draws on.

View the original post
Embed this note
<iframe src="https://founder.myndlabs.tech/notes/embed/the-cheapest-model-just-won-the-routing-vote-DeVZ0kpjE7K" width="480" height="420" style="border:0;max-width:100%" loading="lazy" title="The cheapest model just won the routing vote. langchain's open recipe for picking models per task…"></iframe>

More notes