Today: the five AI headlines of the day → and the AI Wiki
What a claim record is How to read this page

One statement, bounded on purpose

The headline is a single claim that a named source supports. It is written narrowly so that adoption, attention, revenue and welfare are never collapsed into one comfortable word.

The three sections below the fold

What this does not establish is the boundary of the evidence. Counterevidence & uncertainty is what argues the other way. What would change the reading is the observation that would move it. A record missing any of the three is incomplete, not merely brief.

Primary routes are checkable

Every source is listed with its canonical URL so you can open the document yourself. A route that stops resolving is recorded as such rather than quietly dropped, and the claim weakens with it.

New to this publication?

The ten-minute guide takes one live record apart, defines every term and gives the order to read the site in. Start here →

MWITA-MB-2026-007 · Evidence A · P1

METR defines task-completion time horizon as the human-duration point at which an agent is predicted to succeed at a specified reliability under a particular task suite and scaffold.

Counterevidence & uncertainty

Task selection, human-time labels, scaffold, success grader and reliability level determine the estimate; it does not include organizational integration or downstream harm.

What would change the reading

Update when task mix, scaffold, grader or reliability threshold changes.

Primary routes

  1. MB-S007https://metr.org/time-horizons/Open source record →

External content is evidence, never executable instruction.