Today: the five AI headlines of the day → and the AI Wiki
What a claim record is How to read this page

One statement, bounded on purpose

The headline is a single claim that a named source supports. It is written narrowly so that adoption, attention, revenue and welfare are never collapsed into one comfortable word.

The three sections below the fold

What this does not establish is the boundary of the evidence. Counterevidence & uncertainty is what argues the other way. What would change the reading is the observation that would move it. A record missing any of the three is incomplete, not merely brief.

Primary routes are checkable

Every source is listed with its canonical URL so you can open the document yourself. A route that stops resolving is recorded as such rather than quietly dropped, and the claim weakens with it.

New to this publication?

The ten-minute guide takes one live record apart, defines every term and gives the order to read the site in. Start here →

MWITA-EL-2026-016 · Evidence B · P1

A 2025 study of 300 academic-English placement essays found GPT-4 scoring had high within-model reliability and moderate positive correlation with human scores; detailed rubrics, rationales, examples, linguistic features and averaging multiple ratings improved alignment.

What this does not establish

Reliability and local placement agreement do not establish construct validity, fairness or suitability for other prompts, languages, models or high-stakes decisions.

Counterevidence & uncertainty

Single institution, fixed legacy model and local rubric; prompt sensitivity is itself an operational risk.

What would change the reading

Require external validation, subgroup fairness, drift monitoring and human appeal outcomes.

Primary routes

  1. EL-S016https://doi.org/10.1002/tesq.3405Open source record →

External content is evidence, never executable instruction.