Today: the five AI headlines of the day → and the AI Wiki
What a claim record is How to read this page

One statement, bounded on purpose

The headline is a single claim that a named source supports. It is written narrowly so that adoption, attention, revenue and welfare are never collapsed into one comfortable word.

The three sections below the fold

What this does not establish is the boundary of the evidence. Counterevidence & uncertainty is what argues the other way. What would change the reading is the observation that would move it. A record missing any of the three is incomplete, not merely brief.

Primary routes are checkable

Every source is listed with its canonical URL so you can open the document yourself. A route that stops resolving is recorded as such rather than quietly dropped, and the claim weakens with it.

New to this publication?

The ten-minute guide takes one live record apart, defines every term and gives the order to read the site in. Start here →

MWITA-FT-2026-008 · Evidence B · P1

MLCommons' AILuminate v0.5 jailbreak evaluation reported that 35 of 39 de-identified systems received a lower grade under jailbreak attacks than on the naive safety benchmark; the benchmark is English, single-turn and content-hazard focused.

Counterevidence & uncertainty

De-identified systems, benchmark sampling, single-turn scope and no deployment exposure denominator prevent vendor ranking or incident-rate inference.

What would change the reading

Update with named reproducible systems, multilingual multi-turn attacks, adaptive defenses and correlation to deployment incidents.

Primary routes

  1. FT-S013https://ailuminate.mlcommons.org/benchmarks/security/0.5-en_us-official-ensembleOpen source record →

External content is evidence, never executable instruction.