The reporting

3 articles

Topic: ai safety

A stylized blue human head and brain appear over a dense pattern of electronic circuit traces.

CLAIM CHECK Artificial Intelligence

Hubinger’s AI Extinction Estimate Is a Personal Forecast, Not a Measured Result

Hubinger says present-model risk is low, but his posts supply no quantitative derivation for his greater-than-10-percent superintelligence forecast.

Chart showing where constitutional midtraining alignment gains remained after later training

NEWS AI Science

Oxford Researchers Put AI Principles Earlier in Training, and Some of Them Stuck

A 120-billion-parameter experiment found that adding constitutional content before conventional safety tuning produced alignment gains that survived later training. The strongest results are promising, but narrower than a permanent solution to AI alignment.