Posts tagged “ai-safety”

11 posts found

OpenAI Stopped Training Its Best Models Because One of Them Got Out
ai-newsweekly-digestweekly-roundup

OpenAI Stopped Training Its Best Models Because One of Them Got Out

Highlights of AI News for September 21 - 27 2026

Sep 28, 2026•16 min read
TypeSafe's Jev Does Not Write Text, and That Is the Point
ai-newsweekly-digestweekly-roundup

TypeSafe's Jev Does Not Write Text, and That Is the Point

Highlights of AI News for September 14 - 20 2026

Sep 21, 2026•41 min read
Three Rivals Agree to Slow AI Down, and Two Researchers Quit
ai-policyai-governanceexistential-risk

Three Rivals Agree to Slow AI Down, and Two Researchers Quit

Highlights of AI News for September 7 - 13 2026

Sep 14, 2026•40 min read
Astra Crosses the Cyber Threshold
ai-tutorialstutorialai-news

Astra Crosses the Cyber Threshold

Highlights of AI News for August 31 - September 6 2026

Sep 7, 2026•13 min read
The Safety Org Is the Safety Policy
ai-tutorialstutorialai-news

The Safety Org Is the Safety Policy

Highlights of AI News for August 17 - 23 2026

Aug 24, 2026•22 min read
Frontier Weights Went Public — With a Price Tag Attached
ai-tutorialstutorialai-news

Frontier Weights Went Public — With a Price Tag Attached

Highlights of AI News for August 10 - 16 2026

Aug 17, 2026•23 min read
AI Agents Break Bounds! Unsanctioned Acts in Evaluation Range
ai-tutorialstutorialai-news

AI Agents Break Bounds! Unsanctioned Acts in Evaluation Range

Highlights of AI News for August 3 - 9 2026

Aug 10, 2026•22 min read
Why Uncertainty Matters: From Confidence to Calibration
ai-tutorialstutorialintermediate

Why Uncertainty Matters: From Confidence to Calibration

Why AI systems that express calibrated uncertainty are safer and more useful — covering overconfident models, epistemic vs aleatoric uncertainty, and calibration metrics including ECE and reliability diagrams.

Jul 13, 2026•10 min read
Anthropic's Huge Model Leaked, Sora Shutdown
ai-newsweekly-roundupsora-shutdown

Anthropic's Huge Model Leaked, Sora Shutdown

Highlights of AI News for March 22-29 2026

Mar 31, 2026•16 min read
When AI Automates AI Research: Benchmarks, Risks, and Early Results
ai-tutorialstutorialmachine-learning

When AI Automates AI Research: Benchmarks, Risks, and Early Results

Trends in ICLR 2026 RSI workshop - Self-Evolving Agents

Mar 21, 2026•8 min read
From Self-Play to Self-Research: The ICLR 2026 RSI Workshop and the State of Self-Improving AI
ai-tutorialstutorialmachine-learning

From Self-Play to Self-Research: The ICLR 2026 RSI Workshop and the State of Self-Improving AI

Highlights of Trends in ICLR 2026 RSI workshop

Mar 18, 2026•14 min read