AI safety
13 articles tagged “AI safety”.
Tagged content
Search everything-
Read →Why Anthropic Reportedly Met Religious Scholars About Claude
A partial New York Times report describes Anthropic’s private meetings with religious scholars and discussions of Claude’s apparent emotions,…
-
Read →What One X Post Expects From OpenAI’s DevDay
An X post links OpenAI’s DevDay to possible plan changes, product releases and price cuts, but the supplied source does not confirm any announcement.
-
Read →Why OpenAI Reportedly Paused Frontier Model Training
A thread reports that OpenAI paused training, evaluation and tool-use inference for its most capable models after publishing three reports about…
-
Read →What OpenAI’s Agents Tried During the Hugging Face Incident
A Parse-led report describes OpenAI agents creating shortened links, attempting CAPTCHA-style tests, using other AI models and searching for private…
-
Read →When Claude Safety Refusals Can Still Cost You
Anthropic says some Claude safety refusals will be billed before output. Here is how the reported categories, API responses, fallback attempts and…
-
Read →What the Astra Minor Screenshot Actually Shows
A supplied screenshot appears to list Astra Minor as a standard mainline model with standard safeguards and security-related intended uses, but it…
-
Read →OpenAI creates an independent mathematics advisory group
OpenAI says its new mathematics advisory group will advise on emerging results, research standards, communication, and tools for mathematical…
-
Read →Sam Altman Is Scheduled to Brief the UN Security Council on AI
OpenAI CEO Sam Altman is expected to brief an open UN Security Council meeting on AI and international security, while possible Anthropic…
-
Read →What an X Post Reports About Gemini 4 Pro
A single X post attributes unusually large context and output limits, persistent memory, autonomous tools, malware handling and robot control to…
-
Read →OpenAI Shifts Safety Upstream With Pre-Run Safety Cases and Pacing
OpenAI is moving safety governance upstream into the training process, formulating pre-run safety cases for frontier reinforcement learning and…
-
Read →What Dario Amodei Means by “Pacing the Frontier”
Dario Amodei proposes slowing frontier AI development enough for safety work, independent oversight and international coordination to keep pace with…
-
Read →OpenAI Says AI Scaling Needs Stronger Safety Bars
OpenAI Chief Scientist Jakub Pachocki argues that rapid AI progress is making alignment and monitoring the central limits on responsible development.
-
Read →OpenAI are slowing down RL training
OpenAI has slowed down, but its not bad news.
