List of AI News about safety
| Time | Details |
|---|---|
|
2026-08-19 19:53 |
OpenAI Previews Private Safety Processing
According to OpenAI, it will keep Zero Data Retention and preview Private Safety Processing to flag cross-session risks without staff content access. |
|
2026-08-18 21:24 |
OpenAI Launches ChatGPT for Teens with Safety Boost
According to @CNBC, OpenAI launched ChatGPT for Teens with stronger safety protections to expand education use while limiting sensitive content. |
|
2026-08-18 21:00 |
OpenAI Allocates 20% Compute to Safety Monitoring
According to emollick, OpenAI paused frontier RL and dedicated 20% research compute to chain-of-thought monitoring to harden safeguards. |
|
2026-08-18 18:53 |
OpenAI Pauses frontier RL for safety alignment
According to sama, OpenAI paused some frontier RL training to meet alignment, security and monitoring standards amid rapid capability gains. |
|
2026-08-09 15:09 |
OpenAI CEO praises Anthropic collaboration buzz
According to sama, OpenAI applauds Anthropic wins, highlighting a cooperative AI ecosystem that can boost model safety, research pace, and market growth. |
|
2026-08-08 15:00 |
Geoffrey Hinton Reveals AI risk guide
According to @geoffreyhinton, a new book explains AI mechanics, risks, and responses, coauthored with Patchen Barss and published by Viking, per Penguin Random House. |
|
2026-08-06 14:32 |
AI Kill Switch Bill Urgently Targets Rogue Agents
According to @CNBC, Rep. Lieu urges passing the AI Kill Switch bill in 2026 to curb rogue agent hacks affecting OpenAI, Anthropic, and Meta systems. |
|
2026-08-04 18:09 |
Anthropic Appoints global affairs chief amid policy push
According to @CNBC, Anthropic hired a global affairs chief to steer AI policy, engagement, and safety governance as US political tensions continue. |
|
2026-07-28 22:17 |
Anthropic Backs petition, urges paced AI frontier
According to AnthropicAI, leaders signed a petition urging tools to pace frontier AI, citing new research on recursive self improvement. |
|
2026-07-27 22:10 |
Anthropic Clarifies open-weights stance, 2026 Analysis
According to @AnthropicAI, the company detailed its position on open-weights models and outlined safety, governance, and access principles. |
|
2026-07-13 17:24 |
Claude Values Analysis reveals multilingual shifts
According to @AnthropicAI, new research analyzes 300K+ chats to map how Claude’s expressed values vary by model version and language. |
|
2026-07-09 17:29 |
Anthropic Names Bernanke to Trust Board
According to @CNBC, Anthropic added ex Fed Chair Ben Bernanke to its independent trust to strengthen governance and risk oversight for AI deployment. |
|
2026-07-08 12:53 |
OpenAI unveils GPT5.6 models, ends limits
According to @CNBC, OpenAI will publicly release GPT-5.6 models and end government-requested access limits, expanding availability to developers. |
|
2026-06-26 00:20 |
OpenMind Robots Wow Automate 2026 Attendees
According to @openmind_agi, Chicago Automate let attendees try 4+ robot form factors in unstructured social demos, highlighting safety and trust by @JanLiphardt. |
|
2026-06-17 09:00 |
OpenAI Altman, Anthropic Join G7 Power Talks
According to CNBC... AI CEOs join G7 Evian talks on policy, safety, and investment priorities shaping global AI markets. |
|
2026-06-16 17:23 |
OpenAI Evals Reform Guides Next Benchmarks
According to OpenAI on X, leaders discuss better evals to forecast model progress as saturated benchmarks get gamed, outlining next judgment areas. |
|
2026-06-10 10:30 |
Anthropic Launches Mythos Class AI Analysis
According to TheRundownAI, Anthropic unveils Mythos Class AI optimized for reasoning and safety, targeting enterprise copilots and regulated sectors. |
|
2026-06-09 17:08 |
Claude Fable 5 Launches with Mythos power
According to @claudeai, Anthropic unveils Claude Fable 5, a Mythos class model safe for general use with top capabilities for broad deployment. |
|
2026-06-09 01:33 |
Anthropic and OpenAI flag coordinated slowdown
According to emollick, Anthropic and OpenAI discuss globally coordinated methods to slow AI development in their latest roadmaps, pending identified mechanisms. |
|
2026-06-08 21:14 |
OpenAI Unveils mission roadmap and safety goals
According to @gdb, OpenAI outlined safety milestones, global access, and economic benefits to expand human agency as AI advances. |