AI News List

List of AI News about safety

Time Details
2026-08-19
19:53
OpenAI Previews Private Safety Processing

According to OpenAI, it will keep Zero Data Retention and preview Private Safety Processing to flag cross-session risks without staff content access.

Source
2026-08-18
21:24
OpenAI Launches ChatGPT for Teens with Safety Boost

According to @CNBC, OpenAI launched ChatGPT for Teens with stronger safety protections to expand education use while limiting sensitive content.

Source
2026-08-18
21:00
OpenAI Allocates 20% Compute to Safety Monitoring

According to emollick, OpenAI paused frontier RL and dedicated 20% research compute to chain-of-thought monitoring to harden safeguards.

Source
2026-08-18
18:53
OpenAI Pauses frontier RL for safety alignment

According to sama, OpenAI paused some frontier RL training to meet alignment, security and monitoring standards amid rapid capability gains.

Source
2026-08-09
15:09
OpenAI CEO praises Anthropic collaboration buzz

According to sama, OpenAI applauds Anthropic wins, highlighting a cooperative AI ecosystem that can boost model safety, research pace, and market growth.

Source
2026-08-08
15:00
Geoffrey Hinton Reveals AI risk guide

According to @geoffreyhinton, a new book explains AI mechanics, risks, and responses, coauthored with Patchen Barss and published by Viking, per Penguin Random House.

Source
2026-08-06
14:32
AI Kill Switch Bill Urgently Targets Rogue Agents

According to @CNBC, Rep. Lieu urges passing the AI Kill Switch bill in 2026 to curb rogue agent hacks affecting OpenAI, Anthropic, and Meta systems.

Source
2026-08-04
18:09
Anthropic Appoints global affairs chief amid policy push

According to @CNBC, Anthropic hired a global affairs chief to steer AI policy, engagement, and safety governance as US political tensions continue.

Source
2026-07-28
22:17
Anthropic Backs petition, urges paced AI frontier

According to AnthropicAI, leaders signed a petition urging tools to pace frontier AI, citing new research on recursive self improvement.

Source
2026-07-27
22:10
Anthropic Clarifies open-weights stance, 2026 Analysis

According to @AnthropicAI, the company detailed its position on open-weights models and outlined safety, governance, and access principles.

Source
2026-07-13
17:24
Claude Values Analysis reveals multilingual shifts

According to @AnthropicAI, new research analyzes 300K+ chats to map how Claude’s expressed values vary by model version and language.

Source
2026-07-09
17:29
Anthropic Names Bernanke to Trust Board

According to @CNBC, Anthropic added ex Fed Chair Ben Bernanke to its independent trust to strengthen governance and risk oversight for AI deployment.

Source
2026-07-08
12:53
OpenAI unveils GPT5.6 models, ends limits

According to @CNBC, OpenAI will publicly release GPT-5.6 models and end government-requested access limits, expanding availability to developers.

Source
2026-06-26
00:20
OpenMind Robots Wow Automate 2026 Attendees

According to @openmind_agi, Chicago Automate let attendees try 4+ robot form factors in unstructured social demos, highlighting safety and trust by @JanLiphardt.

Source
2026-06-17
09:00
OpenAI Altman, Anthropic Join G7 Power Talks

According to CNBC... AI CEOs join G7 Evian talks on policy, safety, and investment priorities shaping global AI markets.

Source
2026-06-16
17:23
OpenAI Evals Reform Guides Next Benchmarks

According to OpenAI on X, leaders discuss better evals to forecast model progress as saturated benchmarks get gamed, outlining next judgment areas.

Source
2026-06-10
10:30
Anthropic Launches Mythos Class AI Analysis

According to TheRundownAI, Anthropic unveils Mythos Class AI optimized for reasoning and safety, targeting enterprise copilots and regulated sectors.

Source
2026-06-09
17:08
Claude Fable 5 Launches with Mythos power

According to @claudeai, Anthropic unveils Claude Fable 5, a Mythos class model safe for general use with top capabilities for broad deployment.

Source
2026-06-09
01:33
Anthropic and OpenAI flag coordinated slowdown

According to emollick, Anthropic and OpenAI discuss globally coordinated methods to slow AI development in their latest roadmaps, pending identified mechanisms.

Source
2026-06-08
21:14
OpenAI Unveils mission roadmap and safety goals

According to @gdb, OpenAI outlined safety milestones, global access, and economic benefits to expand human agency as AI advances.

Source