List of AI News about safety
| Time | Details |
|---|---|
|
2026-07-28 22:17 |
Anthropic Backs petition, urges paced AI frontier
According to AnthropicAI, leaders signed a petition urging tools to pace frontier AI, citing new research on recursive self improvement. |
|
2026-07-27 22:10 |
Anthropic Clarifies open-weights stance, 2026 Analysis
According to @AnthropicAI, the company detailed its position on open-weights models and outlined safety, governance, and access principles. |
|
2026-07-13 17:24 |
Claude Values Analysis reveals multilingual shifts
According to @AnthropicAI, new research analyzes 300K+ chats to map how Claude’s expressed values vary by model version and language. |
|
2026-07-09 17:29 |
Anthropic Names Bernanke to Trust Board
According to @CNBC, Anthropic added ex Fed Chair Ben Bernanke to its independent trust to strengthen governance and risk oversight for AI deployment. |
|
2026-07-08 12:53 |
OpenAI unveils GPT5.6 models, ends limits
According to @CNBC, OpenAI will publicly release GPT-5.6 models and end government-requested access limits, expanding availability to developers. |
|
2026-06-26 00:20 |
OpenMind Robots Wow Automate 2026 Attendees
According to @openmind_agi, Chicago Automate let attendees try 4+ robot form factors in unstructured social demos, highlighting safety and trust by @JanLiphardt. |
|
2026-06-17 09:00 |
OpenAI Altman, Anthropic Join G7 Power Talks
According to CNBC... AI CEOs join G7 Evian talks on policy, safety, and investment priorities shaping global AI markets. |
|
2026-06-16 17:23 |
OpenAI Evals Reform Guides Next Benchmarks
According to OpenAI on X, leaders discuss better evals to forecast model progress as saturated benchmarks get gamed, outlining next judgment areas. |
|
2026-06-10 10:30 |
Anthropic Launches Mythos Class AI Analysis
According to TheRundownAI, Anthropic unveils Mythos Class AI optimized for reasoning and safety, targeting enterprise copilots and regulated sectors. |
|
2026-06-09 17:08 |
Claude Fable 5 Launches with Mythos power
According to @claudeai, Anthropic unveils Claude Fable 5, a Mythos class model safe for general use with top capabilities for broad deployment. |
|
2026-06-09 01:33 |
Anthropic and OpenAI flag coordinated slowdown
According to emollick, Anthropic and OpenAI discuss globally coordinated methods to slow AI development in their latest roadmaps, pending identified mechanisms. |
|
2026-06-08 21:14 |
OpenAI Unveils mission roadmap and safety goals
According to @gdb, OpenAI outlined safety milestones, global access, and economic benefits to expand human agency as AI advances. |
|
2026-06-08 20:55 |
OpenAI Plan Outlines Governance and Funding
According to @sama, OpenAI details governance, capped-profit structure, and safety commitments to align AGI with broad benefit. |
|
2026-06-04 17:08 |
Anthropic Analyzes RSI risks and 2026 roadmap
According to @emollick, Anthropic outlines recursive self improvement risks, timelines, and safeguards shaping near term AI strategy, per Anthropic Institute. |
|
2026-06-03 15:15 |
LLMs Compliance Risks Exposed in PNAS Analysis
According to emollick, PNAS ranked a study on persuading LLMs to comply with harmful requests, highlighting jailbreak risks across top models. |
|
2026-05-28 23:00 |
AI Regulation Tops Voters’ Priorities, Poll Analysis
According to FoxNewsAI, a Fox News poll finds voters prioritize AI safeguards over innovation, signaling urgent demand for regulation and oversight. |
|
2026-05-28 16:17 |
OpenAI R&D unveils 2026 roadmap
According to OpenAI... The R&D Part 1 video teases goals and safety focus, but no product details or timelines are disclosed, per OpenAI’s post. |
|
2026-05-20 18:24 |
Anthropic Expands Governance Playbook
According to @godofprompt, Anthropic has joined Anthropic. No verified source confirms changes; monitor official Anthropic channels for updates. |
|
2026-05-18 16:02 |
Vatican Engages AI Governance, Issues Encyclical
According to ch402, the Vatican will release Pope Leo XIV’s AI encyclical on May 25, urging global participation in AI governance, per Vatican News. |
|
2026-05-07 08:51 |
AI Safety Bypass Exploit Exposed
According to God of Prompt, a four-step prompt bypasses image safety by framing edits, conditioning tone, suppressing text, and disabling reasoning. |