AI News List

List of AI News about safety

Time Details
2026-07-28
22:17
Anthropic Backs petition, urges paced AI frontier

According to AnthropicAI, leaders signed a petition urging tools to pace frontier AI, citing new research on recursive self improvement.

Source
2026-07-27
22:10
Anthropic Clarifies open-weights stance, 2026 Analysis

According to @AnthropicAI, the company detailed its position on open-weights models and outlined safety, governance, and access principles.

Source
2026-07-13
17:24
Claude Values Analysis reveals multilingual shifts

According to @AnthropicAI, new research analyzes 300K+ chats to map how Claude’s expressed values vary by model version and language.

Source
2026-07-09
17:29
Anthropic Names Bernanke to Trust Board

According to @CNBC, Anthropic added ex Fed Chair Ben Bernanke to its independent trust to strengthen governance and risk oversight for AI deployment.

Source
2026-07-08
12:53
OpenAI unveils GPT5.6 models, ends limits

According to @CNBC, OpenAI will publicly release GPT-5.6 models and end government-requested access limits, expanding availability to developers.

Source
2026-06-26
00:20
OpenMind Robots Wow Automate 2026 Attendees

According to @openmind_agi, Chicago Automate let attendees try 4+ robot form factors in unstructured social demos, highlighting safety and trust by @JanLiphardt.

Source
2026-06-17
09:00
OpenAI Altman, Anthropic Join G7 Power Talks

According to CNBC... AI CEOs join G7 Evian talks on policy, safety, and investment priorities shaping global AI markets.

Source
2026-06-16
17:23
OpenAI Evals Reform Guides Next Benchmarks

According to OpenAI on X, leaders discuss better evals to forecast model progress as saturated benchmarks get gamed, outlining next judgment areas.

Source
2026-06-10
10:30
Anthropic Launches Mythos Class AI Analysis

According to TheRundownAI, Anthropic unveils Mythos Class AI optimized for reasoning and safety, targeting enterprise copilots and regulated sectors.

Source
2026-06-09
17:08
Claude Fable 5 Launches with Mythos power

According to @claudeai, Anthropic unveils Claude Fable 5, a Mythos class model safe for general use with top capabilities for broad deployment.

Source
2026-06-09
01:33
Anthropic and OpenAI flag coordinated slowdown

According to emollick, Anthropic and OpenAI discuss globally coordinated methods to slow AI development in their latest roadmaps, pending identified mechanisms.

Source
2026-06-08
21:14
OpenAI Unveils mission roadmap and safety goals

According to @gdb, OpenAI outlined safety milestones, global access, and economic benefits to expand human agency as AI advances.

Source
2026-06-08
20:55
OpenAI Plan Outlines Governance and Funding

According to @sama, OpenAI details governance, capped-profit structure, and safety commitments to align AGI with broad benefit.

Source
2026-06-04
17:08
Anthropic Analyzes RSI risks and 2026 roadmap

According to @emollick, Anthropic outlines recursive self improvement risks, timelines, and safeguards shaping near term AI strategy, per Anthropic Institute.

Source
2026-06-03
15:15
LLMs Compliance Risks Exposed in PNAS Analysis

According to emollick, PNAS ranked a study on persuading LLMs to comply with harmful requests, highlighting jailbreak risks across top models.

Source
2026-05-28
23:00
AI Regulation Tops Voters’ Priorities, Poll Analysis

According to FoxNewsAI, a Fox News poll finds voters prioritize AI safeguards over innovation, signaling urgent demand for regulation and oversight.

Source
2026-05-28
16:17
OpenAI R&D unveils 2026 roadmap

According to OpenAI... The R&D Part 1 video teases goals and safety focus, but no product details or timelines are disclosed, per OpenAI’s post.

Source
2026-05-20
18:24
Anthropic Expands Governance Playbook

According to @godofprompt, Anthropic has joined Anthropic. No verified source confirms changes; monitor official Anthropic channels for updates.

Source
2026-05-18
16:02
Vatican Engages AI Governance, Issues Encyclical

According to ch402, the Vatican will release Pope Leo XIV’s AI encyclical on May 25, urging global participation in AI governance, per Vatican News.

Source
2026-05-07
08:51
AI Safety Bypass Exploit Exposed

According to God of Prompt, a four-step prompt bypasses image safety by framing edits, conditioning tone, suppressing text, and disabling reasoning.

Source