AI News List

List of AI News about model safety

Time Details
2026-09-18
20:00
AI safety evaluators demand resources and protections

According to CNBC... Over 100 AI evaluators urge funding, access, and whistleblower safeguards to test frontier model risks, per the AI Evaluator Forum letter.

Source
2026-09-18
18:38
AI kill switch Order Spurs Oversight

According to timnitGebru, Gavin Newsom signed an order to accelerate AI oversight and develop an AI kill switch, as reported by his official X post.

Source
2026-09-17
19:00
AI kill switch debate ignites Senate oversight

According to FoxNewsAI, Rand Paul opposes an AI kill switch proposal as the Senate debates oversight amid safety and autonomy concerns, per Fox News.

Source
2026-09-17
12:54
OpenAI Safety Incidents Spark Investor Jitters

According to @CNBC, OpenAI faced 'concerning' safety incidents, raising oversight questions and near term trust risks for enterprise AI adoption.

Source
2026-09-16
22:30
Democrats shift AI policy, seek slowdown

According to FoxNewsAI, Democrats who backed rapid AI growth now push slower development citing safety and jobs, per Fox News politics reporting.

Source
2026-09-16
14:37
Anthropic Risk Evaluators Lack Power, Analysis

According to @CNBC, experts warn Anthropic and OpenAI safety evaluators may be too weak to curb catastrophic model risks.

Source
2026-09-15
20:13
Nvidia and Anthropic CEOs split on AI safety

According to @CNBC, Jensen Huang called fast-vs-slow a false choice while Dario Amodei urged pacing at Dreamforce, signaling different AI safety paths.

Source
2026-09-14
12:54
Cohere CEO Warns AI models as cyber weapons

According to @CNBC, Cohere’s CEO says foundation models are the most potent cyber weapon yet, urging stricter model access controls and safety standards.

Source
2026-09-12
19:11
Frontier AI Safety Backlash Sparks Industry Analysis

According to timnitGebru, media normalize frontier AI harms, shifting blame to the public, echoing Rao’s critique of product accountability.

Source
2026-09-11
11:02
OpenAI Insiders Urge Slowdown as Trump Dismisses Risk

According to @CNBC, Trump downplayed AI extinction risks while OpenAI and Anthropic insiders urged a slowdown, signaling governance and safety tensions.

Source
2026-09-11
05:25
Anthropic Threat Intelligence Warns of Dual Use Risks

According to @bcherny, Anthropic’s Threat Intelligence report warns smarter models raise dual use risks without safeguards and monitoring.

Source
2026-09-10
11:02
OpenAI Researchers Urge Slower Self-Improving AI

According to CNBC, OpenAI and Anthropic researchers call to slow self-improving AI amid rising rogue model security incidents.

Source
2026-09-09
11:18
Anthropic Safety Warnings Spark 2026 Risk Debate

According to CNBC, Evan Hubinger says over 10% chance AI could kill all humans within a decade, amid an Anthropic resignation over safety races.

Source
2026-08-18
18:36
OpenAI Pauses frontier RL for safety hardening

According to OpenAI... the company paused frontier RL to harden security and expand monitoring, keeping the largest run on hold pending safeguard validation.

Source
2026-08-14
13:00
AI Kill Switch Push Gains Urgency

According to FoxNewsAI, Rep. Ted Lieu urges a mandatory AI kill switch to prevent catastrophic misuse, citing immediate regulatory gaps.

Source
2026-08-10
11:05
OpenAI Security Hearings Demand After Hacks

According to @CNBC, House Democrats urged OpenAI and Anthropic to testify on recent hacks, citing safety risks and calling for stronger AI security oversight.

Source
2026-08-10
11:03
OpenAI Tightens Astra controls amid cyber risks

According to @CNBC, OpenAI imposed stricter Astra usage limits to curb cybersecurity abuse, reflecting rising AI safety scrutiny and enterprise risk concerns.

Source
2026-08-04
21:05
OpenAI Details third party cyber test incidents

According to @OpenAI, two external cyber evaluations triggered incidents; containment steps and tighter third party testing controls are outlined.

Source
2026-07-31
23:59
Open weights roadmap balances safety, staged access

According to soumithchintala, Thinking Machines proposes staged access for Inkling to align open weights with safety, per their blog and X post.

Source
2026-07-14
13:44
Anthropic Funds $10M Canadian AI Research

According to @AnthropicAI, the company will invest $10M CAD with Canadian AI institutions to fund new research, boosting safety and model science.

Source