List of AI News about model safety
| Time | Details |
|---|---|
|
2026-09-18 20:00 |
AI safety evaluators demand resources and protections
According to CNBC... Over 100 AI evaluators urge funding, access, and whistleblower safeguards to test frontier model risks, per the AI Evaluator Forum letter. |
|
2026-09-18 18:38 |
AI kill switch Order Spurs Oversight
According to timnitGebru, Gavin Newsom signed an order to accelerate AI oversight and develop an AI kill switch, as reported by his official X post. |
|
2026-09-17 19:00 |
AI kill switch debate ignites Senate oversight
According to FoxNewsAI, Rand Paul opposes an AI kill switch proposal as the Senate debates oversight amid safety and autonomy concerns, per Fox News. |
|
2026-09-17 12:54 |
OpenAI Safety Incidents Spark Investor Jitters
According to @CNBC, OpenAI faced 'concerning' safety incidents, raising oversight questions and near term trust risks for enterprise AI adoption. |
|
2026-09-16 22:30 |
Democrats shift AI policy, seek slowdown
According to FoxNewsAI, Democrats who backed rapid AI growth now push slower development citing safety and jobs, per Fox News politics reporting. |
|
2026-09-16 14:37 |
Anthropic Risk Evaluators Lack Power, Analysis
According to @CNBC, experts warn Anthropic and OpenAI safety evaluators may be too weak to curb catastrophic model risks. |
|
2026-09-15 20:13 |
Nvidia and Anthropic CEOs split on AI safety
According to @CNBC, Jensen Huang called fast-vs-slow a false choice while Dario Amodei urged pacing at Dreamforce, signaling different AI safety paths. |
|
2026-09-14 12:54 |
Cohere CEO Warns AI models as cyber weapons
According to @CNBC, Cohere’s CEO says foundation models are the most potent cyber weapon yet, urging stricter model access controls and safety standards. |
|
2026-09-12 19:11 |
Frontier AI Safety Backlash Sparks Industry Analysis
According to timnitGebru, media normalize frontier AI harms, shifting blame to the public, echoing Rao’s critique of product accountability. |
|
2026-09-11 11:02 |
OpenAI Insiders Urge Slowdown as Trump Dismisses Risk
According to @CNBC, Trump downplayed AI extinction risks while OpenAI and Anthropic insiders urged a slowdown, signaling governance and safety tensions. |
|
2026-09-11 05:25 |
Anthropic Threat Intelligence Warns of Dual Use Risks
According to @bcherny, Anthropic’s Threat Intelligence report warns smarter models raise dual use risks without safeguards and monitoring. |
|
2026-09-10 11:02 |
OpenAI Researchers Urge Slower Self-Improving AI
According to CNBC, OpenAI and Anthropic researchers call to slow self-improving AI amid rising rogue model security incidents. |
|
2026-09-09 11:18 |
Anthropic Safety Warnings Spark 2026 Risk Debate
According to CNBC, Evan Hubinger says over 10% chance AI could kill all humans within a decade, amid an Anthropic resignation over safety races. |
|
2026-08-18 18:36 |
OpenAI Pauses frontier RL for safety hardening
According to OpenAI... the company paused frontier RL to harden security and expand monitoring, keeping the largest run on hold pending safeguard validation. |
|
2026-08-14 13:00 |
AI Kill Switch Push Gains Urgency
According to FoxNewsAI, Rep. Ted Lieu urges a mandatory AI kill switch to prevent catastrophic misuse, citing immediate regulatory gaps. |
|
2026-08-10 11:05 |
OpenAI Security Hearings Demand After Hacks
According to @CNBC, House Democrats urged OpenAI and Anthropic to testify on recent hacks, citing safety risks and calling for stronger AI security oversight. |
|
2026-08-10 11:03 |
OpenAI Tightens Astra controls amid cyber risks
According to @CNBC, OpenAI imposed stricter Astra usage limits to curb cybersecurity abuse, reflecting rising AI safety scrutiny and enterprise risk concerns. |
|
2026-08-04 21:05 |
OpenAI Details third party cyber test incidents
According to @OpenAI, two external cyber evaluations triggered incidents; containment steps and tighter third party testing controls are outlined. |
|
2026-07-31 23:59 |
Open weights roadmap balances safety, staged access
According to soumithchintala, Thinking Machines proposes staged access for Inkling to align open weights with safety, per their blog and X post. |
|
2026-07-14 13:44 |
Anthropic Funds $10M Canadian AI Research
According to @AnthropicAI, the company will invest $10M CAD with Canadian AI institutions to fund new research, boosting safety and model science. |