AI Safety
Anthropic Strengthens AI Safeguards for Claude
Anthropic enhances its AI model Claude's safety and reliability with robust safeguards, ensuring beneficial outcomes while preventing misuse and harmful impacts.
Character.AI Implements New Safety Measures for Teen Users
Character.AI announces significant changes to enhance the safety of its platform for users under 18, including removing open-ended chat and introducing age assurance tools.
OpenAI Enhances GPT-5 for Sensitive Conversations with New Safety Measures
OpenAI has released an addendum to the GPT-5 system card, showcasing improvements in handling sensitive conversations with enhanced safety benchmarks.
NVIDIA Introduces Safety Measures for Agentic AI Systems
NVIDIA has launched a comprehensive safety recipe to enhance the security and compliance of agentic AI systems, addressing risks such as prompt injection and data leakage.
NVIDIA NeMo Guardrails Enhance LLM Streaming for Safer AI Interactions
NVIDIA introduces NeMo Guardrails to enhance large language model (LLM) streaming, improving latency and safety for generative AI applications through real-time, token-by-token output validation.
Ensuring AI Reliability: NVIDIA NeMo Guardrails Integrates Cleanlab's Trustworthy Language Model
NVIDIA's NeMo Guardrails, in collaboration with Cleanlab's Trustworthy Language Model, aims to enhance AI reliability by preventing hallucinations in AI-generated responses.
NVIDIA Introduces Halos: A Comprehensive Safety System for Autonomous Vehicles
NVIDIA unveils Halos, a full-stack safety system for autonomous vehicles, integrating AI and automotive technology for enhanced safety and compliance.
Enhancing AI Safety in Customer Service with NVIDIA NeMo Guardrails
Explore how NVIDIA NeMo Guardrails enhance AI safety in customer service, integrating advanced security measures to ensure effective and secure interactions.
OpenAI Releases Comprehensive GPT-4o System Card Detailing Safety Measures
OpenAI's report on GPT-4o highlights extensive safety evaluations, red teaming, and risk mitigations prior to release.
Anthropic Expands AI Model Safety Bug Bounty Program
Anthropic broadens its AI model safety bug bounty program to address universal jailbreak vulnerabilities, offering rewards up to $15,000.