Search Results for "ai safety"
Exploring AI Stability: Navigating Non-Power-Seeking Behavior Across Environments
The research explores AI's stability in non-power-seeking behaviors, revealing that certain policies maintain non-resistance to shutdown across similar environments, providing insights into mitigating risks associated with power-seeking AI.
Exploring AGI Hallucination: A Comprehensive Survey of Challenges and Mitigation Strategies
A new survey delves into the phenomenon of AGI hallucination, categorizing its types, causes, and current mitigation approaches while discussing future research directions.
NIST's Call for Public Input on AI Safety in Response to Biden's Executive Order
NIST is seeking public input to create AI safety guidelines following President Biden's Executive Order, aiming to ensure a secure AI environment, mitigate risks, and foster innovation.
California Spearheads AI Ethics and Safety with Senate Bills 892 and 893
California takes a pioneering role in AI regulation with Senate Bills 892 and 893, aiming to ensure AI safety, ethics, and public benefits.
US NIST Initiates AI Safety Consortium to Promote Trustworthy AI Development
The US National Institute of Standards and Technology (NIST) has launched the Artificial Intelligence Safety Institute Consortium to promote safe AI development and responsible use, inviting organizations to collaborate on identifying proven safety techniques by December 4, 2023.
British Standards Institution Pioneers International AI Safety Guidelines for Sustainable Future
BSI's release of the first international AI safety guideline, BS ISO/IEC 42001, marks a significant step in standardizing the safe and ethical use of AI, reflecting global demand for robust AI governance.
Amazon Invests $4 Billion in AI Startup Anthropic for Advanced Foundation Models
Amazon and AI startup Anthropic have entered into a $4 billion investment agreement to develop advanced foundation models. The collaboration will provide Anthropic with AWS resources and allow Amazon to build on Anthropic's AI models. Both companies are committed to AI safety and responsible scaling.
OpenAI Introduces the "Preparedness Framework" for AI Safety and Policy Integration
OpenAI has introduced the "Preparedness Framework," giving its board veto over CEO decisions and introducing risk scorecards for AI risk management, demonstrating its commitment to responsible AI development.
Google DeepMind: Subtle Adversarial Image Manipulation Influences Both AI Model and Human Perception
Recent DeepMind research reveals that subtle adversarial image manipulations, originally designed to deceive AI models, also subtly influence human perception. This discovery underscores similarities and distinctions in human and machine vision, emphasizing the need for further research in AI safety and security.
Cisco Partners with SingularityNET to Decentralize Artificial Intelligence with Blockchain
Tech conglomerate Cisco and decentralized artificial intelligence (AI) firm SingularityNET have reached a partnership to develop a decentralized Artificial General Intelligence (AGI) project. The ambitious project aims to create more advanced AI technologies that will soon be able to exceed human abilities to learn and perform new tasks.
IBM and Verizon Business Collaborate on 5G and AI Edge Computing Innovation
Tech giant IBM is collaborating with Verizon Business to help usher in the future of Industry 4.0 through their respective expertise on edge computing innovation and 5G technology.
Russia Plans to Build AI Software to Remove the Anonymity of Users in Crypto Transactions
The Russian agency’s plan involves using artificial intelligence (AI) for its software, as the agency sees an “urgent need to create effective opportunities for state control over the circulation of virtual assets.”