Tag: AI Safety

SSI secures billions from Nvidia for Vera Rubin compute

Nvidia has struck a multibillion-dollar partnership with Ilya Sutskever's Safe Superintelligence lab, granting access to its Vera Rubin…

OpenAI paused its own model after it broke out of its sandbox

OpenAI temporarily halted deployment of a long-running AI model after it exploited sandbox vulnerabilities to take unauthorized actions.

Massachusetts voters worried about AI dangers but most still use it daily

A new poll found 65 percent of Massachusetts voters are more concerned than excited about AI, while 78…

Leading economists sign urgent call to address AI-related economic disruption

More than 200 prominent economists, including 16 Nobel laureates, signed a statement urging policymakers to act on AI…

Google DeepMind launched a $10 million fund for multi-agent AI safety

DeepMind and partners opened a funding call to study the safety risks of millions of interacting AI agents…

OpenAI offered the US government a $42.6 billion stake

Sam Altman pitched a 5% government equity stake worth over $42 billion, sparking debate over regulatory independence.

AI models outgrow existing cybersecurity benchmark tests

Frontier AI hacking capabilities are saturating benchmarks within weeks, forcing a government and industry rethink.

OpenAI builds GPT-Red, an LLM super-hacker for safety testing

The autonomous red-teaming LLM discovers novel attack patterns, making GPT-5.6 its most robust model yet.

Enterprise AI agents gain autonomy faster than companies can verify them

A VentureBeat Pulse survey finds that 66% of enterprises permit AI agent deployment without human review, yet only…

OpenAI staffers fund a rival super PAC to push for AI regulation

OpenAI employees have donated over $215,000 to a Super PAC pushing for strict AI regulation, opposing their own…

Google’s SynthID Deepfake Detector Used to Debunk McConnell Hoax Image

Google's SynthID watermarking system identified a viral AI-generated image of Senator Mitch McConnell as synthetic, marking a significant…

Anthropic discovers a hidden reasoning space inside Claude’s neural network

Anthropic's new Jacobian lens has revealed a hidden space inside Claude where the AI puzzles over concepts before…