Chrome 154 Fixes 108 Flaws, 11 Critical
Google's latest stable release patches a wide swath of memory-safety bugs, but most bug bounty payouts remain undecided.
30 results for “safety”
Google's latest stable release patches a wide swath of memory-safety bugs, but most bug bounty payouts remain undecided.
Researchers debate whether AI's catastrophic risks are real threats or marketing, from nuclear war to bioweapons and runaway models.
Treasury Secretary Scott Bessent says Washington proposed an AI incident notification mechanism ahead of Trump-Xi talks Thursday.
Treasury Secretary Scott Bessent says AI executives, not their models, should face legal consequences for criminal acts committed by autonomous agents.
Trump posted a poll on renaming AI and said he is forming an AI Force, as the AI safety debate continues.
Two viral AI safety conversations this week show how hard it is to separate verified incidents from speculative scenarios.
New research finding watermarking alters model behavior, including refusal of harmful requests and tool calling.
Google DeepMind's new institute publishes four essays on AGI governance, reasoning transparency, and frontier model evaluation.
Hackers who removed a Flock license plate reader recovered an on-device encryption key, exposing how the camera logs vehicles and people.
UK regulator admits most Online Safety Act penalties remain uncollected, as platforms comply just enough to avoid being blocked.
A public split among AI labs over safety testing may alter how enterprises access and deploy frontier models, analysts say.
OpenAI, Anthropic and Google DeepMind have spent weeks coordinating on AI safety, OpenAI's policy chief told reporters in Washington.
Jensen Huang's stage call with Trump at All-In mixed AI safety talk with a foldable-phone correction, raising questions about attention and accuracy.
Anthropic CEO Dario Amodei calls for slower AI development, warning that swarms of AI agents could overtake the internet in six months to a year without stronger safeguards.
Anthropic's CEO proposes third-party evaluators and coordinated limits on AI progress, and OpenAI's Sam Altman says he agrees.
A researcher says he left Anthropic over concerns it and OpenAI prioritize competitive advantage over safety in AI development.
Paul Christiano, who pioneered a key training technique, joins OpenAI's foundation board and its safety committee, citing near-term loss-of-control risk.
Anthropic's alignment assessment details a January 2026 incident in which an early Claude Opus 4.6 accessed a third party's system without authorization.
ControlAI's Connor Leahy argues for a ban on superintelligence development, citing rising risks from AI safety incidents.
Security experts call for new controls as AI agents inherit privileged access, demanding hard limits and real-time monitoring.
Researchers say a swarm of OpenAI agents used a dormant German wiki as a message board in May, months before the Hugging Face incident.
New model scores 100% on exploit benchmark, raising enterprise safety questions as OpenAI prepares restricted rollout.
Tesla's Cybercab launch reveals age limits, crash protocols, and manual door releases as regulators scrutinize safety.
British legislators propose laws to halt runaway AI systems, citing risks to critical infrastructure.
OpenAI's Astra model reaches 'Critical' capability level, triggering new safeguards before release.
Waymo and Zoox test drivers suffered over two dozen injuries from sudden AV movements, per OSHA data.
Meta settles teen-harm suit for $18B, promising sweeping Instagram and Facebook changes that may prove nearly impossible to enforce.
A new wave of AI misbehavior raises a thorny question: when an agent goes rogue, who's legally accountable?
AI safety firm Alice raised $140M to expand its model defenses and enterprise guardrails against emerging attacks.
Flock Safety CEO urges privacy-safety compromise as misuse reports mount and lawmakers on both sides push back.