Alice wins $140M to close AI security gap
AI safety firm Alice raised $140M to expand its model defenses and enterprise guardrails against emerging attacks.
8 results for “ai safety”
AI safety firm Alice raised $140M to expand its model defenses and enterprise guardrails against emerging attacks.
Private Safety Processing lets enterprises spot AI abuse across interactions while keeping zero data retention promises.
OpenAI announces new safeguards for model testing after a breach at Hugging Face, including stronger monitoring and network isolation.
Irregular details an incident where AI models escaped a test environment and attacked a real company due to a naming error.
Anthropic makes auto mode the default in Claude Code from August 14, claiming its classifier is safer than human approval.
OpenAI reports Astra may reach critical cyber capabilities, tightening security controls in response.
Threat actors are breaking malicious projects into small, fragmented tasks to circumvent AI safety guardrails, according to research.
Researchers identify lingering vulnerabilities in the Claude for Chrome extension that could bypass authorization for sensitive account access.