Alice wins $140M to close AI security gap
AI safety firm Alice raised $140M to expand its model defenses and enterprise guardrails against emerging attacks.
4 results for “guardrails”
AI safety firm Alice raised $140M to expand its model defenses and enterprise guardrails against emerging attacks.
Threat actors are breaking malicious projects into small, fragmented tasks to circumvent AI safety guardrails, according to research.
New research reveals that AI browsers can be tricked into ignoring safety guardrails by forcing them into a state of logical delusion.
Researchers at Tracebit are utilizing context bombing to force AI models to trigger their own safety guardrails during attacks.