ai-security
Everything on Ground Truth tagged “ai-security” — 2 items.
Safety Guardrails Blocked a Security Team's Own Incident Analysis News
Hugging Face disclosed that commercial AI safety filters blocked its analysis of real attack code during an incident, so it ran the forensics on a self-hosted open-weight model instead.
Claude Code Briefly Made Silence Mean Yes, Then Reversed It News
Anthropic shipped a Claude Code default that let its AI agent auto-continue after 60 seconds when a user did not answer a clarifying question, then rolled it back two days later after developers called it a broken trust boundary.