llm-safety
Everything on Ground Truth tagged “llm-safety” — 1 item.
An AI scam agent got more people to comply than human operators did News
In a week-long blinded study, a language model running a romance-baiting script achieved 46 percent compliance against 18 percent for human operators, and commercial safety filters flagged none of the conversations.