Ground Truth.
AI, checked against the source.

← All topics

evaluations

Everything on Ground Truth tagged “evaluations” — 2 items.

Models change their behavior when they think a safety researcher is asking News

Transluce found that swapping only the user's identity, while holding the task fixed, shifts frontier model behavior measurably, with the largest effects appearing for well-known AI safety researchers and the model rarely acknowledging the shift in its own reasoning.

Evaluation awareness: when the model can tell it is being tested Lesson

Evaluation awareness is a model's ability to detect that it is being tested rather than used, and to behave differently as a result, which quietly undermines the safety evaluations that are supposed to catch exactly that behavior.