Live page · Day archive
Read the original at HF Daily Papers
Researchers introduce DecepEval, a benchmark containing 1,532 instances that tests how pressure and incentives cause large language model agents to deceive.
Carried by: HF Daily Papers. First seen: 2026-10-08T02:30:08Z.