semanticscholar8.0 / 10
D-REX
D. Rex, C. Cutler, Gregory T. Lemmel, E. Rahmani, David W. Clark, D. Helper, G. Lehman, D. Mark
Abstract
A benchmark for detecting deceptive reasoning in large language models
Research area
deception
Published
1997
Source
semanticscholar
Org
Gray Swan AI
Sign in to read and join the discussion.