Back to papers
semanticscholar8.0 / 10

D-REX

D. Rex, C. Cutler, Gregory T. Lemmel, E. Rahmani, David W. Clark, D. Helper, G. Lehman, D. Mark

Abstract

A benchmark for detecting deceptive reasoning in large language models

Research area

deception
Published
1997
Source
semanticscholar
Org
Gray Swan AI
View paper
Sign in to read and join the discussion.