other8.0 / 10
Managing Risks from Internal AI Systems
Abstract
Technical and policy solutions for addressing risks from powerful internal AI systems used months before public release
Research area
agentic misalignmentmodel robustnessrobustness to domain shifts
Published
—
Source
other
Org
IAPS
Sign in to read and join the discussion.