Skip to yearly menu bar Skip to main content


When can we trust untrusted monitoring? An AI control safety case sketch across collusion strategies

Morgan Sinclaire ⋅ Nelson Gardner-Challis ⋅ Georgiy Kozhevnikov ⋅ Jonathan Bostock ⋅ Charlie J Griffin ⋅ Joan Velja ⋅ Alessandro Abate

Abstract

Chat is not available.