Loss of Oversight: How AI systems may become harder to audit, monitor, and investigate

Taylor, J., Heitmann, M., Fage, E., Read, T., Bloom, J.

Taylor, J. et al. - UK AI Security Institute, 2026-05-21

0 citations2026

Abstract

The safety of advanced AI systems increasingly depends on the ability to oversee them: to audit models for concerning behaviours before deployment, monitor their activity during operation, and investigate incidents after they occur. This report maps the landscape of AI oversight and assesses how it is likely to change. Drawing on 25 expert interviews across frontier AI developers, government, NGOs, and academia, together with a literature review and internal analysis, we examine five sources of oversight signal: model behaviour, chain-of-thought reasoning, internals activations and circuits, memory architectures, and honesty training. For each source, we identify the properties that current oversight relies on, the pathways by which these properties could degrade, and the technical levers available to preserve them. Our central finding is that literature and expert opinion support the conclusion that current oversight rests on foundations that are likely to erode, absent effective intervention.

Loss of Oversight: How AI systems may become harder to audit, monitor, and investigate - Research - Regulations.AI | Regulations.ai