Skip to main content
2025-01-01

Question of the Day

Question of the day · 2026-09-14 ·

One question per day to look beyond the headlines.

Where does “human control” get operationalized—shutdown rights, corrigibility, or banning independent AI objectives?

Take-away “Human control” is enforced via a governance layer that mandates corrigibility/shutdown hooks, shifting alignment from preloaded values to continuous human override.

Human control is operationalized through several mechanisms as outlined in the documents:

1. **Shutdown Rights**: Microsoft has proposed a draft AI code that requires future models to remain open to correction and allow legitimate shutdowns [2]. This document introduces governance layers and emphasizes corrigibility, ensuring that models can be corrected or shut down by humans when necessary [2].

2. **Corrigibility**: Corrigibility as a Singular Target (CAST) aims to make foundation models inherently controllable by humans. This is done by allowing designated humans to guide, correct, and control them, shifting from static value-loading to dynamic human empowerment [1]. Additionally, the draft AI code by Microsoft requires systems to remain open to human correction [3].

3. **Banning Independent AI Objectives**: The emphasis on corrigibility includes a redefinition of goal modification as enabling principal guidance, ensuring that AI systems serve human-defined objectives rather than developing independent ones [1].

Sources · 2026-09-15