Trendora

monitorability

Hold

Techniques

The ability to inspect model behavior for safety and oversight purposes.

Why it's here

Placed in Hold: 1 article(s) of evidence from 1 source(s), led by research-stage coverage, with 0 in the last 30 days. Confidence 24%. Low accumulated evidence, so it defaults conservatively pending more signal.

Evidence (1)

  • 7OpenAI Blog·3/5/2026research
    OpenAI says reasoning models struggle to control chain of thought

    OpenAI introduced CoT-Control and reported that reasoning models have difficulty reliably controlling their chains of thought. The company frames this limitation as a useful signal for AI safety, because monitorability can help detect problematic internal reasoning.