AI-Wiki

Tag: gpt-4o

1 item with this tag.

  • Aug 30, 2026

    Monitoring Reasoning Models for Misbehavior and the Risks of Promoting Obfuscation

    • openai
    • reward-hacking
    • chain-of-thought
    • cot-monitoring
    • obfuscation
    • monitorability-tax
    • o3-mini
    • gpt-4o
    • alignment
    • oversight
    • type/source
    • kind/paper

Created with Quartz v0.1.0 © 2026

254 sources · 171 entities · 47 concepts · 7 threads · 5 syntheses

  • GitHub