Ant's Sante model tops GPT-5.6 Sol on diagnosis with 5.1B active parameters
OpenAI chief scientist Jakub Pachocki says chain-of-thought monitoring — watching models write out their reasoning to catch misalignment — is becoming unreliable, and no lab is ready to scale safely.