OpenAI's GPT-5.6 Sol scored 7.8% on ARC-AGI-3. After turning on two settings — retained reasoning and compaction — the score tripled to 38.3% with 6x fewer tokens. Same model, different harness. Here's what that means for anyone running AI agents in production.