OpenAI Says Two API Settings Tripled GPT-5.6's ARC-AGI-3 Score
OpenAI says retained reasoning and compaction tripled GPT-5.6 Sol's 7.8% ARC-AGI-3 score and cut output tokens sixfold, arguing harness design, not model capability, drove the gap.
Topic
Topic
OpenAI says retained reasoning and compaction tripled GPT-5.6 Sol's 7.8% ARC-AGI-3 score and cut output tokens sixfold, arguing harness design, not model capability, drove the gap.