How enabling two settings tripled our scores on the ARC-AGI-3 benchmark Post published:2026-07-30 How two API settings improved GPT-5.6 performance on ARC-AGI-3, boosting scores and efficiency by retaining reasoning and enabling compaction. Read more articles Previous PostAccelerating scientific discovery with ChatGPT for Academic Researchers Next PostGuideSkill: Evolving Executable LLM Agent Skills for Guideline-Grounded Clinical Reasoning