The Harness Is Doing the Work: How Small Open Models Climbed ARC-AGI-3
Top scores in a Kaggle contest on AI reasoning rose from about 1% to nearly 56%. Much of the gain appears to come from the code around small open models, but the numbers need a careful read.