Probably some combination of: some of the 372 problems were easier than the rest; the AI got lucky on these 372; there were existing papers out there in the literature which proved especially helpful for these 372; and other similar factors.
In the past 2 years the AI's started solving math problems in roughly the order of "hardness" as ranked by humans.