Inside OpenAI’s big play for science
OpenAI launched its OpenAI for Science team in October 2025 to test how GPT-5-class models can support scientists. Kevin Weil said GPT-5.2 scored 92% on GPQA versus GPT-4’s 39%; the piece also notes OpenAI deleted posts that overstated old-paper retrieval as solving unsolved math problems.
Why it matters: Strong HKR-H/K/R: the piece has an insider-angle hook, a concrete GPQA 92% vs 39% data point, and a real tension between scientific ambition and overclaim risk. It stays at 80 because this is reported strategy analysis, not a new model release or shipped capability.