GPT-6.1 Sol brings near-Astra task performance to a cheaper model with Critical cyber capability
OpenAI released GPT-6.1 Sol on September 29, reporting Astra-matching performance on DeepSWE v1.1 at roughly one-fifth of the task cost and more than double GPT-6 Sol's maximum-effort score on Terminal-Bench Science. Its system card classifies it as Critical for cybersecurity and High for biological and chemical capability, with the same safeguards stack as Astra.
Why it made the cut: Substantially cheaper access to strong coding and scientific agents, accompanied by a consequential capability-risk classification, matters beyond a routine model refresh. The comparisons are company-reported and depend on reasoning effort, tools, and evaluation setup; they do not establish equal real-world research ability. Astra remains the stronger scientific model in the reported tests, and lower token prices do not guarantee proportionately cheaper completed work.
System card addendum · Official release and evaluation summary
Link to this post