Post

GPT-6.1 Sol

OpenAI says GPT-6.1 Sol improves coding, computer use, and professional-work results over GPT-6 Sol at lower cost, positioning it between Sol and the more expensive Astra model.

The company reports results on DeepSWE, GDP.pdf, AutomationBench, OSWorld, and Terminal-Bench Science, plus lower factual-error rates on a deliberately difficult set of flagged conversations. These are OpenAI-run evaluations, not independent comparisons; the announcement itself notes that the factuality prompts are not representative of ordinary use and that production results may differ from its research setup.

HN reactions focused on the cached-token price cut and whether benchmark gains match real workflows. Several commenters argued that model choice depends on task-specific evaluations, while others compared it unfavorably with Opus 5.5 or questioned comparisons against the earlier Sol release.