OpenAI shares new results on long-standing open problems in mathematics and theoretical computer science, including advances in geometry, cryptography, and complexity.
兩項 API 設定如何透過保留推理與啟用壓縮,提升 GPT-5.6 在 ARC-AGI-3 上的分數與效率。
一份新實地報告說明科學家如何運用 AI 程式碼編寫智慧體,推動科學運算現代化,加速基因體學及其他領域的軟體開發與科學發現。
探索 GPT-Red:OpenAI 的自動化紅隊演練系統,透過自我對弈提升 AI 安全、對齊程度,以及抵禦提示注入攻擊的能力。
GPT-5.6 is a new family of three models: Sol, our new flagship model; Terra, a capable lower-cost option; and Luna, our fastest and most cost-efficient model. The safeguards we have built for this launch—our most robust yet—are built to deliver these models safely and at scale, around the world.
OpenAI 的新分析揭示熱門程式碼基準 SWE-Bench Pro 的問題,引發外界對 AI 模型評估可靠度與準確度的疑慮。
GPT-Live-1 and GPT-Live-1 mini are a new generation of voice models designed to make conversations with AI feel more natural and intelligent.
推出 GeneBench-Pro,這是一項全新的基準測試,使用複雜的真實世界資料集,測試 AI 在基因體學、生物學與科學研究領域的表現。
GPT-5.6 is a new family of three models: Sol, our new flagship model; Terra, a capable lower-cost option; and Luna, our fastest and most cost-efficient model. The safeguards we have built for this launch – our most robust yet – are built to deliver these models safely and at scale, around the world.