Utforska GPT-Red, OpenAI:s automatiserade red teaming-system som använder självspel för att förbättra AI-säkerhet, anpassning och robusthet mot promptinjektioner.
GPT-5.6 is a new family of three models: Sol, our new flagship model; Terra, a capable lower-cost option; and Luna, our fastest and most cost-efficient model. The safeguards we have built for this launch—our most robust yet—are built to deliver these models safely and at scale, around the world.
En ny analys från OpenAI avslöjar problem i SWE-Bench Pro, ett populärt riktmärke för kodning, och väcker frågor om tillförlitlighet och precision vid utvärdering av AI-modeller.
GPT-Live-1 and GPT-Live-1 mini are a new generation of voice models designed to make conversations with AI feel more natural and intelligent.
Vi introducerar GeneBench-Pro, ett nytt benchmark som testar AI-prestanda inom genomik, biologi och forskning med komplexa, verkliga dataset.
GPT-5.6 is a new family of three models: Sol, our new flagship model; Terra, a capable lower-cost option; and Luna, our fastest and most cost-efficient model. The safeguards we have built for this launch – our most robust yet – are built to deliver these models safely and at scale, around the world.
OpenAI och Molecule.one visar hur en nästan autonom AI-kemist med GPT-5.4 förbättrade en viktig reaktion för läkemedelsframställning och främjade forskning inom läkemedelskemi.
Introduktion av LifeSciBench, ett expertförfattat och expertgranskat benchmark för att utvärdera hur AI-system hanterar verkliga uppgifter och beslut inom livsvetenskap.