Jelajahi GPT-Red, sistem red teaming otomatis OpenAI yang memakai self-play untuk meningkatkan keamanan, keselarasan, dan ketahanan terhadap injeksi prompt.
GPT-5.6 is a new family of three models: Sol, our new flagship model; Terra, a capable lower-cost option; and Luna, our fastest and most cost-efficient model. The safeguards we have built for this launch—our most robust yet—are built to deliver these models safely and at scale, around the world.
Analisis baru OpenAI mengungkap masalah di SWE-Bench Pro, benchmark pengodean populer, yang memunculkan kekhawatiran atas keandalan dan akurasi evaluasi model AI.
GPT-Live-1 and GPT-Live-1 mini are a new generation of voice models designed to make conversations with AI feel more natural and intelligent.
Memperkenalkan GeneBench-Pro, benchmark baru yang menguji performa AI dalam genomika, biologi, dan penelitian ilmiah menggunakan kumpulan data dunia nyata yang kompleks.
GPT-5.6 is a new family of three models: Sol, our new flagship model; Terra, a capable lower-cost option; and Luna, our fastest and most cost-efficient model. The safeguards we have built for this launch – our most robust yet – are built to deliver these models safely and at scale, around the world.
OpenAI dan Molecule.one menunjukkan bagaimana ahli kimia AI yang hampir otonom yang menggunakan GPT-5.4 meningkatkan reaksi kunci dalam pembuatan obat, memajukan penelitian kimia medisinal.
Memperkenalkan LifeSciBench, benchmark yang ditulis dan ditinjau pakar untuk menilai cara sistem AI menangani tugas dan keputusan riset ilmu hayati nyata.