셀프 플레이를 활용해 AI의 안전성, 정렬, 그리고 프롬프트 인젝션에 대한 견고성을 향상하는 OpenAI의 자동화 레드팀 시스템 GPT-Red를 소개합니다.
토큰 하나하나에서 더 높은 지능을 이끌어내고, 비용 대비 성능은 더욱 높이며, 가장 어려운 작업에서는 필요에 따라 더 강력한 성능을 제공합니다.
GPT-5.6 is a new family of three models: Sol, our new flagship model; Terra, a capable lower-cost option; and Luna, our fastest and most cost-efficient model. The safeguards we have built for this launch—our most robust yet—are built to deliver these models safely and at scale, around the world.
OpenAI의 새로운 분석 결과에 따르면 널리 사용되는 코딩 벤치마크인 SWE-Bench Pro에서 여러 문제가 확인되었으며, 이로 인해 AI 모델 평가의 신뢰성과 정확성에 대한 우려가 제기되고 있습니다.
사람과 AI가 더욱 자연스럽게 소통할 수 있도록 설계된 차세대 음성 모델로, 이제 ChatGPT 음성 대화를 지원합니다.
GPT-Live-1 and GPT-Live-1 mini are a new generation of voice models designed to make conversations with AI feel more natural and intelligent.
GeneBench-Pro는 실제 연구 환경의 복잡한 데이터를 활용해 유전체학, 생물학, 과학 연구 분야에서 AI의 성능을 평가하는 새로운 벤치마크입니다.
OpenAI는 코딩, 과학, 사이버 보안 분야에서 한층 강화된 성능과 OpenAI 역사상 가장 발전된 안전 체계를 갖춘 차세대 모델 GPT-5.6 Sol을 선보입니다.
GPT-5.6 is a new family of three models: Sol, our new flagship model; Terra, a capable lower-cost option; and Luna, our fastest and most cost-efficient model. The safeguards we have built for this launch – our most robust yet – are built to deliver these models safely and at scale, around the world.