OpenAI is strengthening monitoring, alignment, and security for frontier AI models. See how new safeguards are guiding the pace of model development.
OpenAI는 기하학, 암호학, 계산 복잡도를 비롯해 수학과 이론 컴퓨터 과학 분야에서 오랫동안 해결되지 않은 미해결 문제에 대한 새로운 연구 성과를 공개합니다.
추론 유지와 컨텍스트 압축을 통해 점수와 효율을 높여 GPT-5.6의 ARC-AGI-3 성능을 개선한 두 가지 API 설정을 소개합니다.
새로운 현장 보고서는 연구자들이 AI 코딩 에이전트를 활용해 과학 컴퓨팅을 현대화하고, 유전체학을 비롯한 다양한 분야에서 소프트웨어 개발과 연구 성과 창출을 어떻게 가속화하고 있는지 보여줍니다.
셀프 플레이를 활용해 AI의 안전성, 정렬, 그리고 프롬프트 인젝션에 대한 견고성을 향상하는 OpenAI의 자동화 레드팀 시스템 GPT-Red를 소개합니다.
GPT-5.6 is a new family of three models: Sol, our new flagship model; Terra, a capable lower-cost option; and Luna, our fastest and most cost-efficient model. The safeguards we have built for this launch—our most robust yet—are built to deliver these models safely and at scale, around the world.
OpenAI의 새로운 분석 결과에 따르면 널리 사용되는 코딩 벤치마크인 SWE-Bench Pro에서 여러 문제가 확인되었으며, 이로 인해 AI 모델 평가의 신뢰성과 정확성에 대한 우려가 제기되고 있습니다.
GPT-Live-1 and GPT-Live-1 mini are a new generation of voice models designed to make conversations with AI feel more natural and intelligent.
GeneBench-Pro는 실제 연구 환경의 복잡한 데이터를 활용해 유전체학, 생물학, 과학 연구 분야에서 AI의 성능을 평가하는 새로운 벤치마크입니다.