GPT-Red જુઓ: OpenAI ની સ્વચાલિત રેડ ટીમિંગ સિસ્ટમ, જે AI સલામતી, સંરેખણ અને પ્રોમ્પ્ટ ઇન્જેક્શન દૃઢતા સુધારવા સેલ્ફ-પ્લેનો ઉપયોગ કરે છે.
GPT-5.6 is a new family of three models: Sol, our new flagship model; Terra, a capable lower-cost option; and Luna, our fastest and most cost-efficient model. The safeguards we have built for this launch—our most robust yet—are built to deliver these models safely and at scale, around the world.
OpenAI નું નવું વિશ્લેષણ લોકપ્રિય કોડિંગ બેન્ચમાર્ક SWE-Bench Pro માં સમસ્યાઓ દર્શાવે છે, જે AI મોડલોના મૂલ્યાંકનની વિશ્વસનીયતા અને ચોકસાઈ અંગે ચિંતા ઊભી કરે છે.
GPT-Live-1 and GPT-Live-1 mini are a new generation of voice models designed to make conversations with AI feel more natural and intelligent.
પ્રસ્તુત છે GeneBench-Pro, એક નવો બેન્ચમાર્ક જે જટિલ, વાસ્તવિક દુનિયાના ડેટાસેટ્સનો ઉપયોગ કરીને જિનોમિક્સ, જીવવિજ્ઞાન અને વૈજ્ઞાનિક સંશોધનમાં AIના પ્રદર્શનનું પરીક્ષણ કરે છે.
GPT-5.6 is a new family of three models: Sol, our new flagship model; Terra, a capable lower-cost option; and Luna, our fastest and most cost-efficient model. The safeguards we have built for this launch – our most robust yet – are built to deliver these models safely and at scale, around the world.
OpenAI અને Molecule.one દર્શાવે છે કે GPT-5.4 નો ઉપયોગ કરતો લગભગ સ્વાયત્ત AI રસાયણશાસ્ત્રી કેવી રીતે કામ કરે છે. દવા-નિર્માણની એક મહત્વપૂર્ણ પ્રતિક્રિયામાં સુધારો કર્યો, જેના કારણે ઔષધીય રસાયણશાસ્ત્ર સંશોધન આગળ વધ્યું.
પ્રસ્તુત છે LifeSciBench: વાસ્તવિક લાઇફ સાયન્સ સંશોધન કાર્યો અને નિર્ણયો AI સિસ્ટમો કેવી રીતે સંભાળે છે, તેના મૂલ્યાંકન માટે નિષ્ણાતો દ્વારા રચિત અને સમીક્ષિત બેન્ચમાર્ક.