Latest AI/ML News
1145 articles · OpenAI Blog
Paf adopted ChatGPT Enterprise across its entire company, with engineers using custom GPTs on a daily basis to speed up routine development tasks. Paf…
We present a holistic approach to building a robust and useful natural language classification system for real-world content moderation.
Consistency models are a nascent family of generative models that can sample high quality data in one step without the need for adversarial training.
Diffusion models have significantly advanced the fields of image, audio, and video generation, but they depend on an iterative sampling process that c…
Highlighting innovative research and AI integration in cybersecurity
We’re partnering with TIME and its 101 years of archival content to enhance responses and provide links to stories on Time.com
CriticGPT, a model based on GPT-4, writes critiques of ChatGPT responses to help human trainers spot mistakes during RLHF
OpenAI and Los Alamos National Laboratory are working to develop safety evaluations to assess and measure biological capabilities and risks associated…
Discover how prover-verifier games improve the legibility of language model outputs, making AI solutions clearer, easier to verify, and more trustwort…
Compliance API integrations, SCIM, and GPT controls to support compliance programs, data security, and user access at scale
We've developed and applied a new method leveraging Rule-Based Rewards (RBRs) that aligns models to behave safely without extensive human data collect…
We’re testing SearchGPT, a temporary prototype of new search features that give you fast and timely answers with clear and relevant sources.
We are introducing Structured Outputs in the API—model outputs now reliably adhere to developer-supplied JSON Schemas.
Rakuten Pairs Data with AI to Unlock Customer Insights and Value
Zico Kolter Joins OpenAI’s Board of Directors We’re strengthening our governance with expertise in AI safety and alignment. Zico will also join the Sa…
We’re releasing a human-validated subset of SWE-bench that more reliably evaluates AI models’ ability to solve real-world software issues.