Introducing SimpleQA
A factuality benchmark called SimpleQA that measures the ability for language models to answer short, fact-seeking questions.
Topics: Models
Entities: Models
Source feed
1260 items
Stay up to speed on the rapid advancement of AI technology and the benefits it offers to humanity.
A factuality benchmark called SimpleQA that measures the ability for language models to answer short, fact-seeking questions.
Topics: Models
Entities: Models
Decagon and OpenAI deliver high-performance, fully automated customer support at scale
Topics: Models
We’ve simplified, stabilized, and scaled continuous-time consistency models, achieving comparable sample quality to leading diffusion models, while using only two sampling steps.
Topics: Models
Entities: Models
OpenAI and the Lenfest Institute AI Collaborative and Fellowship program
Topics: Models
We've analyzed how ChatGPT responds to users based on their name, using AI research assistants to protect privacy.
Topics: Policy
We introduce MLE-bench, a benchmark for measuring how well AI agents perform at machine learning engineering.
Topics: Agents
Entities: Agents
Hearst’s iconic brands bring curated lifestyle and local news content to OpenAI’s products.
Topics: Models
Introducing canvas
In addition to securing $6.6 billion in new funding from leading investors, we have established a new $4 billion credit facility with leading banks, including JPMorgan Chase, Citi, Goldman Sachs, Morgan Stanley, Santander, Wells Fargo, SMBC, UBS, and HSBC.
We are making progress on our mission to ensure that artificial general intelligence benefits all of humanity.
Developers can now build fast speech-to-speech experiences into their applications
Developers can now fine-tune GPT-4o with images and text to improve vision capabilities
Topics: Models
Entities: Models