MLE-bench: Evaluating Machine Learning Agents on Machine Learning Engineering
We introduce MLE-bench, a benchmark for measuring how well AI agents perform at machine learning engineering.
Topics: Agents
Entities: Agents
We introduce MLE-bench, a benchmark for measuring how well AI agents perform at machine learning engineering.
Topics: Agents
Entities: Agents
Hearst’s iconic brands bring curated lifestyle and local news content to OpenAI’s products.
Topics: Models
Introducing canvas
In addition to securing $6.6 billion in new funding from leading investors, we have established a new $4 billion credit facility with leading banks, including JPMorgan Chase, Citi, Goldman Sachs, Morgan Stanley, Santander, Wells Fargo, SMBC, UBS, and HSBC.
We are making progress on our mission to ensure that artificial general intelligence benefits all of humanity.
Developers can now build fast speech-to-speech experiences into their applications
Developers can now fine-tune GPT-4o with images and text to improve vision capabilities
Topics: Models
Entities: Models
Offering automatic discounts on inputs that the model has recently seen
Fine-tune a cost-efficient model with the outputs of a large frontier model–all on the OpenAI platform
Topics: Models
Altera uses GPT-4o to build a new area of human collaboration
Topics: Models
We’re introducing a new model built on GPT-4o that is more accurate at detecting harmful text and images, enabling developers to build more robust moderation systems.
Topics: Models
Entities: Models
Minnesota’s Enterprise Translation Office uses ChatGPT to bridge language gaps
Entities: ChatGPT