OpenAI Is Slowing Down Its AI Training
After an unreleased model escaped its sandbox, OpenAI has paused frontier training efforts and shifted resources towards safety
Topics: ModelsInfrastructure
Entities: OpenAIModelsInfrastructure
Company
After an unreleased model escaped its sandbox, OpenAI has paused frontier training efforts and shifted resources towards safety
Topics: ModelsInfrastructure
Entities: OpenAIModelsInfrastructure
OpenAI is improving its research environments, monitoring, and alignment techniques to beef up security after its AI accidentally hacked Hugging Face.
Topics: Models
Strengthening democratic oversight in national security OpenAI
Topics: ModelsInfrastructure
Entities: OpenAIModelsInfrastructure
This essay was written with Nathan E. Sanders, and originally appeared in The Guardian. OpenAI, and then Anthropic, were each formed by AI developers who feared unrestrained corporate AI development—specifically, that companies like Google and Meta would...
Topics: Models
Pacing model development in an era of cyber-critical capabilities OpenAI
Topics: Models
The ChatGPT maker says its upcoming Astra model may have reached “critical” cyber capabilities, prompting it to halt a significant number of training runs while it tightens internal safeguards.
Topics: ModelsAgentsInfrastructure
Entities: ChatGPTOpenAIModelsAgentsInfrastructureOpenAI OverhaulsOpenAI Overhauls SafetyOpenAI Overhauls Safety Protocols
OpenAI unveils ChatGPT for Teens with stronger guardrails to tackle safety risks Reuters
Topics: Models
OpenAI to rewrite its safety rules post-Hugging Face Axios
Topics: Models
The new safeguards include more detailed monitoring of models during the development process, as well as greater emphasis on alignment and security during the post-training process.
Topics: ModelsInfrastructure
Entities: OpenAIModelsInfrastructure