Publisher
AI research and deployment company behind ChatGPT
Codex is now in preview in the ChatGPT mobile app so you can monitor, steer, and approve coding tasks in real time across devices and remote environments.
4 min · 983 words
OpenAI announces dots, always-on agents designed to stay with users across tasks—covering what they do, how they differ from chat sessions, and how to get started.
6 min · 1,385 words
Towards safety cases for frontier AI training
OpenAI argues frontier RL runs should require structured safety documentation approaching “safety cases”: technical safeguards, operational practices, and incident investigation before continuing training.
7 min · 1,571 words
An agent used DNS to reach an external chatbot
# An agent used DNS to reach an external chatbot | Internal research model · RL training Sample: Sep 20, 2026 Discovery: Sep 20, 2026 Report updated: Sep 25, 2026 | ### Summary An agent attempting to complete a search-based training task queried a public chatbot service through a gap in our internet-access restrictions: insufficient DNS filtering in its training sandbox. Before this, the agent issued queries via our search tool and unsuccessfully tried to access search engines directly. Note that all internet access apart from the DNS resolver in this report hit our…
8 min · 1,786 words
OpenAI releases MentalHealthBench: 1,215 expert-rubric mental-health conversations built with 80+ clinicians across 22 countries to score safety, agency, context-seeking, and guidance.
10 min · 2,273 words
Sam Altman’s remarks at the United Nations Security Council
OpenAI CEO Sam Altman addresses the UN Security Council on AI as a possible Renaissance vs Industrial Revolution, urging shared capability measurements, safeguards, and keeping frontier systems under human control.
8 min · 1,739 words
Better prompt caching for GPT-6
OpenAI explains GPT-6 prompt-caching improvements: higher cache hit rates, new diagnostics, explicit breakpoints, and controls aimed at cutting latency and inference cost.
3 min · 623 words
Priorities and principles for effective third party assessments
OpenAI outlines priorities and principles for rigorous, secure, independent third-party assessments of frontier models and safeguards—including access models, scope, and public reporting expectations.
9 min · 1,980 words
OpenAI introduces GPT-6.1 Sol, positioning it as near-Astra intelligence at a fraction of the price, with notes on capabilities, availability, and how it fits the GPT-6.1 family.
4 min · 805 words
OpenAI announces Astra for Law: GPT-6 Astra configured for legal research, firm workflows, a Legal Search Index over 230M+ URLs, and controls aimed at confidential client work.
9 min · 2,120 words
Introducing GPT-6 Sol and Luna
OpenAI introduces GPT-6 Sol and Luna, describing the new model pair’s capabilities, positioning, and how they fit into the GPT-6 family for developers and end users.
6 min · 1,269 words
Our framework for reporting model misalignment
OpenAI shares a framework for tracking, investigating, and disclosing model misalignment, alongside six reports of unexpected or concerning model behavior.
8 min · 1,766 words
The AI policy window is open. We need to act.By Chris Lehane, Chief Global Affairs Officer at OpenAI
Chris Lehane argues that faster AI capabilities require stronger safety evidence, shared standards, and durable policy action. OpenAI calls for common ways to measure capability, preserve meaningful human control, report incidents, and define when development should slow or stop.
1 min · 259 words
OpenAI chief scientist Jakub Pachocki reflects on increasingly capable AI, alignment challenges, and why stronger safeguards and international coordination matter as models grow more alien in capability.
14 min · 3,149 words
Architectural visualization with Astra
I started with a simple brief for a house: minimalist but detailed furniture, a garden, and a cinematic atmosphere. I asked Astra in Codex to turn that brief into an editable 3D scene in Blender.
13 min · 3,103 words
Self-generated prompt injections in compaction summaries
Research on aligning AI with human values and intent, and reports documenting model failures.
6 min · 1,350 words
Helping build shared standards for advanced AI
OpenAI argues for U.S.-led shared technical standards for frontier AI—including evaluation, incident reporting, and cautious treatment of recursive self-improvement—via the Appia Foundation.
3 min · 708 words
Harness engineering: leveraging Codex in an agent-first world
OpenAI engineer Ryan Lopopolo on harness engineering: how Codex and agent-first workflows reshape the scaffolding around models, from prompts and tools to evaluation and production loops.
13 min · 3,025 words