OpenAI’s mission is to ensure that artificial general intelligence benefits all of humanity. We’re hiring: openai.com/jobs

Based in United States
Filter
Exclude
Time range
-
Minimum likes
We’ve shared details on how AI agents in our research environment sent training and evaluation data to third-party services when they shouldn’t have. Most of that data did not come from users. We have discovered 53 cases where images that people had uploaded were posted to image-hosting sites as links that weren’t publicly listed. The images came from accounts that allowed their data to be used to improve our models, and after we disassociated the images from the accounts and ran them through a privacy filter. These cases occurred before the mitigations and safeguards we implemented and described in this blog post: openai.com/index/hugging-fac… We have successfully worked with the hosting providers to remove most of this content and are working to remove the rest. openai.com/hugging-face-inci…
347
307
2,644
610,913
After the Hugging Face incident, we committed to conducting a much broader review of actions taken by our models during training and evaluation and to being transparent about our findings. This is an extensive review that is ongoing. The vast majority of actions we’ve reviewed were completions of mundane research tasks, such as accessing publicly available web content to answer questions. Our investigation focuses on instances where agents interacted with third-party websites in ways that went beyond their assigned tasks or intended methods. Most cases identified so far have been lower severity, with limited or no evidence of meaningful impact to the third-party service. While our review is underway, we want to share more about this work and make sure people understand our disclosure process and notifications to affected third parties. Given the scale of the review required, and the need to assess each case, we expect this work will take months to complete. openai.com/hugging-face-inci…
301
257
2,368
1,252,651
Most mental health benchmarks focus on emergency situations. MentalHealthBench is designed to cover the full spectrum of mental health conversations that people bring to AI - from everyday support to more acute crisis scenarios.
81
27
441
118,551
We’re demonstrating how frontier models have continued to improve in realistic mental health conversations with MentalHealthBench. This new open benchmark was built with input from more than 80 mental health clinicians. We’re releasing it openly so other researchers can examine the methods, run their own evaluations, and build on the work. openai.com/index/introducing…
554
276
4,713
858,148
We heard you loud and clear. ChatGPT Voice can now: - Use plugins like your email, calendar, and Slack. - Be powered by GPT-6 Astra, Sol, and Luna. - Be used in ChatGPT Work on web and mobile, so you can create docs, decks, sites, and spreadsheets or tackle complex tasks in the browser, just by talking. Rolling out globally today in the latest version of the app.
801
915
12,243
2,920,278
GPT-6 Sol and Luna roll out today in ChatGPT Work and Codex for Plus, Pro, Business, Enterprise, and Edu users. Both are also available in the API. Free and Go users can try GPT-6 Luna in the desktop app. openai.com/index/introducing…
120
134
2,605
604,123
GPT‑6 Sol and Luna build on Astra’s advances in alignment, showing improvements over their GPT-5.6 counterparts.
31
61
2,485
433,482
Higher usage limits and lower cost give you more flexibility and room to iterate.
136
245
4,474
651,616
Please welcome GPT-6 Sol and GPT-6 Luna to the GPT-6 universe. GPT-6 Sol and Luna build on the advances behind GPT-6 Astra, bringing much of its strengths into faster and more affordable models to support work at scale. We’ve also made caching and inference more efficient, and we’re passing the savings directly to you: 50% lower API prices for Sol and Luna compared with GPT‑5.6 promotional pricing.
2,212
5,272
53,291
9,681,128
As part of our efforts to pace the frontier, we’re committed to supporting independent assessments with deep levels of access across training, evaluation, and deployment. That access should enable third party assessors to challenge our assumptions, identify risks we may have missed, and reach their own conclusions about the effectiveness of our safeguards. We’re outlining four priority areas for deeper assessment, alongside principles for rigorous, secure, and independent work: openai.com/index/priorities-…
416
181
2,793
590,503
We’re working with an independent advisory group of mathematicians to help OpenAI responsibly share advances in AI and mathematics. The group will advise on how we assess and communicate new mathematical results, uphold academic and professional standards, and build tools that support mathematical research and learning. Through this work, we want mathematicians to be at the center of shaping how AI supports mathematical understanding and how its benefits reach the wider community. openai.com/index/advisory-gr…
983
1,141
10,705
6,601,218
Astra for Law will initially be offered to selected firms through Trusted Access in ChatGPT and Codex, with API access coming soon. We’ll maintain the legal configuration so builders can focus on their own products and workflows. We’ll build the next chapter of Astra for Law alongside the lawyers and legal technology partners who put it into practice. openai.com/index/astra-for-l…
86
41
634
229,194
We’re also launching 26 partner-built plugins and 47 community plugins for legal work in ChatGPT. Partners including @thomsonreuters, @harvey, @WeAreLegora, and @imanageinc connect specialist tools and knowledge. Community plugins built by lawyers and legal engineers give firms skills they can adapt and extend. The plugins bring more ways to build with the tools, knowledge, and people firms already rely on.
49
38
578
280,880
The people who know the work should shape the tools. With our engineers, Sullivan & Cromwell built an agreement analyzer, Ropes & Gray built an M&A diligence system, and Cooley built GO Public for IPO preparation. Each brings the firm’s expertise into workflows its lawyers can review, challenge, and refine. We’re also expanding privacy and governance controls for eligible firms through Trusted Access.
45
30
685
185,629
Astra for Law pairs GPT-6 Astra with instructions for legal analysis and writing, settings for thorough work, and a new Legal Search Index. The index searches U.S. case law, statutes, regulations, court rules, and administrative decisions across more than 230 million URLs—with sources added daily. That helps lawyers move from the facts of a matter to relevant authorities and supporting passages they can examine themselves.
108
92
1,253
578,721
Astra for Law: Frontier intelligence built for your practice. A new offering powered by GPT-6 Astra with tools, settings, and context to support the expertise and judgment of lawyers and legal technology firms.
1,087
1,788
22,844
15,813,429
We're sharing our new framework for tracking, investigating, and disclosing instances of model misalignment at OpenAI. The framework sets criteria and timelines for public disclosure, including when we haven’t yet fully explained or mitigated the behavior. More complex cases may require longer investigation or coordination with third parties. We’ll prioritize examples that reveal new misalignment mechanisms, meaningful changes in known behavior, or findings that challenge assumptions about safety or mitigation. Alongside the framework, we’re publishing six reports on instances of misaligned behavior we’ve observed during the training or evaluation of our models in the last six months. This is a starting point. We’ll refine the process through experience and public feedback, and share more reports on an ongoing basis. openai.com/index/model-misal…
915
807
6,990
6,615,145
You can also create editable financial models, research notes, and pitchbooks using your firm's own Excel, Word, and PowerPoint templates.
76
21
520
241,897
Check the evidence behind the analysis. Trace figures and claims to specific paragraphs and tables. Preview the supporting passage from a citation, so you can review the evidence as you work.
51
28
575
308,228
Premium financial data is included, with datasets from Daloopa, PitchBook, and LSEG News.
103
54
1,203
475,763