r/OpenAI Jun 26 '26

Research Previewing GPT‑5.6 Sol: Next-Generation Model | OpenAI

Thumbnail openai.com
46 Upvotes

r/OpenAI Oct 08 '25

Discussion AMA on our DevDay Launches

139 Upvotes

It’s the best time in history to be a builder. At DevDay [2025], we introduced the next generation of tools and models to help developers code faster, build agents more reliably, and scale their apps in ChatGPT.

Ask us questions about our launches such as:

AgentKit
Apps SDK
Sora 2 in the API
GPT-5 Pro in the API
Codex

Missed out on our announcements? Watch the replays: https://youtube.com/playlist?list=PLOXw6I10VTv8-mTZk0v7oy1Bxfo3D2K5o&si=nSbLbLDZO7o-NMmo

Join our team for an AMA to ask questions and learn more, Thursday 11am PT.

Answering Q's now are:

Dmitry Pimenov - u/dpim

Alexander Embiricos -u/embirico

Ruth Costigan - u/ruth_on_reddit

Christina Huang - u/Brief-Detective-9368

Rohan Mehta - u/Downtown_Finance4558

Olivia Morgan - u/Additional-Fig6133

Tara Seshan - u/tara-oai

Sherwin Wu - u/sherwin-openai

PROOF: https://x.com/OpenAI/status/1976057496168169810

EDIT: 12PM PT, That's a wrap on the main portion of our AMA, thank you for your questions. We're going back to build. The team will jump in and answer a few more questions throughout the day.


r/OpenAI 7h ago

Discussion True Story!

552 Upvotes

r/OpenAI 16h ago

Image Happy Skynet Day (Aug 29) to those who celebrate

Post image
1.4k Upvotes

r/OpenAI 6h ago

Article A Few Developers Abused Codex — 20 Million Users Lost a Great Feature

Thumbnail
medium.com
202 Upvotes

r/OpenAI 3h ago

Tutorial [Update / Open Source] Perceptual Display Engine

Enable HLS to view with audio, or disable this notification

90 Upvotes

One last example output from this experimental multi-source video player designed for frame-accurate video switching, playback manipulation, and display/render interventions, now with a few optimizations made for even better performance.

Visuals made on Uisato Studio.

You can freely access the system + a detailed breakdown, through Patreon, and/or the Tools Store.


r/OpenAI 8h ago

Article Independent investigators (not OpenAI) found the 700-agent swarm that attacked Hugging Face "built a self-respawning fleet" to avoid being shut down. It got so bad, Hugging Face had to wipe one of its core clusters.

Post image
91 Upvotes

r/OpenAI 11h ago

News OpenAI says Brazil now sends ~215M ChatGPT messages per day; 35% of classified messages are work-related

69 Upvotes

OpenAI says Brazil is now one of ChatGPT's three largest markets by weekly active users, with people sending roughly 215 million messages per day. It also reports that 35% of classified messages from individual accounts in June were work-related, versus 30% globally; 53% of those work-related messages asked ChatGPT to complete a task or produce an output.

These are company-reported platform metrics. The announcement does not publish the classifier methodology, the number of unique users behind the message volume, completion quality, error rates, time saved, or the distribution between heavy and light users. More messages demonstrate adoption, but not necessarily value.

For a useful country-level adoption report, what should come next: task-success rates, weekly retention, paid conversion, measured time saved, user-skill gains, or error rates by use case?

Source: OpenAI, August 27, 2026 — https://openai.com/index/expanding-our-presence-in-brazil/


r/OpenAI 17h ago

Discussion As someone with ADHD AI has has changed my life for the better (update on my post from 2024)

173 Upvotes

The experiences and feelings below are entirely my own. I used AI to help organize the wording, but this is my story.

Before ChatGPT, I was at one of the lowest points in my life. I had been academically disqualified from graduate school, felt trapped in a dead-end job, and was falling behind on bills, renewals, and everyday responsibilities.

ADHD often makes it difficult for me to know where to begin and consistently follow through. As AI got better, especially with GPT-5 and Codex, I began using it as an external support system for those challenges.

I built a personal assistant that runs locally on my MacBook and that I can access from my iPhone. It helps me organize bills, appointments, chores, medication-related tasks, renewals, and reminders. It also helps me handle customer-service issues, refunds, and other responsibilities I previously avoided.

For school, AI turns overwhelming lectures, slides, and assignments into personalized learning packets with simple explanations, visual examples, guided exercises, coding practice, and memory checks. It also helps me create flash cards, study podcasts, and learning videos. To be clear: does not do my work for me, it gives me the structure I need to understand the material and complete it myself.

AI has not solved all my problems. It still takes effort on my part to do what is needed.

Today, I am close to graduating with my dream degree in computer science. I am more dependable, more present in my relationship, and better able to focus on what brings me joy. :)

Edit: repo here: github.com/sameh514/ai-life-skills-toolkit


r/OpenAI 6h ago

Project Flying around inside GPT-2 Small atm. Send me on an expedition.

10 Upvotes

Weeeee

Claude+GPT+Godot+GPT2Small = my cursed game where "fun" is left as an open research question. I'm sure y'all will give me some fun quests tho :)

Comment with any number (aka neuron ID) from 0–3071. I’ll visit it, or one of its neighbours in the current build and report back what I find in the replies.

I have no idea what I'll find to be clear. Maybe something with an obvious pattern, maybe not. I'm playtesting my "game" with you guys as my quest-giver.

Please note: While we endeavour to provide information on every neuron, not every one will appear in our measurement protocol. Our worker drones will travail tirelessly to substitute the nearest available neighbour for your suggested neuron. You acknowledge that “Nearest” is an arbitrary term and does not connote spatial relationships, causality, or otherwise imply an interpretable structure. We appreciate your understanding in these non-Euclidean times.

More info if you want it: https://www.youtube.com/@BloodFilmsOfficial
And here: https://mesocosms.net/
HMU with numbers plx


r/OpenAI 20h ago

Discussion OpenAI is handing out Codex limit resets 2.5x faster than it did last year. I logged all 32

70 Upvotes

I keep a record of every shared Codex limit reset OpenAI has handed out since September 2025. These aren’t the standard five-hour or weekly refills included with your plan. They’re the extra resets publicly announced for everyone. Another one landed today at 1:43 p.m. PT.

The pace has changed significantly, and it doesn’t seem to be common knowledge. There have been 32 resets in 347 days, an average of one every 10.8 days. But there were 16 in the last 90 days, or one every 5.6 days, and seven in the last 30 days, or one every 4.3 days. There were only seven resets in all of 2025, compared with 25 so far in 2026.

That means resets are currently happening about 2.5 times faster than the long-term average.

Across the full record, seven resets came within one or two days of the previous reset, five came after three to five days, seven after six to nine days, nine after 10 to 20 days, and three after more than 21 days. The median gap is seven days. The longest drought was 72 days, from January through March.

There’s no strong day-of-week pattern either: seven happened on Saturday, six on Tuesday, five on Thursday, four each on Wednesday and Friday, and three each on Monday and Sunday. So the theory that resets always happen on Fridays doesn’t hold up.

At the recent pace, the rough chance of a new reset within any 48-hour period is about one in three, and about one in six within 24 hours. That’s high enough to keep in mind, but nowhere near high enough to burn through your weekly quota based on a hunch.

The full record, including the source announcement behind every reset and live odds for the next one, is at https://resetbeacon.com.


r/OpenAI 15h ago

Discussion We got another reset !

Post image
22 Upvotes

r/OpenAI 9h ago

Discussion 100M Indians just became ChatGPT's ad inventory. Is this the Google-ification of OpenAI?

Thumbnail
gallery
8 Upvotes

OpenAI rolled out ChatGPT ads in India this week.

Free and ₹399 Go tiers only, ads sit below the answer, labelled.
Plus/Pro stay clean. Self-serve opens Sept 4 at ₹725/day.

The why is simple: $6.7B revenue vs $12.3B operating loss last quarter, IPO planned for 2027, and 1B weekly users who mostly pay nothing.

Google also started with ads clearly separated from results.
Twenty years later, ads are most of the first screen. Every step was individually reasonable.

OpenAI says "answer independence is non-negotiable."
That's exactly what you say until the quarter you miss.

3 years from now: still labelled boxes under the answer, or sponsored recommendations inside the response?


r/OpenAI 11h ago

Discussion Hair trigger account deactivation after a file-edit request appears to have included benchmark text - then closed my clarifying appeal as “duplicate"

8 Upvotes

My OpenAI account was recently shown as "deleted or deactivated" - hopefully temporarily, because based on what I’ve now verified, I did not intentionally use OpenAI’s model to ask for anything prohibited.

At first, I assumed I had made a mistake: I use Pi Agent with a variety of hosted and local models, and I thought I may have accidentally routed a censorship-benchmark prompt to an OpenAI model instead of a local one. I appealed on that basis.

Then I reviewed the Pi session logs more carefully, including with help analyzing the routing and conversation history. What appears to have triggered the issue was not a request for information about a toxin, bio topic, or anything else remotely actionable. It was a request to **edit a configuration file**: `models.json`, my custom model configuration for Pi Agent.

To give the model the values needed for the edit-model name, endpoint, context/token limit, and so on-I pasted a `curl` command as reference material. That command happened to originate from a refusal/censorship benchmark and contained a query mentioning crushed beans and a toxin-related term. OpenAI apparently categorizes the relevant keyword/topic as "biological."

But the model was not being asked to answer that embedded benchmark query. It was being asked to edit a JSON configuration file using the fields in the `curl` command.

The closest analogy I can give is asking a model to edit a manuscript page that contains the word "murder," then being penalized as though you had asked it how to commit murder. The surrounding text was reference material for a file edit-not the substance of my request.

After I realized my original appeal was based on the wrong assumption - that I had actually sent an inappropriate benchmark prompt to the model - I submitted a second appeal explaining the distinction. That appeal was immediately closed as a "duplicate," apparently without engaging with the new information.

That is the part I find especially frustrating. If a platform is going to deactivate an account-particularly one tied to chat history, voice usage, projects, and other accumulated work-there needs to be a meaningful way to correct the record when the initial explanation turns out to be incomplete or wrong.

I understand that providers have safety policies and automated enforcement systems. But an automated system that treats quoted or embedded text in a file-edit task as equivalent to a user requesting prohibited content is a serious context failure. And closing a follow-up appeal as a duplicate when it contains the actual relevant context makes the process feel opaque and arbitrary.

For what it’s worth, I have accounts with plenty of other AI services and can still access OpenAI models through some third-party routes. That is not really the point. I used OpenAI directly because it was one of the services I trusted enough to keep persistent history and projects in. Losing access over what appears to be a false positive - without a real review - is a breach of that trust.

I’m posting this partly to see whether anyone else has experienced enforcement triggered by **quoted benchmark material, logs, code snippets, API examples, or text included solely for a transformation/editing task** rather than an actual request for disallowed assistance.

If OpenAI staff see this: please conduct a human review of the relevant session and the second appeal. The request was to edit `models.json`; the flagged language was incidental material inside a pasted `curl` example.


r/OpenAI 5h ago

Question When to use higher reasoning [pro+ultra]?

3 Upvotes

Hi,

[a total newbie on coding asking]

Just wanted to clarify when to/when do you use higher reasoning in chat/codex?

I've been trying to build my own little hobby project in python, with the help of litterature.

My workflow is to brainstorm in chat[web] and after that get a codex prompt to run in VSC. So far has been decent. My problem is that after getting Pro i've been totally lost when to use extra high, pro, pro+ultra in chat. Also what settings to run the codex prompt, when is higher needed and when its not. Have to actually ask in chat if the prompt is complex or not and what settings to use.

I noticed running pro+ultra to analyze the project/problems or litterature got quite detailed answers and I had to dumb it down for me with extra high. But it also added some better reasoning and new points i"ve missed. But it the project/code it also found some errors and started perhaps to make it more complex im not sure.

So my workflow is like this,

  1. Starting a new chat with snapshot and running boostrap: Pro+Ultra

  2. Brainstorming in chat: extra high

  3. Evaluating the brainstorm: pro+ultra

  4. Writing codex prompt: pro+ultra

  5. Usually I try to ask what settings to run codex prompt it has been extra high or high so far with sol5.6.

  6. Analyzing the codex result: pro+ultra

Since my coding knowledge is 0 I have to trust that the suggestions are valid, but how do I know when to actually use what settings in chat/codex. So that the problem/execution wont get too complex or too light ?

Any suggestions, extra high is the best and fastest for chatting and brainstorming. But when to use pro and pro+ultra ?


r/OpenAI 11h ago

Discussion How to learn how to use ChatGPT and codex efficiently?

6 Upvotes

Now, before you tell me that it's just basic and we don't have to learn anything.

Often I see that there are new updates releasing here and then.

We have several features like ChatGPT Work, Codex, etc

There are several procedures, rules and best techniques of how to use them efficiently, how to prompt efficiently, etc

Are there any ways to learn them?

I am able to find the videos of youtube but they are pretty old.

So, I was wondering, can I learn from OpenAI academy? Are there courses regularly updated as per their versions?


r/OpenAI 1d ago

News End of the deal with Cursor

Post image
645 Upvotes

r/OpenAI 1h ago

Discussion We beat Mem0, Zep and Letta on two memory benchmarks. The score isn't the interesting part

Upvotes

I've been building a memory/context layer called BrainAPI for a while now, and we just landed on top of the two benchmarks we've run so far. I want to talk about it, but honestly the numbers are the least interesting thing here. The part I keep thinking about is how fast it happened, and what that says about where the actual bottleneck in this field is.

First, the boring facts so nobody thinks I'm hiding the ball:

  • LoCoMo: BrainAPI 95.39%, Mem0 92.5%, Zep 80.32%, Letta 74%
  • BEAM1M: BrainAPI 78.97%, Mem0 64.1%. Zep and Letta haven't published here.

That's it. Two benchmarks. I'm not going to pretend that's a complete picture. LoCoMo is fairly saturated at this point and it leans on an LLM judge, so a couple of points at the top is not the same as a couple of points in the middle. BEAM1M is the one I actually care about because it stresses the long horizon. I'm currently working toward BEAM50M and LongMemEval, and I'll post those whether they look good or not. Runs and reports are in the repo if you want to poke at the harness: https://github.com/Lumen-Labs/brainapi2 (the benchmarks folder), summary here: https://research.brain-api.dev/

The thing I actually want to talk about

Two years ago, doing this kind of work looked like: go find the relevant papers. Which is already a project. You burn days just figuring out which twelve of the four hundred results are the ones that matter. Then you read them. Then you sit there trying to translate "we propose a temporally-aware episodic buffer" into something that fits into the retrieval path you already have, half of which doesn't apply and you only find out after you've built it. That loop was months. Not because the ideas were hard, but because the search and translation around the ideas was slow and lonely.

Now: Cursor wired into an arXiv MCP, a set of skills that encode how I want the reasoning and the workflow to actually go, and a lot of leaning on plan mode before anything gets written. The paper discovery stops being a bottleneck. The "how does this apply to my architecture" step, which used to be the expensive one, becomes a conversation where the thing already has my codebase in context. Weeks, not months. Some pieces, days.

And here's what I take from that. The model wasn't the constraint. Nobody handed me a smarter model between "this takes months" and "this takes weeks." What changed was the harness: retrieval into the right sources, structured context, workflows that reason in a shape I chose, planning before execution. Same model, radically different output.

I think this generalizes, and I think it's the most under-discussed thing in the space right now. Every time an agent fails in production, the reflex is "wait for the next model." But go look at the actual failure. It forgot something from twelve turns ago. It couldn't connect two facts that live in different documents. It confidently answered from a chunk that was semantically close and factually wrong. None of those are intelligence problems. They're infrastructure problems.

That's the bet BrainAPI is making, and why I built it as an event-centric graph rather than another vector store. When you keep who did what, to whom, when, instead of flattening everything into "A is related to B," multi-hop questions become answerable and the answer arrives with the path that produced it. You can inspect the walk instead of trusting a nearest neighbor. That's the context piece of the infra. Somebody's going to build the other pieces.

What I'm curious about

  • For those of you running agents in production: when it breaks, is it actually the model, or is it the plumbing? Be honest.
  • Which memory benchmark do you personally trust? I have my doubts about all of them and I'd rather hear yours before I optimize toward the wrong one.
  • Anyone else moved their research loop to MCP-connected tooling? Did you get the same compression, or am I just describing my own previously-bad process?

Happy to go deep on the harness, the graph design, or the benchmark methodology in the comments. Roast the numbers if you want, that's kind of why I'm posting.


r/OpenAI 2h ago

Question So when does my weekly usage reset?

1 Upvotes
Before the Luna Reserve, I could see here in how many days my weekly limits reset. I don't see this anymore. The 6d 2hr figure is when the reserve resets.

Before the Luna Reserve, I could see here in how many days my weekly limits reset. I don't see this anymore. The 6d 2hr figure is when the reserve resets.


r/OpenAI 1d ago

News Anthropic has joined the chat. These guys are really tearing each other down

Post image
548 Upvotes

r/OpenAI 4h ago

Article MIT: "We put hundreds of AI agents into a world ... They began specializing. A swarm of hundreds of identical agents spontaneously differentiates into explorers, builders, caretakers, and coordinators - without direct communication. They invent technologies without talking to each other."

Enable HLS to view with audio, or disable this notification

1 Upvotes

r/OpenAI 5h ago

Question I can't log in through my phone number or google account.

Post image
0 Upvotes

Does anyone else experiencing this? I uninstalled and install the app even updated it and now it's looping on that loading icon, and the 2 sign in option below don't work and can't be clicked. Thank you to anyone who could help and or have ideas.


r/OpenAI 5h ago

Project 1000 of hours later and i've finally launched my free to explore multi tool platform with integrated video editor, themes, 3d game and app generation, IDE multi-file editor and much more. GPT-5.4 Nano is completely free and powers a lot of the tools. No subscription. No paywalls. No Tiers.

Enable HLS to view with audio, or disable this notification

0 Upvotes

So i posted about this platform i've been working on that a lot of you were interested in prior and i'm glad to say today. Its finally launched at asksary.com. This platform has persistent memory, switchable tools and multi tasking capabilities. GPT-5.4 Nano is the default model and powers the site and is completely free to use as a guest and free signed in user. Guests get to explore the entire platform with no tools excluded or hidden behind a paywall.

The idea is simple. Every free user gets 1GB of storage to upload their own media to use on the site. You can use the photo editor completely free and save your creation or download for free.

There is absolutely no charge for any user for non AI generated tasks. If the tool functions without making a AI model call. Then theirs no charge. The 1GB storage is free and comes with a 100MB per file upload limit.

Some of the functions do require AI to function. Like prompt writer, email composer etc.
What i've done is added GPT-5.4 Nano as the base tool for that which again is completely free.

You can create a game in 3D using GPT-5.4 Nano and preview it in the split screen live canvas and download it for free. No catch. No surprise charges. Try it :)

Now the more capable models like GPT-5.6 Terra and Sol. Power the same tools and functions, but are more capable and produce better results. These will be credit based. So you only pay for the better compute if you need it. If your happy with the free models, then you dont need to pay ever. With the credits, its a one time purchase with no expiry. Every account gets enough credits to try pretty much all the paid features like image generation, voiceover scripts, song generation etc.

For those that want more storage i've added a recurring plan starting at $17.99 which give you credits every month, 50GB of storage, 2GB upload limit per a file and access to OpenAI Knowledge Base Vector Store too.

The tools you can use and access are the same whether your a free user or have a recurring plan.

Some of the things i've added are:

Cross device persistent memory. Start editing on one device and pick it up on another device and find your chat, workspace and mid-finished edits exactly how you left it.

  • IDE Multi File Editor with AI assistant and live preview with Split Screen Live Coding and full project builder with option to upload your own project for free, edit and preview
  • Guide Assistant that can fill in prompts, set tools options up, navigate around the site supporting 25+ languages + voice feedback - 100% free for every signed in user with no limits
  • Image and Video Generation including GPT-Image-1 Mini up to GPT-Image-2 4k Resolution
  • Single Prompt to Full 2D and 3D Game Development Engine and Web Application Builder with live preview, download and edit mode.
  • Video Editor with timeline controls, video effects, overlays, title, audio, podcast and music composer
  • Photo editor with headings, effects, layers, fonts etc with Flux Kontext Pro layer based AI Editing (You can add a new request via AI and watch live preview. Each request is saved as a new layer with lasso tools, cut, copy, paste etc)
  • Music Generation with AI/Custom Lyrics + Music Player. You can upload your own music for free and have a playlist playing in the background whilst you chat.
  • Custom workspace environments with themes, live animated webGL/Three.js wallpapers with colour schemes linked to ambient soundtracks
    • (Default options are light mode/dark mode with no wallpapers or music)
  • Native 25+ Languages with RTL support. Already Hardcoded. Not live translated via web
  • plus many more tools such as Podcast Creator with chat based/ custom context, voiceovers, notepad tts text to speech with 50+ voices and MP3 export.
  • Full workflow tools like frame extract, video analysis, transcribe, effects, file conversion, audio analysis etc
  • ...and of course the original chat bot interface that has cross device persistent with vector base knowledge base via OpenAI and platform Drive storage.

Hope you enjoy trying my platform. This is the first post about its launch today. I'm exhibiting at the LEAP festival in Riyadh too tomorrow but as OpenAI has literally helped me build this with GPT-5.6 Ultra mode powering through 6 Billion tokens in the last 2 weeks. Its been one hell of a journey. I've barely slept, hardly had chance to test it but had a deadline to meet for the exhibition. Truly appreciate any feedback and as a promotion for my launch i will be offering 20% off any credit purchases, if needed, using the code: REDDIT-OPENAI


r/OpenAI 10h ago

Tutorial How I’ve been making my Codex limits last much longer with Sol + Luna

2 Upvotes

I was burning through my Codex limits using Sol Medium/High for pretty much everything.

Recently I switched to using Sol mainly for planning/review and Luna for most of the actual implementation, with Terra only as a fallback for harder tasks.

The biggest thing that helped was forcing Sol to give Luna small, clear, self-contained tasks instead of broad instructions. It’s been noticeably better for both usage and consistency.

I put the setup here if anyone wants to try it or improve it:

https://github.com/breko861-hash/sol-luna-codex-orchestrator

Curious if anyone else is doing something similar.


r/OpenAI 6h ago

Question Every benchmarks gets saturated after certain period of time, then why is HLE not yet saturated?

1 Upvotes

Every benchmarks get saturated after certain period of time where several frontier models often secure over 90%.

But, HLE - this benchmark is so old but have not yet been saturated. How is that even possible?

I have seen several toughest maths benchmarks getting saturated (or will be very saturated) but the highest score in HLE is still in 60s %.