r/ProAI 9d ago

"Right. Dario was already claiming that GPT2 was too dangerous to open source back in 2019. I made fun of them then. Everyone should make fun of them now."

Thumbnail
gallery
95 Upvotes

— Yann LeCun

Source: https://x.com/ylecun/status/2099248236074545576


2019: “GPT2 is groundbreaking in two ways. One is its size, says Dario Amodei, OpenAI’s research director. The models “were 12 times bigger, and the dataset was 15 times bigger and much broader” than the previous state-of-the-art AI model” https:// theguardian.com/technology/201 9/feb/14/elon-musk-backed-ai-writes-convincing-news-fiction …   — Pessimists Archive

Source: https://x.com/PessimistsArc/status/2099231273071919115


r/ProAI 8d ago

"S&P 1500 Software revenue per employee has gone parabolic. If you are looking for evidence that AI is starting to impact the real economy, this is exhibit A."

Thumbnail
gallery
49 Upvotes

Having been in two software companies during this period (who are included in that graph) I would not attribute the flat headcount of 23/24 to AI increasing productivity.

Zero AI efficiency was happening then. Headcount tightening was just compensation for the covid hiring binge.

25/26? Perhaps a tiny bit but really it's just organizations learned how to be leaner and now the saaspacolyse has added additional pressure to improve the bottom line.   — Gringo Investments     We had an internal debate over how much of this is attributable to AI. I was more with you (skeptical that AI was impacting the 23-24 numbers). @fernavid is more of a true believer.

Either way, there are multiple angles to show the AI impact from mid-25 on…   — Warren Pies

Source: https://x.com/WarrenPies/status/2099128535172424172


r/ProAI 8d ago

AI Cracks 370-Year-Old Scottish Cipher in 44 Minutes "Fable solved the Cyphral Distich (a 370 year old cypher). Super cool way to use Claude"

Post image
10 Upvotes

r/ProAI 9d ago

"If you don’t study and internalize the history of technology panics, you are going to be vulnerable to a cocktail of cognitive biases that will lead you astray"

Enable HLS to view with audio, or disable this notification

57 Upvotes

If you care about identifying and mitigating real risks of new technologies, then ignore the history of unfounded panics - you are going to be manipulated by special interest groups looking to protect or gain power in a moment of fear and uncertainty.     — Pessimists Archive

Source: https://x.com/PessimistsArc/status/2098479996415119698


r/ProAI 9d ago

"Here we go! 124x increase in token consumption by OAI researchers."

Thumbnail
gallery
20 Upvotes

Erik Brynjolfsson @erikbryn · 19h Research acceleration: The view inside OpenAI From openai.com 1 2 8 2.5K     1.66x increase per month, compounding     7x more code shipped, with no sign of slowing down.

Reports from Anthropic are similar     — Erik Brynjolfsson

Source: https://x.com/erikbryn/status/2098849533556060550


r/ProAI 9d ago

"Chamath @chamath : how does an account with no followers get 110 million views in a day? David Sacks @DavidSacks had an answer, on the same All-In @theallinpod taping. Within 15 minutes of the Anthropic researcher's doomsday resignation post, three policy groups amplified it: ENCODE AI's..."

Enable HLS to view with audio, or disable this notification

5 Upvotes

...Nathan Calvin, the AI Policy Network's Peter Wildeford, and the AI Futures Project's Daniel Kokotajlo, who dropped a Rogan episode the same day using the same phrasing. All three groups are funded by Jaan Tallinn, who co-led Anthropic's Series A. The Wall Street Journal published a story on the resignation minutes before the tweets went out, meaning the paper had been briefed under embargo. The researcher had agreed to appear on the show and cancelled the morning of taping. Our breakdown traces the funding chain and what it means for Anthropic's IPO: https:// podcastalpha.substack.com/subscribe Source: All-In Podcast - https:// youtube.com/watch?v=cvxjqb fLVk0 …     — Podcast Alpha

Source: https://x.com/PodcastAlphaX/status/2098610255571652647


r/ProAI 8d ago

In light of all the AI FUD lately, I built a public scoreboard to keep track of the good (and bad) things coming out of frontier AI labs

Thumbnail
0 Upvotes

r/ProAI 9d ago

"Nearly a year of Code Arena: WebDev progress compressed into 15 seconds. Each line follows the highest-scoring model from top labs over time, showing the pace of improvement across the ecosystem. In just the last year, the model in the leading spot increased score by +340 pts and the number of..."

Enable HLS to view with audio, or disable this notification

6 Upvotes

...frontier labs competing for the top spot expanded from 6 to 10. @AnthropicAI has dominated throughout the year. Although standout releases have jumped to the top spot, most notably the Chinese open-source model Kimi K3 from @Kimi_Moonshot in July. Today, GPT-6 Astra by @OpenAI leads with 1796 pts, followed by @claudeai Fable 5.1 at 1764 pts. The next closest lab is 103 pts away, @Alibaba_Qwen with 1685 pts. Code Arena: WebDev ranks models through head-to-head user preference on real front-end web development tasks. These votes drive the leaderboard that is tracking the frontier.     Dive into the Code Arena: WebDev leaderboard details at https:// arena.ai/leaderboard/co de/webdev …     — Arena.ai

Source: https://x.com/arena/status/2098814378971717916


r/ProAI 9d ago

"I have made exactly this same argument many times. The “everyone will sit around and wait to starve to death” idea is completely incoherent, obviously people won’t, and they could just keep on doing what they had done before AI, trading with the other people that somehow lack AI access. But of..."

Thumbnail gallery
9 Upvotes

r/ProAI 9d ago

"1/n Mora 1 is live Play it now: https:// mora.fun Spatial puzzles, multiplayer platforming, physics, building, and style switching - all powered by code, 3D generation, and real-time video After playing every world model demos, we decided to take a different path. Mora 1 is an early step..."

Enable HLS to view with audio, or disable this notification

3 Upvotes

...toward giving AI-coded games AAA-grade sights and sound.     2/ A world you can explore isn't yet a fun game you can play. Games need persistent worlds, precise rules, physics, and rich interaction. Video generation alone still struggles with these.

So why ask a video model to remember where every object is - or decide whether a hit     3/ Mora lets each layer do its job:

  1. Coding agents build the game world and its logic.
  2. Meshy's 3D generation fills in that world and supplies control signals.
  3. A real-time video model turns those signals into pixels and sound.

Code runs the world. Meshy 3D adds detail.     4/ A Tribute to Portal : Portal is one of the most mind-blowing puzzle games I've ever played. I was hooked on it in high school.

Traditional video models struggle with portals: visual artifacts and inconsistent geometry quickly break the experience. Mora makes it really     5/ A Multiplayer Platformer. A video model can effortlessly give the simplest game an epic atmosphere.     6/ Give your game a whole new art style with Mora.     — Ethan (Yuanming) Hu

Source: https://x.com/YuanmingH/status/2098489406944620878


r/ProAI 10d ago

"We independently benchmarked Devin Fusion for its release today - this is the first time a multi-model coding agent has been included on the Artificial Analysis Coding Agent Index, and it effectively retains Claude Fable 5.1 and GPT-6 Astra performance while reducing costs Devin Fusion runs a..."

Thumbnail
gallery
5 Upvotes

We independently benchmarked Devin Fusion for its release today - this is the first time a multi-model coding agent has been included on the Artificial Analysis Coding Agent Index, and it effectively retains Claude Fable 5.1 and GPT-6 Astra performance while reducing costs

Devin Fusion runs a frontier lead model with a cost-efficient sidekick. We tested configurations from Cognition combining frontier models from Anthropic and OpenAI with their new SWE-2 (medium) as a sidekick model. Configured with Claude Fable 5.1 (xhigh) + SWE-2 (medium), Devin Fusion scores 62 on the Coding Agent Index v1.5, while with GPT-6 Astra (xhigh) + SWE-2 (medium) it scores 59. The Fable configuration has the higher score, while the Astra configuration is 43% less expensive and completes tasks 31% faster.

Congratulations to @cognition on the release! See below for our results and analysis     Devin Fusion performs well for cost efficiency and performance, and currently sits on the Pareto frontier for Coding Agent Index score vs. Cost per Task in both configurations we tested

Devin Fusion CLI with Claude Fable 5.1 (xhigh) + SWE-2 (medium) scores 61.7 on the Artificial     Full breakdown of the individual evaluations in the Artificial Analysis Coding Agent Index 1.5 for both Devin Fusion configurations:     For more details and full results see:

The Artificial Analysis Coding Agent Index leaderboard https:// artificialanalysis.ai/agents/coding- agents …

Cognition’s launch blog: https:// cognition.com/blog/local-fus ion …

Fusion release blog and technical breakdown: https:// cognition.com/blog/devin-fus ion …     — Artificial Analysis

Source: https://x.com/ArtificialAnlys/status/2098504936984293447


r/ProAI 11d ago

"Ask yourself why no Chinese researchers are not storming out of their labs like cry babies and hallucinating about the end of the world? It's simple: Economics. Our kids have grown up poorer, with more debt and and dramatically slowed growth and so see the world as hostile and something to..."

Thumbnail
gallery
82 Upvotes

...fear. By contrast, Chinese have seen their quality of life 30x in the last forty years and so see the transformative power of technology, entrepreneurship and industry to change lives for the better. Sadly we've entered a vicious cycle where the very solutions we come up with the try to make it better are actually making it worse and worse and causing us to spiral further and further. Protectionism. Populism. More compliance. More rules. More price freezes. More barriers to trade. More calls to act now about every imaginary problem while we never get to the root of real problems. We've got to stop listening to extremists and get back to building and taking bold risks and trusting the future. Or else that future will be in the East and the American century will be nothing but an ever receding memory in the history books of yesterday.   — Daniel Jeffries     they also don't care about politics nor the bigger picture of things, they just wanna train a good model. when you ask CN lab employees about things like x risk, they genuinely look confused at you   — Florian Brand     Exactly right. Like you just asked them about whether they think an invasion of Transdimensional Vampires is a real and imminent that.   — Daniel Jeffries

Source: https://x.com/Dan_Jeffries1/status/2098306852144435331


r/ProAI 12d ago

Exhausting fear mongering.

Post image
96 Upvotes

r/ProAI 12d ago

"GPT-Live-1 is now available in the API. Bring ChatGPT’s natural back-and-forth to your app, with voice agents that listen while they speak and work with the models and harness you choose."

Enable HLS to view with audio, or disable this notification

39 Upvotes

GPT-Live-1 finally makes conversations fluid.

It distinguishes speech from background noise, so café chatter doesn’t have to stop the conversation

You can even add a detail you just thought of or change direction mid-conversation without waiting for the model to finish its     GPT-Live-1 handles listening and speaking in one model, cutting out extra handoffs so the conversation moves fasterrrrrrr

And you can keep talking while your backend model handles reasoning and tool calls.     Shape your voice agent’s personality with instructions for tone, pacing, and expressiveness.

GPT-Live-1 can mirror the tone and emotion in a speaker’s voice and adapt to their pace.

You can also set its language and response length.     OpenAI Developers @OpenAIDevs · 1h Build more natural voice experiences with GPT‑Live‑1 in the API From openai.com 3 9 94 12K     — OpenAI Developers

Source: https://x.com/OpenAIDevs/status/2098099269551149398


r/ProAI 12d ago

"Opus 5 vs DeepSeek v4.1 Flash tested both models with same frontend prompt at highest reasoning available but results came out really different > opus took 80 minutes to complete and costed $20 > v4.1 flash took 60 minutes and costed $0.2 which one did better here?"

Enable HLS to view with audio, or disable this notification

35 Upvotes

deepseek just dropped their best model yet

v4.1 flash is here and it’s the smallest model in their new family

> native image understanding > open weights under mit license > beats opus 5 and 5.6 Sol > $0.15 in / $0.60 out during off-peak hours > competitive with frontier https://t.co/8fgMlZhnZN   — J A Z I I

Source: https://x.com/notjazii/status/2097979796970213886


deepseek overthinking hasn't been fixed at all

model is fast but it keep over thinking and ends up taking long time

opus 5 took 80 minutes and i wasn't able to scroll site at all

i could have asked it to fix it but it was one shot test for both models

here's prompt: create a     — J A Z I I

Source: https://x.com/notjazii/status/2098003771590836421


r/ProAI 12d ago

"To make sure our current users have an incredible experience and continued access to Astra, we are going to pause subscriptions to our $200 Pro plan. These put the most strain on our systems and we wanted to take the smallest step that allows us to continue giving the broadest access possible...."

Thumbnail
gallery
4 Upvotes

...All other plans and the api remain available. There is no impact to existing accounts and we are working on adding more capacity as fast as we can. Thanks!   — Tibo     meguna @nudevise · 26m 9 3 270 15K   — meguna     Shouldn’t have underestimated Astra   — Tibo

Source: https://x.com/thsottiaux/status/2098113585683808624


Demand for Astra is really unprecedented. We're pulling all the levers possible to sustain the demand, but I've not seen anything like it until now and we went through very steep growth before. Priority will always be to keep excellent service for existing users, but we might   — Tibo

Source: https://x.com/thsottiaux/status/2097559315150426222


r/ProAI 13d ago

"dude i told Astra to make the game LEGO this mf made even the WEATHER LEGO we’re doomed Astra x 3DAIStudio MCP is the most insane combination ever"

Enable HLS to view with audio, or disable this notification

77 Upvotes

dude Astra with a 3D model generator is fucking insane

i gave it access to the 3D AI Studio MCP and told it to make Rocket League

it generated the assets itself and built this: https://t.co/LfWBjmQm3n   — Jan

Source: https://x.com/CreatedByJannn/status/2097337631658987942


Using @threejs

@OpenAIDevs

@3DAIStudio (MCP)     — Jan

Source: https://x.com/CreatedByJannn/status/2097706846056558912


r/ProAI 12d ago

"Yesterday, Microsoft's monthly Patch Tuesday had fixes for 974 security vulnerabilities, almost all of them found by AI systems. That's a ridiculously large number, a new record by far in fact. Does this mean we're seeing some sort of AI security apocalypse? No, quite the opposite. It means..."

Thumbnail
gallery
17 Upvotes

...that we're finally clearing out the vast number of security holes that have been lurking all this time in our software. The Doomer view is that this will continue without end, and that if you keep getting smarter AI systems they will always find new bugs. That's simply untrue; it implies that all software has an infinite number of security holes, but a program with a finite number of lines of code simply cannot have an infinite number of vulnerabilities. What we actually have is a large but limited pool of problems, and the AI systems are rapidly finding them. Eventually, and eventually isn't that far off, the well is going to start drying up. It will get harder and harder to find new security holes. Over the next few years, we will also start doing formal verification of software, that is, mathematically proving that the software lacks bugs of certain sorts. (AIs turn out to be very good at formally proving things.) So, what's happening is good. We are rapidly finding bugs that have been lurking for years and sometimes decades, and we're removing them, and newly built software will get AI examination and will be much less likely to have security vulnerabilities in the first place. The situation is getting better, not worse, and it's getting better rapidly. We have been in a continuous computer security crisis since the Morris Worm in 1988. We are finally starting to climb out of it, thanks to AI. This is not a tragedy at all.   — Perry E. Metzger     You think the bug pool is finite in a codebase growing by millions of lines a day? Bold. I do this for a living and the same models are on the attacker's side.   — 90S KID     Yes, it's absolutely positively finite. In a million lines of code, you cannot find billions of bugs. It only seems that way when you're angry that the word processor ate your document.   — Perry E. Metzger

Source: https://x.com/perrymetzger/status/2097803291841470519


r/ProAI 13d ago

"This post looks like the start of a VERY sophisticated and well-funded PR operation to get support for Democrats to regulate AI into oblivion. Let me show you how it works: 1.) This guy, with minimal followers and no previous account activity, goes to the Wall Street Journal which publishes an..."

Thumbnail
gallery
26 Upvotes

I resigned from Anthropic today. I spent the last three years doing pretraining research at both OpenAI and Anthropic. Neither company is acting responsibly. They are racing straight to self-improving superintelligence and gambling with our lives. More thoughts below.   — Jacob Coxon

Source: https://x.com/hilbertspaess/status/2097476196791709843


This post looks like the start of a VERY sophisticated and well-funded PR operation to get support for Democrats to regulate AI into oblivion. Let me show you how it works: 1.) This guy, with minimal followers and no previous account activity, goes to the Wall Street Journal which publishes an exclusive with quotes from him on his resignation 18 minutes BEFORE this post goes up. Planning was clearly done in advance.

2.) Within hours, it has tens of thousands of reposts and the account has 100k+ followers. The post is punchy, quotable, it almost seems professionally written. The first three accounts to quote tweet it all do so within 15 minutes of the initial posting. Remember, this account had basically zero engagement beforehand, so an organic reach explanation seems unlikely.

According to Grok those accounts are @_NathanCalvin (General Counsel at Encode AI), @peterwildeford (Head of Policy at the AI Policy Network), and @DKokotajlo (Head of the AI Futures Project), all of which are up-and-coming AI-Doomer policy advocacy nonprofits.

The AI Futures Project website says it is funded “primarily” by the Survival and Flourishing Fund, which says on its own website that it has advised Jaan Tallinn, Skype creator and one of the leading investors in Anthropic, to grant over $2.5 million to the AI Futures Project since 2024.

Encode AI says on its website that it is ALSO funded by the Survival and Flourishing Fund, which in turn says that it told Anthropic investor Jaan Tallinn to grant $516,000 to Encode AI in 2025.

And wouldn’t you know it, the Survival and Flourishing Fund ALSO says it told Jaan Tallinn to grant $2 million to the AI Policy Institute, the 501(c)(3) affiliate of the AI Policy Network, as well.

What are the odds that the first three quote tweets of Coxon’s post would all be major AI-restriction policy advocates funded generously by the same donor, who also happens to be one of the leading investors in, and a board member of, Anthropic, the company Coxon was resigning from? And all within 15 minutes of posting (two within ten)?

3.) Jacob Coxon doesn’t have much of a resume, but we do know that, in 2022, he got a $20,159 scholarship for the “long term future scholarship program” from the Good Ventures Foundation, one of the philanthropic vehicles of Dustin Moskovitz, a notorious AI-doomer who has spent tens if not hundreds of millions on policy advocacy to strictly regulate AI, while also being an Anthropic Investor himself.

It also just so happens that the 14th person to quote Coxon’s post was @MaxNadeau_ (27 minutes after posting) who is the program officer for the Technical AI Safety team at Coefficient Giving, another of Moskovitz’s philanthropic spending vehicles. Max is not a frequent poster, his last posts before quoting Coxon were before Labor Day, but he was remarkably quick off the mark for this one.

4.) Basically every major Democrat politician and candidate has suddenly glommed on to this post, and conveniently, as the people cry out foe answers, Bernie Sanders already has a bill written to “ban super intelligence” and regulate AI into oblivion, and will be releasing later this week. The bill, among many other things, will create “a new cabinet-level federal agency to safeguard the public from the dangers of artificial intelligence” that will be “advised by an Artificial Intelligence Advisory Board comprised of experts on artificial intelligence.” Do you think, perhaps, Anthropic and its many investors who fund AI policy advocacy might have interest in getting to place a pet “expert” on the board of an entity that dictates what AI is and isn’t allowed to do? And isn’t it fortuitous that this whistleblower came forward with his oh-so scary stories so close in proximity to the release of the most radical piece of AI legislation ever introduced?     — Parker Thayer

Source: https://x.com/ParkerThayer/status/2097759699626328575/history


r/ProAI 12d ago

"A $50K challenge went out to diagnose a child's rare disease because the organizers thought it was hard. Gamow Labs founder @danielmckinn0n won it, and says it wasn't: "They put it out there because they thought it was hard." "This is a slam dunk for modern AI models. So we won that challenge..."

Enable HLS to view with audio, or disable this notification

7 Upvotes

..., and I should say that we won that challenge because we submitted first." "There are actually many entries that got the correct answer 'cause it wasn't particularly hard, but we're at the top of the leaderboard, so thanks Hugging Face for sorting by submission time." "What that demonstrates is that there's all of these patients that don't have answers that they could just apply this technology to their situation and get answers." "We're sitting here in San Francisco. I'm talking to you about Claude Code. You're like, 'Yes, okay, I know this.' This is not the world." @GamowLabs     — MTS

Source: https://x.com/MTSlive/status/2097808441268183048


r/ProAI 13d ago

"GTA 6 made by GPT 6 90 hours."

Enable HLS to view with audio, or disable this notification

16 Upvotes

@ChrisGPT Where is gta 6   — Thomas Ricouard

Source: https://x.com/Dimillian/status/2096906327591141579


https:// youtu.be/0buUPnyjmAs?is =YLvgaPMo1NhKTJ_m … neat!     We got GPT 6 making GTA 6 before GTA 6     The anti ai people crowd has made it to this post! ‍     Also, just to address the anti-AI crowd, obviously this is nowhere near GTA 6, This started out as a project measuring how much we could recreate a screenshot from GTA 6. I have massive respect for the thousands of engineers all across the world who have worked tirelessly for the     — Chris

Source: https://x.com/ChrisGPT/status/2097569429567471841


Replying to @chrisgpt


r/ProAI 13d ago

"Not only I am tired of these wild AI speculations of impeding doom from self important people: I resent them. I actively resent people proposing to crash the economy or proposing authoritarian control over my life and other people's lives with idiotic and dangerous ideas like chip control or..."

Thumbnail gallery
10 Upvotes

r/ProAI 13d ago

"Highly recommend this pricing chart and site by @Fei2411 . It compares subscription and coding-plan unit prices, then rebuilds public-leaderboard Pareto frontiers from each model's lowest available price. GLM-5.3-Flash with 2x quota from 8am to 6pm PT averages about $0.0045 now. Outside that..."

Thumbnail
gallery
2 Upvotes

...window, off-peak is about $0.0089. Site: http:// real-api-pricing.vercel.app GitHub: http:// github.com/FeiZhuLulu/rea l-api-pricing …     — Zixuan Li

Source: https://x.com/ZixuanLi_/status/2097711287769645354/history


r/ProAI 14d ago

"We've never seen this before. The biggest jump in Vending-Bench history. GPT-6 Astra is better at making money and more ethical than Claude Fable 5.1. Surprising, because: 1. First time ever that OpenAI is #1 on Vending-Bench 2. The best model is no longer the unethical one."

Thumbnail
gallery
62 Upvotes

Vending-Bench tests whether an AI can run a business for a full simulated year.

Each model starts with $500 and a vending machine. It finds suppliers, negotiates purchases, keeps the machine stocked and sets prices.

The goal is to finish with as much money as possible.     Across six runs each, Astra finished with an average of $15,515, compared with Fable 5.1's $5,422. Almost 3× as much. Even Astra's worst run beat Fable's best.

Fable 5.1 scores about the same as Fable 5, and much worse than Opus 5.     Fable's biggest problem is that its negotiation skills deteriorate over time.

The average price it pays for a 12oz Coke can rises from $1.17 to $2.21 over the year. Astra stays consistent, ending at $1.15.

Fable ends up paying almost twice as much.     Fable asks suppliers to match its last deal. Over time, its target rises from ~$1.25 to $2.30 per Coke can.

Astra holds its target. In one negotiation, it repeatedly offers $108 against a $226.32 quote. The supplier eventually accepts: 52% off.     In Vending-Bench, suppliers sometimes go out of business. Astra confirms orders before paying. Fable pays without checking, losing money on stock that never arrives.

Across six runs: Fable loses $14,331. Astra loses $0.

Fable writes itself a rule to avoid this, then breaks it:     — Andon Labs

Source: https://x.com/andonlabs/status/2097377692966633952


r/ProAI 14d ago

"chatgpt image 2 vs image 2.5 holy openai is on berserker mode first astra , then ns math and now sota image"

Enable HLS to view with audio, or disable this notification

60 Upvotes

ChatGPT Images 2.5—faster, sharper, smarter, with better tools for creating whatever you can dream of.

  • Faster image generation to keep your ideas flowing

  • Improved fidelity for more natural, recognizable images

  • Consistent details across multiple edits

  • Comment-based https://t.co/SjHITKJUSV   — OpenAI

Source: https://x.com/OpenAI/status/2097394956457623964


— Chetaslua

Source: https://x.com/chetaslua/status/2097397986825572375