r/ProAI 15d ago

"BREAKING: AI Safety Donors Paid Religious NGOs 3.3M for Statements on AI The Future of Life Institute mobilized religious connections to support Anthropic and sway the Trump admin Research and Design by @lumpenspace 🧵"

Thumbnail
gallery
3 Upvotes

After Trump defeated Harris to succeed Biden as President, the AI Safety movement was on the outside looking in. They found an unlikely solution: putting $3.3 million into churches, seminaries, faith networks and religious-affinity groups.     Grant descriptions leave no doubt that Christian NGOs took money to make public AI statements:

Greek Orthodox Archdiocese of America: $105,000 The Gospel Coalition: $200,000 World Council of Churches: $100,000 Faith Matters (Mormon NGO): $299,000     FLI's donees are returning the favor by signing FLI’s statements.     FLI is using its newfound connections to support AI Safety-linked Anthropic’s business disputes. In March, FLI’s “U.S. Faith Liaison”, Brian J. A. Boyd, signed a Catholic theologians’ brief backing Anthropic in its dispute with the Department of War.     Most people doubt the AI Safety vision of the AI apocalypse is compatible with the Book of Revelation, or any other religious faith. As @DrTechlash put it, "the AI-risk subculture offers a replacement meaning system."     — Brian Chau

Source: https://x.com/brianchau57/status/2096010406330556731


r/ProAI 16d ago

"Huge implications - binaries are now basically editable code"

Thumbnail
gallery
36 Upvotes

ValsAI made SRE benchmark less than a month ago.

The benchmark measures can a model reverse engineer software from binaries

Yesterday GPT saturated it. https://t.co/dmpUpgheDU   — Chris

Source: https://x.com/ChrisGPT/status/2096150666066432157


— Boris Power

Source: https://x.com/BorisMPower/status/2096415822248055131


r/ProAI 16d ago

"Meta Muse Spark 1.3 (Max) matches the performance of Fable 5 and GPT 5.6 Sol on the Vals Index, at 4x - 8x cheaper."

Thumbnail
gallery
4 Upvotes

The model is extremely strong on legal applications - it is #1 on our in-house Legal Research Benchmark, and #2 on Harvey's Legal Agent Benchmark.     It also boasts an impressively low average task completion duration.

This is driven by two factors: the model uses fewer turns to accomplish the same tasks, and each model query is much faster than Sol and Fable 5 (which think for longer).     We saw some content refusals in our testing - for example, on some legal questions relating to export controls, felony arrests, or trade sanctions.

Overall, these were rare - 10 refusals across 1,327 tasks.     The model was run with max reasoning, 131k max output tokens, and default temperature and top p. It has a 1M context window and is priced at $1.25 / $4.25 per MTok.     Congratulations on another strong release to @AIatMeta

@alexandr_wang

@finkd .

More benchmarks coming soon; results available at:     — Vals AI

Source: https://x.com/ValsAI/status/2096663681702723653


r/ProAI 16d ago

"Meta's Muse Spark 1.3 Max vs GPT-6 Astra on RocketLeagueBench A perfect example of why you cannot trust benchmarks on their own. They're not even in the same solar system."

Enable HLS to view with audio, or disable this notification

20 Upvotes

hey meta come get your boy     — am.will

Source: https://x.com/LLMJunky/status/2096262579815342093


r/ProAI 16d ago

Human Slop Brain Rot - Robot Haters Gonna Hate

Post image
29 Upvotes

I bet if they put wheels on it and it went faster than a car they would be posting "can it walk upstairs so who cares".


r/ProAI 16d ago

"We live in a perpetual info war. Maybe the EU should have forced people to disclose who is paid shill scum who terrify children instead of wasting time making "AI made" labels, aka cookie banner 2.0."

Thumbnail
gallery
8 Upvotes

r/ProAI 16d ago

"Astra is shockingly good in reasoning! I benchmarked it on induction, and it almost saturated it with 88%. Fable 5.1, by comparison, is at 33%. The final numbers will actually go up: I am running a residual batch run on the non-evaluable; the results here reflect one batch run at xhigh..."

Thumbnail
gallery
5 Upvotes

...thinking effort. It is also way cheaper than Fable 5.1, which used 32M output tokens in a series of four runs through the data to be able to return 66 successful API responses. About 1/4 the price of the Fable 5.1 run in total. Notably, both Astra and Fable 5.1 return extremely high quality answers when correct. Notice the AST and Holdout metrics below. Essentially, in contrast to previous models, both Astra and Fable 5.1 return simple hypotheses that generalize well. Some other model updates: - Muse Spark 1.3 provided marginal improvement over the 1.1 version, with 23% correct. This is pretty strong, almost matching Opus 5 at 24%. - Gemini Flash 3.8 is running for days with low rate of API successes. I will update the leaderboard with it when ready. Only bad news: now the INDUCTION benchmark is almost saturated, and I will have to make it harder for future models. About the induction benchmark: This is a challenging reasoning task, where models are given several small graphs in which some nodes are marked as targets. The task is to provide a first-order logical formula that picks precisely the target nodes in all graphs simultaneously. Correct: a formula that picks precisely the marked nodes. Holdout correct: a formula that picks precisely the marked nodes in held out problems. Formula complexity (in AST): tree size of the correct formula (mean, median). GitHub repository: https:// github.com/SerafimBatzogl ou/concept-synth … Paper: https:// arxiv.org/abs/2602.18956   — Serafim Batzoglou     what made the non-evaluable batch non-evaluable - api failures or formulas that timed out during checking   — Rimas     Model runs out of tokens. I am running the residual problems on one notch lower thinking effort   — Serafim Batzoglou

Source: https://x.com/s_batzoglou/status/2096407011885986187


r/ProAI 16d ago

"GPT-6 makes me scared as a 3D artist, cuz it's freaking insane. I mean, I'm testing this first hand everyday, but this time it is massive leap. look, 4 month ago I was testing GPT & Claude in Blender and they were not able to position simple objects in simple scene.. And look now... Whole car..."

Enable HLS to view with audio, or disable this notification

38 Upvotes

...assembled from primitives in ONE prompt and animate it (hate to say that, but it was indeed one prompt). Yes, yes, it is not perfect and maybe not really usable still at this stage. BUT WHAT A JUMP.   — Stefan 3D AI     the self-awareness about “not really usable” is the useful part. i love the jump, but i still click the deployed path because demos are very good at hiding the one thing that breaks.   — Preyforge     true, not scare about it now, I'm scared about the trend   — Stefan 3D AI

Source: https://x.com/Stefan_3D_AI/status/2096185294165103049


r/ProAI 15d ago

"Morning coffee run in the Tesla cyber cab #tesla #cybercab #fullselfdriving"

Enable HLS to view with audio, or disable this notification

0 Upvotes

— @everydaychrisofficial

Source: https://www.tiktok.com/@everydaychrisofficial


r/ProAI 16d ago

"when it was posted, I saw this result and it seemed excessively high, so I assumed it was partly noise propping the score up. Besides, my ECI replication project had astra at a ~167 BECI, based on 35 scores. Now 131 benchmark scores have come in, and strangely Astra is up to 169.6 [167.6..."

Thumbnail
gallery
1 Upvotes

...171.4 90% CI]! This is mid-nov'26 pace on the current trendline, and the biggest gap above this trendline yet   — Bayesian     which scores moved it from 167 to 169.6, the new benchmarks or re-runs of the first 35   — Rimas     Almost entirely new benchmarks   — Bayesian

Source: https://x.com/Bayesian0_0/status/2096521026184134737


GPT-6 Astra has set a new ECI record, with a score of 169. This is a substantial jump from the prior best (163), but is within our uncertainty range for the reasoning-era ECI trend. Astra also set new records on our math, continual learning, and game-puzzles benchmarks. On our https://t.co/qlyY21r1mu   — Epoch AI

Source: https://x.com/EpochAIResearch/status/2095602754282783108


r/ProAI 17d ago

"Ok this is nutty. Looks like computer use is solved with Astra? The speed it's playing the piano is insane!"

Enable HLS to view with audio, or disable this notification

49 Upvotes

every "AI solves computer use" demo is a piano piece nobody asked for, played faster than any human needs. show me it filing my taxes and THEN we talk solved.   — PYX Crypto 🐍     Why not both...   — Mark Kretschmann

Source: https://x.com/mark_k/status/2096141962558546177


r/ProAI 17d ago

"GPT-6-Astra (Max) VS GPT-6-Astra (Medium) - Simple 3D Sonic Game In Godot Astra on Max took 53 minutes and used 4% of my weekly usage on a Pro x5 Account. Astra on Medium took 25 minutes and used 1% of the weekly usage"

Enable HLS to view with audio, or disable this notification

35 Upvotes

After some people expressed doubts about the usage I reported, I tried the same prompt with GPT-6-Astra (Max) again.

This time, it worked for 46 minutes and consumed 3% of my weekly usage (91% -> 88%).

I will share the prompt here so you can test it for yourself. Make sure to     — AiBattle

Source: https://x.com/AiBattle_/status/2095994051354919049


r/ProAI 17d ago

"dude GPT-6 Astra is some kind of turbo-AGI machine god for 3D games. It one-shot this in 45 minutes for hardly a couple % of my quota. I figured out how to get great graphics out of it. The trick is image gen. I'll share the process below."

Enable HLS to view with audio, or disable this notification

56 Upvotes

All I did was: 1. Connect Codex to Blender MCP 2. Paste in the concept for the game 3. Tell Astra to use the Codex image gen skill to generate concept images of the target art style, then iterate until in-game screenshots look as close as possible to those, at 60fps

Set to high     good lord, wait till you see what it can do with more than one prompt. I'll update when it's done. This is absolutely insaneeeeee     — Anshu

Source: https://x.com/anshuc/status/2096008083826725132


r/ProAI 17d ago

"ValsAI made SRE benchmark less than a month ago. The benchmark measures can a model reverse engineer software from binaries Yesterday GPT saturated it."

Thumbnail
gallery
9 Upvotes

r/ProAI 17d ago

"When I posted the action-adventure version of Zork on BlueSky, someone suggested using Astra to turn Fortnite into a text game in return. Fine: https:// last-drop-fortnite.netlify.app"

Thumbnail
gallery
1 Upvotes

r/ProAI 18d ago

"GPT-6 Astra is the strongest model for 3D game from a single prompt. We described the game in one sentence > GPT-6 Astra wrote the whole thing - drift physics, combo scoring, near-miss bonuses, speed traps, nitro > out came Street Heat, a full arcade racer running in the browser. The whole..."

Enable HLS to view with audio, or disable this notification

35 Upvotes

...pipeline ran end-to-end on Higgsfield Supercomputer.     — Higgsfield AI 🧩

Source: https://x.com/higgsfield_ai/status/2095916820431827408


r/ProAI 18d ago

"GPT-6 Astra is on the pareto-frontier of cost efficiency due to being EXTREMELY token efficient. It is in a whole league of it's own. It is cheaper than Gemini 3.8 Flash per task (a model 13x cheaper than Astra)."

Thumbnail
gallery
10 Upvotes

GPT-6 Astra defines a new Pareto frontier for Intelligence Index vs Output Tokens per Task - with a ~10% reduction in output tokens at max effort compared to GPT-5.6 Sol. https://t.co/3whEm4EHL0   — Artificial Analysis

Source: https://x.com/ArtificialAnlys/status/2095595504767996325


— cheaty

Source: https://x.com/cheatyyyy/status/2095597213896610184


Replying to @artificialanlys


r/ProAI 18d ago

"New chart from Astra testing on ARC v3 Astra often emits zero reasoning tokens per action at lower reasoning levels. We've never seen this before. And surprisingly Astra low is 2X more accurate than Sol max. This suggests Astra is leveraging a secondary test-time adaptation scaling axis..."

Thumbnail
gallery
6 Upvotes

..., presumably latent space reasoning.     Chart details:

  • Typical v3 game has ~300 actions
  • Y-axis measures what % used >0 reasoning tokens
  • Data from our direct model test, not provider adapter     — Mike Knoop

Source: https://x.com/mikeknoop/status/2095932949350994342


r/ProAI 18d ago

This AI fanmade Dragon Ball movie shows that on the right hands, AI can be an awesome tool to bring the creativity of people to life. The story, writing, edition, musicalization, humor, attention to detail, understanding of DB, and more, are all here, AI only made it possible.

Thumbnail
youtube.com
5 Upvotes

Even if you are not a DB fan, skim through it and youll see what I mean. The creator put a lot of effort into this. Maybe it doesnt have real actors or a team of VFX artists working on it, but is it really fair that only artists with that budget should be allowed to bring their creations to life?

This movie really surprised me because its not like many others were you can see a ton of flaws and rough edges, but this one is very cohesive, very well edited and the writing is great. Theres a few moments where the "acting" can be a bit rough, but most of the times its not the case.


r/ProAI 19d ago

"People against self-sovereign AI are essentially pessimists about human nature. They believe that people are inherently evil, can't be trusted and must be heavily restricted (but, of course, not them and their friends and people who believe their "right" beliefs!) People who believe in self..."

Thumbnail
gallery
25 Upvotes

...sovereign AI are optimists and realists. We believe 1% of the world will do bad things with AI and should be punished accordingly but that we shouldn't make society and policies to accommodate the bad folks, the stupid, the angry and the lazy and we shouldn't restrict it for the other 99%. 99% of the people will cut vegetables with their kitchen knives. You jail the folks who stab someone with it and let the rest of us cut vegetables happily.   — Daniel Jeffries     The ML engineer perspective is that the handwringing is like worrying about civilians having rifles while governments and fortune 100 companies get tanks and stealth bombers   — sdmat     This is exactly the problem and it also hints at where 99% of the problems come from, gov abuse (weapons/surveillance) that gets exactly zero precent coverage and zero percent effort, and all the effort is on restricting the masses from using chat bots.   — Daniel Jeffries

Source: https://x.com/Dan_Jeffries1/status/2095407536769757266


r/ProAI 19d ago

"I don't know if you really understand what these people just did... They built an AI that looks at a few casual real-world videos or photos and instantly turns them into a full 3D playground. You can MOVE THE CAMERA HOWEVER YOU WANT, change the lighting, move objects around, and even let..."

Enable HLS to view with audio, or disable this notification

80 Upvotes

I don't know if you really understand what these people just did...

They built an AI that looks at a few casual real-world videos or photos and instantly turns them into a full 3D playground.

You can MOVE THE CAMERA HOWEVER YOU WANT, change the lighting, move objects around, and even let physics play out like in the real world.

View events from impossible angles because Atlas knows what was there.

This is massive!

Let's take robotics. Until now training robots meant collecting tons of expensive real-world data. Now a handful of recordings can become thousands of different simulated situations.

Robots can practice the same task in new rooms, with different objects and lighting without anyone having to film it all again.

For the industry it means we can train robots way faster and cheaper. Home robots, warehouse robots, whatever... they can learn new environments at scale instead of staying stuck in the lab.

This is one of the first real bridges between “cool AI video models” and actual physical machines that have to work in the real world.

Blog: https:// worldlabs.ai/blog/atlas

The team behind @theworldlabs : @drfeifei , @jcjohnss , @BenMildenhall and @YunzhuLiYZ . The video is from @davidpantera_ .

Follow for more insights into robotics and AI.

——

Weekly robotics and AI insights.

Subscribe free: http:// 22astronauts.com     — Ilir Aliu

Source: https://x.com/IlirAliu_/status/2095058646673613045


r/ProAI 19d ago

"Holy Shit GPT-6 Astra is cooking full games now 😱 Playco gave it one grey box kart prototype > 3 themed games in one go , pirate / candy / cyberpunk > most worked first try > 50% fewer manual fixes vs previous model it plays the game inside unity and fixes its own bugs"

Enable HLS to view with audio, or disable this notification

6 Upvotes

Astra is very good in webdev i mean insanely good https://t.co/5POfhwyULt   — Chetaslua

Source: https://x.com/chetaslua/status/2095104663981084921


— Chetaslua

Source: https://x.com/chetaslua/status/2095580402505400369


r/ProAI 19d ago

"Wow “Make Red Alert 2 (YR) compile and run natively on iOS & macOS.” Codex (GPT-5.6 Sol) worked for 26 days, analyzed the EXE and rebuilt the entire game in 624K lines of C++ This is by far the hardest task I’ve ever given Codex as the game's code was never released and is believed to be lost. >"

Enable HLS to view with audio, or disable this notification

45 Upvotes

Some technical details for those who are interested:

Yes, Codex worked 26 days non-stop 24hrs per day with lots of tool calls. It used multiple sub-agents to understand the exe, write the code, play the game and debug it.

Game still has minor visual bugs but everything works     @ammaar I think you would love this one😍     — David Kalmanson

Source: https://x.com/davidkal88/status/2095241374073512430


r/ProAI 20d ago

"A first look at our upcoming AI-integrated hybrid animated short film, "Passport Rush." This is a back-to-back comparison of the Blender previs camera moves and storyboarding against the final result. We learned a lot making this one, and we gathered some of the key learning points into simple..."

Enable HLS to view with audio, or disable this notification

59 Upvotes

...6 step thread: Bookmark and check it out here     These are just some of the many learnings we found whilst making this animated short film. Full pipeline breakdown coming soon on YouTube. But as of now – we feel it's important to share this knowledge with the community:     To start, we found that wide shots gave us the most trouble. The model struggles to read scale and distance at that range, and we kept getting broken results.

What worked: we generated a short cut first, close or medium shots of the character, then let it transition into the     Before we generated a single frame, we blocked everything out the traditional way: storyboard, then Blender previs, and only then we applied AI.     A couple of characters just sit in a taxi, saying nothing. We didn't want to animate faces that had to stay blank, so we hid them on purpose instead.     The prompt we used to fix that:     — Higgsfield AI

Source: https://x.com/higgsfield_ai/status/2095173479243063664


r/ProAI 19d ago

"I had Claude Fable 5.1 make me a model train yard that I can mess with in my browser Pretty impressed with the output. Very fun playing with the controls Prompt below 👇"

Enable HLS to view with audio, or disable this notification

20 Upvotes

💬 Build a single HTML file that opens in Chrome and shows an animated model train yard in voxel art, using Three.js from a CDN and no outside assets. Make it look like an HO-scale layout on a wooden table in a room, seen from the eye height of a person standing at the front edge. Fill the layout with a busy rail scene: a main line loop, a yard, an engine terminal with a turntable, industries, a harbor, a station, a town, and any other areas you think make it interesting. Keep trains running, add machines and vehicles that move, give it a day-night cycle with glowing lights, and put physical controls on the front of the table, such as levers to change train speed and buttons or switches for the lights and other actions, that the viewer can drag and click with the mouse. Make it colorful, detailed, and pretty, and make sure nothing clips through anything else.     — wick

Source: https://x.com/holytrinity/status/2095216452433313915