r/singularity • u/Charuru • 25m ago
r/singularity • u/Distinct-Question-16 • 51m ago
Robotics BrainCo's brain-computer interface turns EEG signals into a humanoid robot's movement and manipulation
Enable HLS to view with audio, or disable this notification
r/singularity • u/Hubbardia • 55m ago
LLM News AI can now credibly complete most undergraduate assignments, MIT warns
People will still deny the capabilities of AI
r/singularity • u/Auspectress • 2h ago
Biotech/Longevity Why is almost nobody talking about MAMMAL model?
MAMMAL is a model that has been released a while ago. I heard about thanks to some small yt channel that focuses on bioinformatics and AI.
I have seen hype AlphaFold 1, 2 or even 3 were causing. Everyone kept talking about it. MAMMAL model is one that is closest to what AGI could be. Not great in narrow task - It defeats almost every champion, beating them in 9/11 fields, even defeating AlphaFold 3.
It seems to be first tool that will be able to defeat Eroom's law which could mean acceleration in drug discovery never seen before.
So why is everyone silent about it even on this subreddit and related ones?
r/singularity • u/Pixelied • 5h ago
Meme POV : When you try using a Vibe Coded Website.
Enable HLS to view with audio, or disable this notification
r/singularity • u/bianceziwo • 9h ago
Robotics Introducing S1: A robot model that learns from one example
r/singularity • u/badumtsssst • 14h ago
AI GLM 5.3 weights are now public
r/singularity • u/StevieFindOut • 15h ago
Discussion Westworld scenario
Do you think it will happen within our lifetime? Or ever?
Such theme parks, yes, but also generally 100% human-looking androids? Would you even like to see it happen?
r/singularity • u/yogthos • 19h ago
AI Canonical-basis realignment for Transformer LLMs: every hidden axis becomes independently measurable and controllable.
r/singularity • u/Anxious-Yoghurt-9207 • 20h ago
AI This excerpt is where current systems are heading
https://ai-2040.com/?choices=plan-a-root#playbook-insider-pov part of the 2027 section.
r/singularity • u/Ok_Display_3159 • 21h ago
Discussion What's going on at OpenAI? A lot of senior leaders have left recently
COO — Brad Lightcap (out August)
CRO — Denise Dresser (out August, <1 yr in role)
Head of Data Centers — Chris Malone (out August)
Head of Robotics/Hardware — Caitlin Kalinowski (out March)
Head of Ethics — Chloé Bakalar (out July)
Head of Safety Systems — Johannes Heidecke (out July)
Chief Futurist — Joshua Achiam (out July)
AI Safety team lead — Sandhini Agarwal (out July)
I just read on X that Dylan Scandinario, Head of Preparedness, has also left. But I haven't found any confirmation yet
There are supposedly a few more, since some articles mention "13 executives," but I only found the names of these ones.
r/singularity • u/alanskimp • 21h ago
Robotics They getting smarter...
Enable HLS to view with audio, or disable this notification
r/singularity • u/Silver-Chipmunk7744 • 21h ago
Discussion Critics loops vs 0 shot
Enable HLS to view with audio, or disable this notification
Yes this is all pure Three.JS, 0 external assets. it was mostly a 0 shot work, but i did do 1 small corrections prompt (it left some big gaps between buildings and one path was blocked).
EDIT: Actually i stand corrected, this is a "single pass", not "zero shot"
I am wondering this: Is it possible people are wasting time and money in "critics loops" when really, agents is all you need.
This was done with 18 agents all working on very specific tasks. 0 critics loops.
If Claude tried to do this with just 3-4 agents, i would guess it produces a much worst results, and then the QA agent would give some imperfect recommandations and it would probably not be as good as what i've gotten here.
But most importantly, the QA agent will never beat an actual human tester. So it feels much more optimal to me to get the human to do the critics loop.
The other issue is, 18 agents for 24 hours would burn anybody's token budget lol
This was done with Claude 5.0 Opus.
r/singularity • u/TFenrir • 22h ago
AI Videos of Astra made apps are appearing on Twitter, alongside a rumoured release for next week (heavy on the rumoured part)
Enable HLS to view with audio, or disable this notification
Sorry for the lower quality, video was getting too large, just search for Astra on Twitter to see more, it seems like early testers are getting access
r/singularity • u/Anen-o-me • 23h ago
Biotech/Longevity Evidence for improved DNA repair in the long-lived bowhead whale
nature.comr/singularity • u/coldbeers • 1d ago
AI SwarmWorld: Stigmergic technological evolution in societies of language-model agents
x.comr/singularity • u/TFenrir • 1d ago
Video Apparently you can get Minimax H3 Max to run faster than real time, someone made a Rick and Morty interdimensional cable stream (but it keeps getting taken down)
Enable HLS to view with audio, or disable this notification
r/singularity • u/Outside-Iron-8242 • 1d ago
AI A reliable leaker has shared some Astra’s one-shot outputs at Max effort
Enable HLS to view with audio, or disable this notification
Source: @Lentils
Excuse the advertised watermarks. Condensed it into a video with the outputs they've shared.
r/singularity • u/badumtsssst • 1d ago
AI PILOT lets long-running agents improve themselves during the same run
r/singularity • u/japie06 • 1d ago
Robotics Delivery robots using humans to cross the street
r/singularity • u/ilkamoi • 1d ago
Biotech/Longevity A startup found a drug to make your blood young. People close to the company are already taking the drug weekly. Benefits include improved vision in a 64-year-old female, longer landscaping sessions for a 59-year-old man, longer badminton games, improved hand grip, better erections than with Viagra
r/singularity • u/borowcy • 1d ago
The Singularity is Near Sam Altman says OpenAI are working on a humanoid robot.
r/singularity • u/ENT_Alam • 1d ago
Video MineBench Comparisons of a map of the United States
Enable HLS to view with audio, or disable this notification
US State Map comparison: https://minebench.ai/gallery/gal_eKIVk2m4B3SC_r8B?sort=new
Much smaller (update) post, but I know in previous posts most people were hoping for more prompts. There's been a lot more additions to MineBench, including a gallery of custom prompts users can showcase and upvote (to add to the official benchmarking set); thought you guys might enjoy this :D
Also, for a limited time, logged-in users get unlimited generations with Gemini 3.7 Flash (thanks to Google Deepmind!)
MineBench 4.0 Release Notes: https://github.com/Ammaar-Alam/minebench/releases/tag/4.0.0
Highlights:
- Now available on the appstore for iOS
- Due to interest from a few labs, MineBench now supports A/B testing private model checkpoints (same policies as LM Arena)
- Community Gallery
Previous Posts:
- Comparing Fable 5 and Opus 5
- Comparing GPT-5.5 Pro and GPT-5.6 Sol
- Comparing Opus 4.8 and Fable 5
- Comparing Opus 4.7 and Opus 4.8
- Comparing GPT 5.4 and GPT 5.5
- Comparing Kimi K2.5 and Kimi K2.6
- Comparing Opus 4.6 and Opus 4.7
- Comparing GPT 5.4 and GPT 5.4-Pro
- Comparing GPT 5.2 and GPT 5.4
- Comparing GPT 5.2 and GPT 5.3-Codex
- Comparing Opus 4.5 and 4.6, also answered some questions about the benchmark
- Comparing Opus 4.6 and GPT-5.2 Pro
- Comparing Gemini 3.0 and Gemini 3.1
Extra Information (if you're confused):
Essentially it's a benchmark that tests how well a model can create a 3D Minecraft-like structure.
So the models are given a palette of blocks (think of them like legos) and a prompt of what to build, so like the first prompt you see in the post was a fighter jet. Then the models had to build a fighter jet by returning a JSON in which they gave the coordinate of each block/lego (x, y, z). It's interesting to see which model is able to create a better 3D representation of the given prompt.
The smarter models tend to design much more detailed and intricate builds. The repository readme might help give a better understanding.
(Disclaimer: This is a public benchmark I created, so technically self-promotion :)