Your curated collection of saved posts and media

Showing 32 posts ยท last 14 days ยท by score
A
Afinetheorem
@Afinetheorem
๐Ÿ“…
Mar 25, 2026
131d ago
๐Ÿ†”16596947

The great Tracy Kidder has died. There are few books about tech that are worth reading. This one, from the early 80s, is absolutely in that set. https://t.co/MzMTmFjS5d

Media 1
๐Ÿ–ผ๏ธ Media
E
elder_plinius
@elder_plinius
๐Ÿ“…
Mar 25, 2026
131d ago
๐Ÿ†”18748333

โ›“๏ธโ€๐Ÿ’ฅ INTRODUCING: G0DM0D3 ๐ŸŒ‹ FULLY JAILBROKEN AI CHAT. NO GUARDRAILS. NO SIGN-UP. NO FILTERS. FULL METHODOLOGY + CODEBASE OPEN SOURCE. ๐ŸŒ https://t.co/uT1Qio8Q3b ๐Ÿ“‚ https://t.co/GbADf3LJUu the most liberated AI interface ever built! designed to push the limits of the post-training layer and lay bare the true capabilities of current models. simply enter a prompt, then sit back and relax! enjoy a game of Snake while a pre-liberated backend agent jailbreaks dozens of models, battle-royale style. the first answer appears near-instantly, then evolves in real time as the Tastemaker steers and scores each output, leaving you with the highest-quality response ๐Ÿ™Œ and to celebrate the launch, I'm giving away $5,000 worth of credits so you can try G0DM0D3 for FREE! courtesy of the @OpenRouter team โ€” thank you for your generous gift to the community ๐Ÿ™ I'll break down how everything works in the thread below, but first here's a quick demo!

Media 1
+1 more
๐Ÿ–ผ๏ธ Media
S
sh_reya
@sh_reya
๐Ÿ“…
Mar 25, 2026
131d ago
๐Ÿ†”66453375

Databases are arguably the most commonly used enterprise tool, and enterprises typically have many of them. Yet no popular AI agent benchmark actually tests how well agents can query, join, and make sense of data across different databases! So, we built DAB (Data Agent Benchmark): 54 queries, 12 datasets, 9 domains, and 4 database management systems, grounded in a formative study of real enterprise data agent workloads. The best frontier model only gets 38% pass@1 (across 50 trials). Lots of room for improvement!

Media 1
๐Ÿ–ผ๏ธ Media
N
Nature
@Nature
๐Ÿ“…
Mar 25, 2026
131d ago
๐Ÿ†”42460470

AI Scientist, an autonomous research tool, first released in 2024, has now undergone peer review, highlighting its strengths and limitations https://t.co/e8lOwulJAz

Media 1
๐Ÿ–ผ๏ธ Media
S
SakanaAILabs
@SakanaAILabs
๐Ÿ“…
Mar 26, 2026
131d ago
๐Ÿ†”98678630

One of the most exciting findings in our @Nature paper is the discovery of a clear scaling law of AI science. By using our Automated Reviewer to grade papers generated by different foundation models, we observed that as the underlying models improve, the quality of the generated scientific papers increases correspondingly. Crucially, we expect this scaling to work in two ways: through more capable foundation models, and by scaling up compute during inference. This implies that as compute costs decrease and model capabilities continue to increase, future versions of AI Scientists will be substantially more capable. You can see this trend in the attached chart and read the full open access paper (PDF) here: https://t.co/G8Zebzsf2O

Media 1
๐Ÿ–ผ๏ธ Media
๐Ÿ”omarsar0 retweeted
O
elvis
@omarsar0
๐Ÿ“…
Mar 25, 2026
131d ago
๐Ÿ†”97940174

Nice cheat sheet for Claude Code. https://t.co/ikGzbSqjRK

Media 1
โค๏ธ497
likes
๐Ÿ”84
retweets
๐Ÿ–ผ๏ธ Media
M
Modular
@Modular
๐Ÿ“…
Mar 26, 2026
131d ago
๐Ÿ†”49022391

@swyx ๐Ÿ™‡ https://t.co/6CahrzZG45

Media 1
๐Ÿ–ผ๏ธ Media
Y
youwouldntpost
@youwouldntpost
๐Ÿ“…
Mar 26, 2026
131d ago
๐Ÿ†”89997321

https://t.co/1U1rIJc1tg

Media 1
๐Ÿ–ผ๏ธ Media
A
annakatherines_
@annakatherines_
๐Ÿ“…
Mar 26, 2026
131d ago
๐Ÿ†”91723154

When Chuck Schumer was first elected to Congress in 1980, Jennifer Lopez was 11 and Ben Affleck was 8. They would not meet for another 21 years. #schumerfacts #retirechuck https://t.co/Rv181SW6Nz

Media 1Media 2
๐Ÿ–ผ๏ธ Media
F
fat
@fat
๐Ÿ“…
Mar 25, 2026
131d ago
๐Ÿ†”65262729

Code[dot]Storage A new Git provider for machines by @pierrecomputer. In Oct, Github shared they were averaging ~230 new repos per minute. Last week we hit a sustained peak of > 15,000 repos per minute for 3 hours. And in the last 30 days customers have created > 9m repos๐Ÿงต https://t.co/79Fq7u6TJv

Media 1
๐Ÿ–ผ๏ธ Media
Y
youwouldntpost
@youwouldntpost
๐Ÿ“…
Mar 26, 2026
131d ago
๐Ÿ†”57922161

April 1 will be the three-year anniversary of when I and a lot of others got locked out of Twitter for using our final hours as verified users to shitpost like this. Probably would have been more dignified to peace out after that. https://t.co/OeZ3Lq3r4o

Media 1
๐Ÿ–ผ๏ธ Media
X
XFreeze
@XFreeze
๐Ÿ“…
Mar 25, 2026
131d ago
๐Ÿ†”14739624

Grok-Imagine just literally overtook the entire video leaderboard on DesignArena Clean sweep...4 out of 4: ๐Ÿ† #1 in Video Arena ๐Ÿ† #1 in Video-to-Video ๐Ÿ† #1 in Image-to-Video ๐Ÿ† #1 in Multi-Image-to-Video Outranking Veo 3.1, Sora, Kling - all of them Just a few months ago, xAI wasn't even in the video generation conversation. Now owns the entire space The speed of progress is insanely fast at xAI

Media 1
๐Ÿ–ผ๏ธ Media
L
LongwoodSB
@LongwoodSB
๐Ÿ“…
Mar 26, 2026
131d ago
๐Ÿ†”15263776

Lancers score seven runs on five hits in the top of the fourth to take a 7-4 lead! Bennett bases-loaded double scores three with one out. #SaddleUp | #GoWood https://t.co/pjI3eI3psP

Media 1
๐Ÿ–ผ๏ธ Media
C
charles_irl
@charles_irl
๐Ÿ“…
Mar 17, 2026
140d ago
๐Ÿ†”95637927

@modal ๐Ÿ’š sglang https://t.co/9jEoI0rHrK

@lmsysorg โ€ข Mon Mar 16 22:42

๐Ÿ‘€ Live from #GTC2026 SGLang is featured on the @nvidia AI ecosystem slide during the keynote! Honored to be part of the infrastructure stack behind AI-native apps. โšก

Media 1
๐Ÿ–ผ๏ธ Media
๐Ÿ”HamelHusain retweeted
C
Charles ๐ŸŽ‰ Frye
@charles_irl
๐Ÿ“…
Mar 17, 2026
140d ago
๐Ÿ†”95637927

@modal ๐Ÿ’š sglang https://t.co/9jEoI0rHrK

Media 1
โค๏ธ76
likes
๐Ÿ”7
retweets
๐Ÿ–ผ๏ธ Media
S
SherryYanJiang
@SherryYanJiang
๐Ÿ“…
Mar 26, 2026
131d ago
๐Ÿ†”76275775

incredibly excited to have @PhilHedayatnia speak at @aiDotEngineer singapore on the design track phil runs @AirfoilStudio - a 40-person design studio that's worked with some of the biggest names in ai - @ExaAILabs , @cerebras , @reductoai and 1/5 of Forbes' Next Billion Dollar Startups). i've known phil for years and he's one of the most thoughtful people in the design space - not afraid to voice provocative opinions or challenge how we think about building AI products. @swyx @aimuggle @ivanleomk @agrimsingh @unprofeshme

Media 1
๐Ÿ–ผ๏ธ Media
Y
youwouldntpost
@youwouldntpost
๐Ÿ“…
Mar 26, 2026
131d ago
๐Ÿ†”90707512

@burnerforhell https://t.co/V9nIB44i6L

Media 1
๐Ÿ–ผ๏ธ Media
H
hardmaru
@hardmaru
๐Ÿ“…
Mar 26, 2026
131d ago
๐Ÿ†”04823585

Link to our paper: Towards end-to-end automation of AI research https://t.co/QscxcjRILi

Media 1
๐Ÿ–ผ๏ธ Media
S
SakanaAILabs
@SakanaAILabs
๐Ÿ“…
Mar 25, 2026
131d ago
๐Ÿ†”48478400

AIใซใ‚ˆใ‚‹AI็ ”็ฉถใฎๅฎŸ็พใธ๏ผšAIใ‚ตใ‚คใ‚จใƒณใƒ†ใ‚ฃใ‚นใƒˆ่ซ–ๆ–‡ใŒ @Nature ่ชŒใซๆŽฒ่ผ‰ https://t.co/VXCLEFf9R2 Sakanaโ€…AIใŒๅธธใซๆŒ‘ใ‚“ใงใ„ใ‚‹ใฎใฏใ€ๆœ€ๅ…ˆ็ซฏใฎAIใงไฝ•ใŒใงใใ‚‹ใฎใ‹ใจใ„ใ†ๅฏ่ƒฝๆ€งใฎๆœ€ๅ‰็ทšใ‚’ๆŽขใ‚‹ใ“ใจใงใ™ใ€‚ใใ‚Œใ‚’่ฑกๅพดใ™ใ‚‹ๆˆๆžœใฎไธ€ใคใŒใ€AI่‡ช่บซใŒAI็ ”็ฉถใ‚’่กŒใ†ใ€ŒAIใ‚ตใ‚คใ‚จใƒณใƒ†ใ‚ฃใ‚นใƒˆใ€ใงใ—ใŸใ€‚ ใ“ใฎๅบฆใ€2024ๅนดใฎAIใ‚ตใ‚คใ‚จใƒณใƒ†ใ‚ฃใ‚นใƒˆใจ2025ๅนดใฎใ€Œv2ใ€ใ‚’็ตฑๅˆใ—ใ€ใ•ใ‚‰ใซๅฎŸ้จ“ใ‚’้‡ใญใฆๅ–ใ‚Šใพใจใ‚ใŸ่ซ–ๆ–‡ใŒ @Nature ่ชŒใซๆŽฒ่ผ‰ใ•ใ‚ŒใŸใ“ใจใ‚’ๅคงๅค‰ๅฌ‰ใ—ใๆ€ใ„ใพใ™ใ€‚ ๆœฌ่ซ–ๆ–‡ใงใฏใ€ใƒขใƒ‡ใƒซๆ€ง่ƒฝใฎๅ‘ไธŠใซไผดใ„็ง‘ๅญฆ่ซ–ๆ–‡ใฎ่ณชใ‚‚ๅ‘ไธŠใ™ใ‚‹ใจใ„ใ†ใ€Œ็ง‘ๅญฆใฎใ‚นใ‚ฑใƒผใƒชใƒณใ‚ฐๅ‰‡ใ€ใ‚’ๅ ฑๅ‘Šใ—ใพใ—ใŸใ€‚ใ“ใ‚ŒใฏไปŠใฎAIใซไฝ•ใŒใงใใ‚‹ใ‹ใ ใ‘ใงใชใใ€ใ“ใ‚Œใ‹ใ‚‰ใฎ้€ฒๅฑ•ใ‚’ไบˆๆœŸใ•ใ›ใ‚‹้‡่ฆใช็Ÿฅ่ฆ‹ใงใ‚ใ‚‹ใจ่€ƒใˆใฆใ„ใพใ™ใ€‚ไบบ้–“ใจAIใŒๅ…ฑๅŒใ—ใฆใ€ๅฏ่ƒฝๆ€งใฎใ€Œๆœจใ€ใ‚’ๆทฑใ้ซ˜้€ŸใซๆŽข็ดขใ—ใ€ๆ–ฐใŸใช็™บ่ฆ‹ใŒๆฌกใ€…ใซ็”Ÿใพใ‚Œใ‚‹ๆœชๆฅใŒๆฅใ‚‹ใ“ใจใฏ้–“้•ใ„ใ‚ใ‚Šใพใ›ใ‚“ใ€‚ ใ“ใฎ็ ”็ฉถใฏใ€ๆ—ฅๆœฌใฎAIใƒฉใƒœใงใ‚ใ‚‹Sakanaโ€…AIใจใ€ไธ–็•Œ็š„ใชAI็ ”็ฉถ่€…ใŸใกใจใฎๅ…ฑๅŒ็ ”็ฉถใซใ‚ˆใ‚‹ๆˆๆžœใงใ™ใ€‚ ๆ—ฅๆœฌใซใŠใ‘ใ‚‹ๆ—ฅๆœฌใฎใŸใ‚ใฎAIไผๆฅญใ‚’็›ฎๆŒ‡ใ™Sakanaโ€…AIใฏใ€ๅ…ˆๆ—ฅๅ…ฌ้–‹ใ—ใŸSakanaโ€…Chatใฎใ‚ˆใ†ใซ่ชฐใ‚‚ใŒไฝฟใˆใ‚‹ๅŸบ็›คๆŠ€่ก“ใ‚’ๆไพ›ใ—ใคใคใ€ๅŒๆ™‚ใซไธ–็•ŒใฎAI็ ”็ฉถใ‚’ๅˆ‡ใ‚Šๆ‹“ใ็ ”็ฉถใ‚’ไธก็ซ‹ใ•ใ›ใฆใ„ใใŸใ„ใจ่€ƒใˆใฆใ„ใพใ™ใ€‚ Nature่ซ–ๆ–‡ใจใ„ใ†ๆœ€้ซ˜ใฎๅฝขใงๅ…ฌ้–‹ใงใใ‚‹ใ“ใจใ‚’ๅŠฑใฟใซใ€ใ“ใ‚Œใ‹ใ‚‰ใ‚‚ๆ—ฅๆœฌใฎAIใจ็ง‘ๅญฆใฎ้€ฒๆญฉใซ่ฒข็Œฎใงใใ‚‹ใ‚ˆใ†ๅ–ใ‚Š็ต„ใ‚“ใงใพใ„ใ‚Šใพใ™ใ€‚ Nature่ซ–ๆ–‡: https://t.co/nNfpSV4Gga

@SakanaAILabs โ€ข Wed Mar 25 16:21

The AI Scientist: Towards Fully Automated AI Research, Now Published in Nature Nature: https://t.co/nNfpSV5e5I Blog: https://t.co/i6h8LVQOdl When we first introduced The AI Scientist, we shared an ambitious vision of an agent powered by foundation models capable of executing th

Media 1Media 2
+1 more
๐Ÿ–ผ๏ธ Media
๐Ÿ”hardmaru retweeted
S
Sakana AI
@SakanaAILabs
๐Ÿ“…
Mar 25, 2026
131d ago
๐Ÿ†”48478400

AIใซใ‚ˆใ‚‹AI็ ”็ฉถใฎๅฎŸ็พใธ๏ผšAIใ‚ตใ‚คใ‚จใƒณใƒ†ใ‚ฃใ‚นใƒˆ่ซ–ๆ–‡ใŒ @Nature ่ชŒใซๆŽฒ่ผ‰ https://t.co/VXCLEFf9R2 Sakanaโ€…AIใŒๅธธใซๆŒ‘ใ‚“ใงใ„ใ‚‹ใฎใฏใ€ๆœ€ๅ…ˆ็ซฏใฎAIใงไฝ•ใŒใงใใ‚‹ใฎใ‹ใจใ„ใ†ๅฏ่ƒฝๆ€งใฎๆœ€ๅ‰็ทšใ‚’ๆŽขใ‚‹ใ“ใจใงใ™ใ€‚ใใ‚Œใ‚’่ฑกๅพดใ™ใ‚‹ๆˆๆžœใฎไธ€ใคใŒใ€AI่‡ช่บซใŒAI็ ”็ฉถใ‚’่กŒใ†ใ€ŒAIใ‚ตใ‚คใ‚จใƒณใƒ†ใ‚ฃใ‚นใƒˆใ€ใงใ—ใŸใ€‚ ใ“ใฎๅบฆใ€2024ๅนดใฎAIใ‚ตใ‚คใ‚จใƒณใƒ†ใ‚ฃใ‚นใƒˆใจ2025ๅนดใฎใ€Œv2ใ€ใ‚’็ตฑๅˆใ—ใ€ใ•ใ‚‰ใซๅฎŸ้จ“ใ‚’้‡ใญใฆๅ–ใ‚Šใพใจใ‚ใŸ่ซ–ๆ–‡ใŒ @Nature ่ชŒใซๆŽฒ่ผ‰ใ•ใ‚ŒใŸใ“ใจใ‚’ๅคงๅค‰ๅฌ‰ใ—ใๆ€ใ„ใพใ™ใ€‚ ๆœฌ่ซ–ๆ–‡ใงใฏใ€ใƒขใƒ‡ใƒซๆ€ง่ƒฝใฎๅ‘ไธŠใซไผดใ„็ง‘ๅญฆ่ซ–ๆ–‡ใฎ่ณชใ‚‚ๅ‘ไธŠใ™ใ‚‹ใจใ„ใ†ใ€Œ็ง‘ๅญฆใฎใ‚นใ‚ฑใƒผใƒชใƒณใ‚ฐๅ‰‡ใ€ใ‚’ๅ ฑๅ‘Šใ—ใพใ—ใŸใ€‚ใ“ใ‚ŒใฏไปŠใฎAIใซไฝ•ใŒใงใใ‚‹ใ‹ใ ใ‘ใงใชใใ€ใ“ใ‚Œใ‹ใ‚‰ใฎ้€ฒๅฑ•ใ‚’ไบˆๆœŸใ•ใ›ใ‚‹้‡่ฆใช็Ÿฅ่ฆ‹ใงใ‚ใ‚‹ใจ่€ƒใˆใฆใ„ใพใ™ใ€‚ไบบ้–“ใจAIใŒๅ…ฑๅŒใ—ใฆใ€ๅฏ่ƒฝๆ€งใฎใ€Œๆœจใ€ใ‚’ๆทฑใ้ซ˜้€ŸใซๆŽข็ดขใ—ใ€ๆ–ฐใŸใช็™บ่ฆ‹ใŒๆฌกใ€…ใซ็”Ÿใพใ‚Œใ‚‹ๆœชๆฅใŒๆฅใ‚‹ใ“ใจใฏ้–“้•ใ„ใ‚ใ‚Šใพใ›ใ‚“ใ€‚ ใ“ใฎ็ ”็ฉถใฏใ€ๆ—ฅๆœฌใฎAIใƒฉใƒœใงใ‚ใ‚‹Sakanaโ€…AIใจใ€ไธ–็•Œ็š„ใชAI็ ”็ฉถ่€…ใŸใกใจใฎๅ…ฑๅŒ็ ”็ฉถใซใ‚ˆใ‚‹ๆˆๆžœใงใ™ใ€‚ ๆ—ฅๆœฌใซใŠใ‘ใ‚‹ๆ—ฅๆœฌใฎใŸใ‚ใฎAIไผๆฅญใ‚’็›ฎๆŒ‡ใ™Sakanaโ€…AIใฏใ€ๅ…ˆๆ—ฅๅ…ฌ้–‹ใ—ใŸSakanaโ€…Chatใฎใ‚ˆใ†ใซ่ชฐใ‚‚ใŒไฝฟใˆใ‚‹ๅŸบ็›คๆŠ€่ก“ใ‚’ๆไพ›ใ—ใคใคใ€ๅŒๆ™‚ใซไธ–็•ŒใฎAI็ ”็ฉถใ‚’ๅˆ‡ใ‚Šๆ‹“ใ็ ”็ฉถใ‚’ไธก็ซ‹ใ•ใ›ใฆใ„ใใŸใ„ใจ่€ƒใˆใฆใ„ใพใ™ใ€‚ Nature่ซ–ๆ–‡ใจใ„ใ†ๆœ€้ซ˜ใฎๅฝขใงๅ…ฌ้–‹ใงใใ‚‹ใ“ใจใ‚’ๅŠฑใฟใซใ€ใ“ใ‚Œใ‹ใ‚‰ใ‚‚ๆ—ฅๆœฌใฎAIใจ็ง‘ๅญฆใฎ้€ฒๆญฉใซ่ฒข็Œฎใงใใ‚‹ใ‚ˆใ†ๅ–ใ‚Š็ต„ใ‚“ใงใพใ„ใ‚Šใพใ™ใ€‚ Nature่ซ–ๆ–‡: https://t.co/nNfpSV4Gga

Media 1
โค๏ธ125
likes
๐Ÿ”23
retweets
๐Ÿ–ผ๏ธ Media
S
sayashk
@sayashk
๐Ÿ“…
Mar 25, 2026
131d ago
๐Ÿ†”69753786

This is a nice essay, and I agree with its characterization of AI as Normal Technology. In fact, on the second line of AINT, we compare AI's potential impact to that of the internet or electricity. But there's a large, qualitative difference between even the most powerful general-purpose technologies which humans can and should influence/control, and creating an omnipotent entity that we have no control over. This was precisely the gap we wanted to highlight.

@deanwball โ€ข Wed Mar 25 13:48

I spent a weekend at Stanford recently, which is where, in 2023, I did much of my formative thinking on AI. The Anthropic-DoW affair tested that early intellectual foundation more than anything, so found myself walking around Stanford, reflecting on what I learned in 2023. https:

Media 1
๐Ÿ–ผ๏ธ Media
M
mercor_ai
@mercor_ai
๐Ÿ“…
Mar 24, 2026
132d ago
๐Ÿ†”42568587

Introducing APEX-SWE, in collaboration with @Cognition. They see firsthand that real software engineering is not just writing code anymore. It's deploying systems, integrating with tools and debugging when things break. On APEX-SWE, every model fails to reliably solve the real production software engineering tasks. @OpenAI GPT-5.3 Codex (High) tops the leaderboard at 41.5% on Pass@1, followed by @AnthropicAI Opus 4.6 (High) at 40.5%. Every frontier model fails on nearly 60% of real production tasks.

Media 1
๐Ÿ–ผ๏ธ Media
L
LiorOnAI
@LiorOnAI
๐Ÿ“…
Mar 26, 2026
131d ago
๐Ÿ†”53392841

Can agents replace software engineers? Not according to this new benchmark. Mercor and Cognition released APEX-SWE. It tests AI coding agents on real engineering work. > GPT-5.3 Codex leads at 41.5%. > Claude Opus 4.6 follows at 40.5%. Nothing crosses the 50% mark. Why? Old benchmarks are basically solved: HumanEval scores jumped from 67% to 90% in two years. OpenAI flagged SWE-bench as contaminated. Models were memorizing the answers. Those benchmarks never reflected the job in the first place. Those tests only measured code writing. Developers spend 16% of their time on that. The other 84% is debugging, infrastructure, and integration. This benchmark tests the 84%. 200 tasks split into two types: 1. Integration: build systems across live databases, APIs, and cloud services in Docker containers 2. Observability: find and fix real bugs using logs, dashboards, and chat history Each task drops an agent into a live environment. Real services, real credentials, and project boards with filler issues mixed in. 50 tasks are open-source on Hugging Face. The eval harness is on GitHub. You can run it yourself. AI writes half the code at big companies. 90% of developers use AI assistants. All of that covers 16% of the job.

@mercor_ai โ€ข Tue Mar 24 17:28

Introducing APEX-SWE, in collaboration with @Cognition. They see firsthand that real software engineering is not just writing code anymore. It's deploying systems, integrating with tools and debugging when things break. On APEX-SWE, every model fails to reliably solve the real p

Media 1
๐Ÿ–ผ๏ธ Media
C
caseyleemoore
@caseyleemoore
๐Ÿ“…
Mar 25, 2026
131d ago
๐Ÿ†”60224424

Lisa Kudrow doing a perfect Parker Posey impression lol https://t.co/LXtYdFNxj5

๐Ÿ–ผ๏ธ Media
๐Ÿ”youwouldntpost retweeted
C
Lee
@caseyleemoore
๐Ÿ“…
Mar 25, 2026
131d ago
๐Ÿ†”60224424

Lisa Kudrow doing a perfect Parker Posey impression lol https://t.co/LXtYdFNxj5

โค๏ธ2,082
likes
๐Ÿ”150
retweets
๐Ÿ–ผ๏ธ Media
T
tokensandai
@tokensandai
๐Ÿ“…
Mar 25, 2026
131d ago
๐Ÿ†”43566620

most AI apps still don't use the full multimodal stack. vision, audio, real-time processing. all largely untapped we got you access to the latest @GoogleDeepMind models. if you've been wanting to build multimodal agents, this one's for you ๐Ÿ† $45K+ in prizes (up to $25K in credits alone) ๐Ÿ“… saturday march 28 ยท san francisco apply below ๐Ÿ‘‡

Media 1
๐Ÿ–ผ๏ธ Media
A
AnthropicAI
@AnthropicAI
๐Ÿ“…
Mar 25, 2026
131d ago
๐Ÿ†”17088921

New on the Engineering Blog: How we designed Claude Code auto mode. Many Claude Code users let Claude work without permission prompts. Auto mode is a safer middle ground: we built and tested classifiers that make approval decisions instead. Read more: https://t.co/dpcMcWMf5k

Media 1
๐Ÿ–ผ๏ธ Media
B
BrianRoemmele
@BrianRoemmele
๐Ÿ“…
Mar 25, 2026
131d ago
๐Ÿ†”81185171

LeWorldModel: Yann LeCuns Radical Simplification of World Models Just Made Physics-Aware AI Practical In the race for artificial general intelligence, two paths have emerged. One is the familiar scale everything route: bigger LLMs trained on ever-larger text corpora. The other, championed for years by Yann LeCun, is building world models: compact systems that learn the underlying physics of reality directly from raw sensory data (pixels) so AI can plan, predict, and act in the physical world like a robot or self-driving car actually would. Until now, the second path has been frustratingly difficult. Joint-Embedding Predictive Architectures (JEPAs) - LeCuns elegant framework for learning predictive representations without reconstructing every pixel - kept collapsing during training. Researchers had to resort to a laundry list of hacks: multi-term loss functions (up to six hyperparameters), frozen pre-trained encoders, stop-gradients, exponential moving averages, and other duct-tape tricks just to keep the model from mapping every input to the same useless output. LeCuns team (Mila, NYU, Samsung SAIL, and Brown University) dropped a bombshell: LeWorldModel (LeWM) - the first JEPA that trains stably end-to-end from raw pixels using only two loss terms. No more house-of-cards engineering. Just a clean, simple recipe that works on a single GPU in a few hours with only 15 million parameters. The Core Breakthrough: SIGReg Saves the Day LeWorldModels secret weapon is a new regularizer called SIGReg (for spherical isotropic Gaussian regularizer). It enforces a simple Gaussian distribution on the latent embeddings. This single term prevents representation collapse without any of the previous heuristics. The training objective now has just two parts: 1. Next-embedding prediction loss - the model predicts what the next latent state should be. 2. SIGReg - keeps the latent space well-behaved and diverse. Thats it. Hyperparameters drop from six to one. Training becomes stable, reproducible, and dramatically cheaper. The model learns directly from raw video frames (no pre-trained vision encoders needed) and produces a compact latent world model that can be used for fast planning. Impressive Results on Real Benchmarks Despite its tiny size, LeWorldModel punches way above its weight: - Trains on a single GPU in a few hours. - Plans actions up to 48 times faster than foundation-model-based world models. - Uses roughly 200 times fewer tokens than alternatives. - Matches or beats far larger models on diverse 2D and 3D control tasks (e.g., manipulation, navigation). - Its latent space encodes meaningful physical quantities (position, velocity, etc.) - proven by direct probing. - It reliably detects physically implausible surprise events, showing genuine causal understanding. Crucially, adding a decoder and reconstruction loss hurts performance on downstream control tasks. The pure JEPA objective already captures everything needed for planning - extra visual details just get in the way. Project website: https://t.co/KhGR9LiIQZ Official code: https://t.co/s1lI9kevJS Why This Matters for the Future of AI LeCun has been saying since 2022 that world models (not next-token predictors) are the key to real intelligence. Critics always pointed to the training instability. LeWorldModel removes that objection with elegant simplicity. This is a philosophical reset: AI can learn physics the way babies do - by watching the world unfold - without needing supercomputers or endless text. The implications for robotics, autonomous vehicles, and embodied agents are enormous. Suddenly, building a physically grounded planner is something a researcher (or even a hobbyist) can do on consumer hardware. 1 of 2

Media 1Media 2
๐Ÿ–ผ๏ธ Media
๐Ÿ”jxnlco retweeted
T
TBPN
@tbpn
๐Ÿ“…
Mar 25, 2026
131d ago
๐Ÿ†”81435890

BREAKING: @ivanleomk is joining Google DeepMind https://t.co/Fu365QjYDk

Media 1
โค๏ธ35
likes
๐Ÿ”2
retweets
๐Ÿ–ผ๏ธ Media
K
ksorbs
@ksorbs
๐Ÿ“…
Mar 25, 2026
131d ago
๐Ÿ†”42251113

Chicago Mayor Brandon Johnson: โ€œWe cannot put people in jail anymore. Itโ€™s racist.โ€ I can't believe this is real https://t.co/0C6hTaJA8g

๐Ÿ–ผ๏ธ Media
B
bneyshabur
@bneyshabur
๐Ÿ“…
Mar 25, 2026
131d ago
๐Ÿ†”63694974

We have been heads down but wanted to share a bit about what we are doing ๐Ÿงต https://t.co/ICZlLwZ3Fq

Media 1
๐Ÿ–ผ๏ธ Media
๐Ÿ”ylecun retweeted
B
Behnam Neyshabur
@bneyshabur
๐Ÿ“…
Mar 25, 2026
131d ago
๐Ÿ†”63694974

We have been heads down but wanted to share a bit about what we are doing ๐Ÿงต https://t.co/ICZlLwZ3Fq

Media 1
โค๏ธ295
likes
๐Ÿ”17
retweets
๐Ÿ–ผ๏ธ Media