Your curated collection of saved posts and media
New paper π§΅ We show that dynamic short convolutions consistently improve Transformers across scales. We make these gains practical with an efficient parameterization and custom Triton GPU kernels. The improvements carry over to MoEs and linear attention variants (Mamba-2/GDN). https://t.co/Py6isYX0LK
In film, "we'll fix it in post" is what you say when something went wrong on set and you don't want to redo it. AI research has made it our entire methodology: train the model, then patch whatever comes out. Our new ICML oral argues this can't be the basis of a science of AI. π§΅ https://t.co/ok11oGRhUQ
We ground discussion in the history and philosophy of science. What did it take for other fields to move from cataloging phenomena to predicting and controlling them? AI can learn from that playbook. https://t.co/rWWMVVFlgn

A common issue with position papers is that they leave the reader wondering βokay, but what should I actually doβ? To address this we provide open problems on a wide variety of topics throughout to illustrate our perspectives and guide future research https://t.co/2FSgK95W4K

@guilhermeotina Yes, if we said that we would be very silly. But that's not what we're talking about. Scaling laws, grokking, and induction heads are some of the best examples of the kind of work we are advocating for. https://t.co/dUfGEGXgp4
@typewriters Literally put it on my calendar for next year https://t.co/tSz9DGVr03
Highlighting the new WebGPU backend in llama.cpp/ggml The work to bring full-fledged WebGPU support in llama.cpp started about an year and a half ago. It has been lead by @reeselevine and team at USCS. For more information, checkout the interactive blog and paper in the quoted post. Here are 2 excerpts from the paper, summarizing the implemented software architecture.
WebGPU support in llama.cpp is here! Check out our blog post introducing it: https://t.co/3OUusMYqIY Run local models in your browser, with GPU acceleration. No data leaves your computer! Thanks to everyone who's made this possible, especially @ggerganov

More info and discussion: https://t.co/w9rujz5SOY
These are some of my LLM assisted contributions from the past month. Nothing amazing, but I'm slowly getting better at it. Atm, using Qwen3.6 27B exclusively. For hardware - switching between M2 Ultra and RTX 5090. Both are good options, though after using the RTX and going back to the Mac, it always feels like a snail. Yet for most tasks, I feel like both hardware can do the job comfortably.
Build on-device personal AI agents on Windows PCs with new tools from NVIDIA and Microsoft, including secure sandboxing, faster local inference, multi-GPU support, and RTX acceleration for Windows AI APIs. Read the technical blog: https://t.co/vNIsEded46 https://t.co/4bPWPDERJO

Highlighting recent advances in multi-GPU and tensor parallel support in llama.cpp Over the last few months llama.cpp maintainers and engineers from NVIDIA collaborated to improve the multi-GPU performance in ggml. This resulted in significant performance gains on RTX systems and laid the groundwork for hardware-agnostic tensor parallelism in ggml. For more information on this and other advancements in the low-level inference engine of llama.cpp, check the technical blog by @NVIDIARTXSpark below
Build on-device personal AI agents on Windows PCs with new tools from NVIDIA and Microsoft, including secure sandboxing, faster local inference, multi-GPU support, and RTX acceleration for Windows AI APIs. Read the technical blog: https://t.co/vNIsEded46 https://t.co/4bPWPDERJO
Building super fast experiences with Gemma just got easier. Gemma 4 MTP is now officially merged into llama.cpp. Developers can now pair MTP with Gemma 4 QAT for a fast, lightweight setup. https://t.co/CIynVMYuZm
i just beat @GoogleDeepMind's turboquant introducing Shard. 10x KV cache compression on Llama-3.1-8B. zero quality loss - 10x @ 8K context, 11.2x @ 32K - NIAH recall 1.000 across 4K-32K - LongBench Ξ β 0 vs FP16 turboquant tops out at 4-6x at the same quality. we doubled it. read more: https://t.co/PAV5WdAzN6 @kirrithan
AI now helps doctors read X-rays, CTs, and MRIs. But once it's deployed in a hospital, almost no one can tell if it's still accurate. Lattice Health watches deployed imaging AI and flags it the moment it starts to slip. Congrats on the launch, @sparkcpark! https://t.co/VWndL8ja8E
@_DavidLKing_ @VirginiaLConn -57 https://t.co/lgmsXLWYkG
Submissive and Breadable https://t.co/l7gZZQh4f5
@16kbps @atomicthumbs There's also this EXTREMELY cursed "standard" https://t.co/dwuk6WFlxQ https://t.co/RGRAc9XaHR
@retronouns https://t.co/Ew7EqgSse0
Jay Mohr's anecdote about meeting Christopher Walken is the kind of anecdote you could only wish for. @FighterNtheKid https://t.co/OoykZtfd52
And the winner is Ali Smithβs Gliff. A wonderful novel. @DublinLitAward https://t.co/lsRA3JNfJf

Spotify Introduces AI-Generated Personal Podcasts https://t.co/NSHw1C6OWW
Playgrounds are much cooler these days... This is a space themed park in Erie CO. In the background you'll see a path with solar system facts on engraved stones. Behind me is a fake half Moon you can climb. https://t.co/ohlfXxkSBr
Put down your phone and be present with those around you or go for a walk... > JOMO - digital wellbeing @google https://t.co/XKuxWkma8M
A recent post from @joeerl (author of erlang) has me thinking about web-futures. I reflect on how far the web has come, and how much utility there is now - but he is right, it should go *much* further, and decentralization & persistence is key. https://t.co/DqvYo2AQVb
When you have plumbing... You must knit it. https://t.co/TPRBhkF6kG
Big Bang AR app from Google+CERN https://t.co/XCkifNospK -- saw this and had to send it to @ladyofsteele
@ladyofsteele I'm happy to share this dot with you. Your troubles are part of your journey, sorry to say... But your journey is epic and inspiring and worth it. I'm happy my own journey gets to interweave with yours a bit. https://t.co/xmQ3kP3yrD
This was a wonderfully written piece, about an inspiring programmer/inventor, and someone I would have loved to have known - Remembering Joe, a Quarter of a Century of Inspiration and Friendship https://t.co/i4NIelq6rP #joearmstrong #rememberingjoe
@MThomps4 @echobind This is Colorado https://t.co/mE1YmsyPK6
If you ever start to lose faith in humanity, go run a race. So thankful for my hype squad! @RunRocknRoll https://t.co/ZfNnSWpdEf
Good morning world π https://t.co/WlInbesOoH
Lululemon dashed into the athletic shoe segment in March with its Blissfeel running shoe for women. Last week, $LULU made its next move, launching Restfeel recovery slides. https://t.co/b5NPLgE0n4