Your curated collection of saved posts and media

Showing 32 posts ยท last 7 days ยท newest first
A
allen_ai
@allen_ai
๐Ÿ“…
May 07, 2026
80d ago
๐Ÿ†”39169940

Today weโ€™re bringing new NSF OMAI compute online with NVIDIA Blackwell Ultra-powered systems, turning a $152M national investment from @NSF & @NVIDIA into a foundation for truly open AI research. ๐Ÿงต https://t.co/qFgtiibgAK

Media 1Media 2
+1 more
๐Ÿ–ผ๏ธ Media
S
stereogum
@stereogum
๐Ÿ“…
Apr 22, 2026
95d ago
๐Ÿ†”97471885

Happy 80th birthday John Waters! โšก๏ธ โ€จ[๐Ÿ“ท: Rajendra Roy] https://t.co/G35qVTHOP4

Media 1
๐Ÿ–ผ๏ธ Media
๐Ÿ”Vallmeister retweeted
S
Stereogum
@stereogum
๐Ÿ“…
Apr 22, 2026
95d ago
๐Ÿ†”97471885

Happy 80th birthday John Waters! โšก๏ธ โ€จ[๐Ÿ“ท: Rajendra Roy] https://t.co/G35qVTHOP4

Media 1
โค๏ธ20,947
likes
๐Ÿ”2,874
retweets
๐Ÿ–ผ๏ธ Media
M
milkkarten
@milkkarten
๐Ÿ“…
Apr 29, 2026
88d ago
๐Ÿ†”32735477

lena dunham reportedly sold 60k copies of her book in the first week. she did a ton of press on substackโ€”guest essays, q&as, etc. thought this was an interesting insight that likely also applies to products other than books. https://t.co/JY3FO0tux7

Media 1
๐Ÿ–ผ๏ธ Media
N
NetflixUK
@NetflixUK
๐Ÿ“…
May 08, 2026
79d ago
๐Ÿ†”80245051

100 years old and still the coolest person alive. Happy birthday, Sir David! https://t.co/dgxQb6LPZ6

๐Ÿ–ผ๏ธ Media
๐Ÿ”Vallmeister retweeted
N
Netflix UK & Ireland
@NetflixUK
๐Ÿ“…
May 08, 2026
79d ago
๐Ÿ†”80245051

100 years old and still the coolest person alive. Happy birthday, Sir David! https://t.co/dgxQb6LPZ6

โค๏ธ51,358
likes
๐Ÿ”7,380
retweets
๐Ÿ–ผ๏ธ Media
S
svpino
@svpino
๐Ÿ“…
May 09, 2026
78d ago
๐Ÿ†”11660309

Python, MCP, A2A, and a whole lot about how to build multi-agent systems. This goes well beyond building a single agent: it focuses on how multiple agents can collaborate towards a goal. Best part: The book will show you how to build everything from scratch. You start with a simple agent and then add complexity as the book progresses. Amazon link: https://t.co/GIIiIOKxgw

Media 1
๐Ÿ–ผ๏ธ Media
K
Kelaivy
@Kelaivy
๐Ÿ“…
Apr 24, 2026
94d ago
๐Ÿ†”58476811

@Santandave1 โ€œHeโ€™s known for changing a key like John Coltraneโ€ https://t.co/6kjL98wfvE

Media 1
๐Ÿ–ผ๏ธ Media
I
itsmeebusybi
@itsmeebusybi
๐Ÿ“…
May 03, 2026
84d ago
๐Ÿ†”88600567

@animeposuto https://t.co/j4RPTYOZ7I

Media 1
๐Ÿ–ผ๏ธ Media
๐Ÿ”Kelaivy retweeted
I
itsmeebeโœจ
@itsmeebusybi
๐Ÿ“…
May 03, 2026
84d ago
๐Ÿ†”88600567

@animeposuto https://t.co/j4RPTYOZ7I

Media 1
๐Ÿ”1
retweets
๐Ÿ–ผ๏ธ Media
C
centregoals
@centregoals
๐Ÿ“…
May 08, 2026
79d ago
๐Ÿ†”09308206

๐Ÿšจ๐Ÿšจ| OFFICIAL: Morgan Gibbs-White has been named the Premier League Player of the Month for April! ๐Ÿด๓ ง๓ ข๓ ฅ๓ ฎ๓ ง๓ ฟ๐ŸŽ–๏ธ https://t.co/VhSfCYkuBQ

Media 1Media 2
๐Ÿ–ผ๏ธ Media
N
nxthompson
@nxthompson
๐Ÿ“…
Jan 01, 2026
206d ago
๐Ÿ†”40302610

Words to live by todayโ€”and every day. https://t.co/qFRIiQcCGH

Media 1
๐Ÿ–ผ๏ธ Media
๐Ÿ”mebrantley2 retweeted
N
nxthompson
@nxthompson
๐Ÿ“…
Jan 01, 2026
206d ago
๐Ÿ†”40302610

Words to live by todayโ€”and every day. https://t.co/qFRIiQcCGH

Media 1
โค๏ธ11,082
likes
๐Ÿ”2,486
retweets
๐Ÿ–ผ๏ธ Media
N
NaderLikeLadder
@NaderLikeLadder
๐Ÿ“…
May 07, 2026
80d ago
๐Ÿ†”05012958

NVIDIA is at SCSP AI EXPO in Washington D.C. Giving a talk with @Baxate about importance of American open source, from internet to Linux to AI, with a demo of OpenClaw paying my parking tickets 3:25pm ET @ Convention center. At Expo hall demo stage https://t.co/QYpIWKrNrF

๐Ÿ–ผ๏ธ Media
Y
yourweirdaunt
@yourweirdaunt
๐Ÿ“…
May 08, 2026
80d ago
๐Ÿ†”15397033

Another packed night at @PubKey ๐Ÿป๐Ÿค– Loved hosting @WashingtonAINet with Tammy Haddad moderating a great conversation with @NaderLikeLadder and @Baxate on the future of AI, innovation, and all the weird exciting stuff ahead. Smart people, strong drinks, big ideas โ€” exactly the kind of nights we love.

Media 1Media 2
+1 more
๐Ÿ–ผ๏ธ Media
N
NaderLikeLadder
@NaderLikeLadder
๐Ÿ“…
May 09, 2026
79d ago
๐Ÿ†”62490922

Super busy week in DC talking to policy makers and think tanks on AI. The space is moving fast for technologists, even faster for people who recently had to learn what โ€œopen sourceโ€ is for the first time https://t.co/3K3BnZLeIE

Media 1Media 2
+2 more
๐Ÿ–ผ๏ธ Media
B
Baxate
@Baxate
๐Ÿ“…
May 09, 2026
79d ago
๐Ÿ†”01826710

Fun time on the hill explaining why Open Source is deeply strategic for America https://t.co/ilPDzCXe1u

Media 1Media 2
+2 more
๐Ÿ–ผ๏ธ Media
๐Ÿ”NaderLikeLadder retweeted
B
Baxate
@Baxate
๐Ÿ“…
May 09, 2026
79d ago
๐Ÿ†”01826710

Fun time on the hill explaining why Open Source is deeply strategic for America https://t.co/ilPDzCXe1u

Media 1Media 2
+2 more
โค๏ธ37
likes
๐Ÿ”1
retweets
๐Ÿ–ผ๏ธ Media
M
mvanhorn
@mvanhorn
๐Ÿ“…
May 07, 2026
80d ago
๐Ÿ†”81611010

Introducing the Printing Press, a CLI-factory and a CLI-library. Built with @trevin. ๐Ÿญ๐Ÿ–จ๐Ÿ“š Most APIs suck for agents. Most MCPs suck for agents. Most official CLIs suck for agents. They waste tokens and time. @steipete started making his own because of this. ๐Ÿ“š A Library of agent-native CLIs you install today (Linear, ESPN, Flight GOAT (Google Flights + Kayak nonstop), Contact Goat (LinkedIn + Happenstance + Deepline more) +30+ more) ๐Ÿญ A factory that prints new ones for any service - just type /printing-press <product name> CLIs are fast, local, SQLite-backed. Work in Claude Code, Codex, OpenClaw, Hermes. ๐ŸŒ https://t.co/GjnN9E9yTH

Media 2
๐Ÿ–ผ๏ธ Media
T
tbpn
@tbpn
๐Ÿ“…
May 08, 2026
79d ago
๐Ÿ†”78560499

"I think the highest turnover job of any Silicon Valley executive position might be the CMO." Airbnb's @bchesky says he has a huge amount of respect for people in marketing because it's one of the hardest functions in a business. "Because once something works, it almost becomes stale. Because everyone does it." "Part of my theory is that what works in marketing changes every few years and you have to be adaptable. Your old playbook gets outdated." "It's this thing called 'banner blindness.' After you see something over and over, you tend to be blind to it, so you need a new tactic."

๐Ÿ–ผ๏ธ Media
N
NaderLikeLadder
@NaderLikeLadder
๐Ÿ“…
May 11, 2026
76d ago
๐Ÿ†”22264595

I wake up every morning to 3 emails summarizing 1) emails from last week, 2) emails from the weekend, and 3) emails from this morning It's so wild that we used to spend manual time curating emails to bring ourselves up to speed. https://t.co/LJOJO4TATk

Media 1
๐Ÿ–ผ๏ธ Media
N
nvidianewsroom
@nvidianewsroom
๐Ÿ“…
May 11, 2026
76d ago
๐Ÿ†”50137605

NVIDIAโ€™s Nader Khalil and Carter Abdallah joined Tammy Haddad and the @WashingtonAINet to discuss how open models, agents and tools like NemoClaw can help build safer, more transparent AI ecosystems. Open source AI is not just a developer story. Itโ€™s a U.S. leadership story. https://t.co/MIiz9hbZVc

Media 1Media 2
๐Ÿ–ผ๏ธ Media
P
PyTorch
@PyTorch
๐Ÿ“…
Apr 30, 2026
88d ago
๐Ÿ†”62850699

The clock is ticking to join the program for #KubeCon + #CloudNativeCon + #OpenInfraSummit + #PyTorchCon China. Share your knowledge with the global ecosystem & submit to speak by 3 May: https://t.co/U0HIhZMoFE And don't forget early bird rates end in 1 week! Save now: https://t.co/bTRuC5Azau

๐Ÿ–ผ๏ธ Media
P
PyTorch
@PyTorch
๐Ÿ“…
Apr 30, 2026
87d ago
๐Ÿ†”66077112

Got a #PyTorch breakthrough to share? ๐ŸŽค The #CallForProposals for #PyTorchCon North America (Oct 20-21 | San Jose) is in full swing. Submit your technical session idea by June 7! Apply: https://t.co/hLlKK7WxLD https://t.co/3RbHVVvQvY

๐Ÿ–ผ๏ธ Media
P
PyTorch
@PyTorch
๐Ÿ“…
Apr 30, 2026
87d ago
๐Ÿ†”44925048

LightSeek Foundation recently released Shepherd Model Gateway (SMG). It came out of a production bottleneck in LLM serving: CPU-bound work sitting on the critical path. Light Seek moved all non-GPU work into a Rust-based gateway, with a minimal gRPC boundary around inference. The project is developed and maintained by @lightseekorg. The result: up to 3.5ร— throughput in long-context scenarios. ๐Ÿ”— More details in our latest blog: https://t.co/S4OzoakVBr #PyTorch #LightSeek #OpenSourceAI #ShepherdModelGateway

Media 1
๐Ÿ–ผ๏ธ Media
V
vllm_project
@vllm_project
๐Ÿ“…
Apr 24, 2026
94d ago
๐Ÿ†”51105796

๐ŸŽ‰ Day-0 support for @deepseek_ai V4 Pro and Flash on vLLM โ€” a new generation of DeepSeek model, purpose-built for tasks up to 1M tokens. Alongside the release, we're publishing a first-principles walkthrough of the new long-context attention and how we implemented it in vLLM. The new attention mechanism, in four moves: โ€ข Shared K/V + inverse RoPE โ†’ 2ร— memory savings โ€ข c4a / c128a KV compression โ†’ 4ร—โ€“128ร— savings โ€ข DeepSeek Sparse Attention over compressed tokens โ€ข Short sliding window for locality across compression boundaries At 1M context, per-layer KV state is ~8.7ร— smaller than a DeepSeek V3.2-style 61-layer stack (9.62 GiB vs 83.9 GiB, bf16). fp8 attention cache + fp4 indexer cache shrink it further. vLLM side: โ€ข Unified hybrid KV cache โ€” single logical block size (256 native positions) across all compression rates; compressor state folded into the SWA KV cache spec so prefix caching, disagg prefill, CUDA graphs and MTP reuse the same abstraction โ€ข Three page-size buckets for the full 5-way cache stack โ†’ no cross-kind fragmentation โ€ข Fused kernels: compressor + RMSNorm + RoPE + cache insert (1.4โ€“3ร—), inverse RoPE + fp8 quant (2โ€“3ร—), Q-norm + KV RoPE + K insert (10โ€“20ร—) โ€ข Multi-stream overlap of indexer vs main-KV compression vs SWA insertion Disaggregated serving is supported out of the box and strongly recommended for best performance. Follow our recipes site for verified commands for @nvidia Blackwell (B200, B300, GB200, GB300) and Hopper (H100/H200/H20) systems. Thanks to the @deepseek_ai team for open-sourcing DeepSeek V4, and to @inferact for landing day-0 support ๐Ÿค ๐Ÿ“ Blog: https://t.co/Eh7vk6xVJy ๐Ÿ“– Recipes: https://t.co/jlWuzYyZeX ๐Ÿค— https://t.co/IA9qAysqJk

@deepseek_ai โ€ข Fri Apr 24 03:24

๐Ÿš€ DeepSeek-V4 Preview is officially live & open-sourced! Welcome to the era of cost-effective 1M context length. ๐Ÿ”น DeepSeek-V4-Pro: 1.6T total / 49B active params. Performance rivaling the world's top closed-source models. ๐Ÿ”น DeepSeek-V4-Flash: 284B total / 13B active params. You

Media 1Media 2
+2 more
๐Ÿ–ผ๏ธ Media
P
PyTorch
@PyTorch
๐Ÿ“…
May 03, 2026
85d ago
๐Ÿ†”61090152

FINAL CALL! ๐Ÿšจ The #CFP for #KubeCon + #CloudNativeCon + #OpenInfraSummit + #PyTorchCon China, 8-9 Sept in Shanghai, closes TODAY at 23:59 China Standard Time. Don't miss your chance to speak on the premier #OpenSource stage. Submit NOW: https://t.co/PKdf9djzqR https://t.co/ELdAu3V4KR

๐Ÿ–ผ๏ธ Media
P
PyTorch
@PyTorch
๐Ÿ“…
May 05, 2026
83d ago
๐Ÿ†”34756723

Registration deadline alert! ๐Ÿšจ Tomorrow is the last day to save ยฅ1,270 (USD$179) on your pass for #KubeCon + #CloudNativeCon + #OpenInfraSummit + #PyTorchCon China in Shanghai (8-9 Sept). Secure early bird rates by 6 May at 23:59 CST: https://t.co/9x7dLq0gVE https://t.co/ttriVff1dg

๐Ÿ–ผ๏ธ Media
P
PyTorch
@PyTorch
๐Ÿ“…
May 05, 2026
82d ago
๐Ÿ†”27690094

Relive the energy of 2025 as we prepare for San Jose! โšก #PyTorchCon North America returns to the Bay Area Oct 20-21. Check out last yearโ€™s highlights &amp; lock in your $400 savings today. ๐Ÿ“บ Recap 2025: https://t.co/yLtZy5VrD9 ๐ŸŽŸ๏ธ Register for 2026: https://t.co/AVHdaIFT20 https://t.co/uwx47Q9EXK

๐Ÿ–ผ๏ธ Media
P
PyTorch
@PyTorch
๐Ÿ“…
May 05, 2026
82d ago
๐Ÿ†”29033592

Weโ€™re excited to share In-Kernel Broadcast Optimization (IKBO)โ€”a kernel-model-system co-design approach driving @Meta's large-scale recommendation system (RecSys) inference. It serves as the scalability backbone for the request-centric, inference-efficient framework underlying the Meta Adaptive Ranking Model (serving LLM-scale models in production). The core idea is simple but powerful: broadcast is a data layout concern, not a computational necessity. Instead of materializing broadcast at the service layer, IKBO pushes it all the way down into the kernel layerโ€”precisely where request and candidate features interact. This enables the system, together with model co-design, to capture the full performance upside. By preserving request and candidate inputs in their native batch sizes and fusing broadcast directly into the kernel, IKBO eliminates replicated memory movement and redundant computation across the stack. From Attention in sequence modeling to GEMM in traditional workloads, the same idea delivers broad inference gains. Results from our co-designed deployments: โšก Up to a 2/3 reduction in compute-intensive net latency on co-designed models. โšก End-to-end deployment across Meta's RecSys stacks on both GPU and MTIA (Meta's in-house accelerator). โšก Up to 4x speedups for linear compression GEMM and up to 2.4x/6.4x (kernel only / kernel + broadcasting) throughput for IKBO attention compared to FlashAttention-4 on H100. ๐Ÿ”— If youโ€™re interested in how rethinking a primitive like broadcast can unlock system-level gains for modern AI inference, take a look: https://t.co/mdl0F1Zj7E By @jiao_jian62798, @liptds_11, Hongtao Yu, Yuanwei Fang, Zhengkai Zhang, Zoran Zhao, Yuxuan Chen, Sijia Chen, Yang Chen #PyTorch #OpenSourceAI #RecSys #AIInfrastructure

Media 1
๐Ÿ–ผ๏ธ Media
P
PyTorch
@PyTorch
๐Ÿ“…
May 05, 2026
82d ago
๐Ÿ†”39396841

The second in the series of Simplifying Sparse Deep Learning with Universal Sparse Tensor in nvmath-python In a previous post, NVIDIA introduced the Universal Sparse Tensor (UST), enabling developers to decouple a tensorโ€™s sparsity from its memory layout for greater flexibility and performance. Weโ€™re excited to announce the integration of the UST into nvmath-python v0.9.0 to accelerate sparse scientific and deep learning applications. Read the follow-on post: https://t.co/sARnNMbz1r #PyTorch #OpenSourceAI #AI #Inference #Innovation

Media 1
๐Ÿ–ผ๏ธ Media
V
vllm_project
@vllm_project
๐Ÿ“…
May 05, 2026
82d ago
๐Ÿ†”16574950

๐Ÿš€ Day-0 MTP support for Gemma4 now available at vLLM with ready-to-use docker image! โšก๏ธEnjoy up to 3x faster decoding performance to supercharge your development with zero quality degradation! Check out the full vLLM recipes for Gemma 4 model series๐Ÿ‘‡ https://t.co/IrCaaa6SIo https://t.co/eFcAZRogLF

@googledevs โ€ข Tue May 05 16:28

Gemma 4: Now up to 3x Faster. โšก Same quality, way more speed. Our new MTP drafters allow Gemma 4 to predict multiple tokens at once, effectively tripling your output speed without compromising intelligence. https://t.co/xyltPFFVMw

Media 1Media 2
๐Ÿ–ผ๏ธ Media