Your curated collection of saved posts and media
ai avatars are boring We built characters that argue, react, and sometimes go completely off-script Powered by @odysseyml live at https://t.co/HrVojqaxbZ https://t.co/vxSAkZ60id
CVPR 2026π₯We built a model that lets you fly through a generated video world. Not just generating frames β but maintaining a consistent 3D world under complex camera motion. Code, ckpt, and even the data pipeline are all open-sourced β #AI #worldmodel #videogen #cvpr #drone https://t.co/AjLXPPM359
I got no other than the Godmother of AI @drfeifei as a coworker at the @huggingface Miami office today! So inspiring! https://t.co/mdn3xxQcsy https://t.co/AF740DX9zN

The new GitHub repository dashboard is now generally available. You can now find, specify, and save custom views of your projects in seconds. π Filter by language or org π Search via the command palette π Bookmark your top queries Get organized. β¬οΈ https://t.co/8K4U4M54ky https://t.co/vNS13R0zap
When Elon mentioned visiting Saturn, it sounded like from a sci-fi movie. And, he is serious, "Now, wouldnβt it be amazing if you could buy a trip to Saturn? Or frankly, if you just have a trip to Saturn. I think things will just be free in the future. It sounds nuts, but you know, if youβve got an AI robotics economy that is anywhere close to a million times the size of the current Earth economy, literally any need you possibly want can be met. If you can think of it, you can have it".
Our latest LiteParse release gives your AI agent access to text bounding boxes within any PDF π LiteParse is our fast/free open-source document parser that can extract text from any document. Besides the extracted text, we also expose bounding boxes for every text block. This means that if you're building an AI agent over PDFs, it can now trace back to the precise line in the document and highlight it to the user, creating an audit trail for any decision made over this unstructured data. We've created a new guide that shows you how to get bounding boxes (and use it to highlight text on the page similar to the example below). Come check it out! https://t.co/YJFRFPyWxJ LiteParse: https://t.co/JNER0mVcB8
Bounding boxes are key for citations, and we just shipped a new guide showing how to use LiteParse for visual citations! LiteParse is our fast and open-source document parser. Using both bounding box extraction and page screenshots, anyone (including agents) can learn how to ass

@imtiaznabi_ I do: https://t.co/DkQIlgrc5h
Anthropic CEO: β50% of all entry-level Lawyers, Consultants, and Finance Professionals will be completely wiped out within the next 1β5 years." grad students and junior hires are cooked. https://t.co/LGpi7kUl1y
You can now enable Claude to use your computer to complete tasks. It opens your apps, navigates your browser, fills in spreadsheetsβanything you'd do sitting at your desk. Research preview in Claude Cowork and Claude Code, macOS only. https://t.co/sVymgmtEMI
@dexhorthy @trychroma https://t.co/DMmtHdqmjF
@_xjdr I don't see it in the app or the CLI, and can't find it in the CLI source. Is it coming from a skill or something? https://t.co/JATXyTy2Hn

Weirdly, open offices also kill socialization! Face-to-face communication among office workers actually drops by 70%(!!) in open plan offices, since people hide away from others. https://t.co/t9Xu05PRHI

Scaling laws are "just" regressions. But a biased fitting method can quietly misallocate millions of $ of compute at frontier scales. My coworker Eric Czech dug into a bias in parabolic IsoFLOP fits used by Meta, DeepSeek, Microsoft, Waymo, et al. for their scaling lawsπ§΅ https://t.co/JlQO3tEECc
Meet Suno v5.5: More expressive, more you. Use your voice, your sound, and your taste to make music that's unmistakably yours, in the best and most personal Suno experience yet. https://t.co/fKC7VAFLok
The Pulse of Motion Measuring Physical Frame Rate from Visual Dynamics paper: https://t.co/oQ3KAPx225 https://t.co/Zg86MJt0qW
It's over. Been using Tweetdeck (aka X Pro) to follow several of @Scobleizer's lists, but I guess this is the end. https://t.co/qNPNFtBxED
I explained in todayβs @Pragmatic_Eng newsletter: Thatβs a repo where OpenAI is merging external contributions. Some are made by Claude Code, some with GitHub Copilot, some with Codex. Codex doesnβt add itself as a contributor - on purpose - thatβs why. https://t.co/11aiShxeAb https://t.co/z68AZ3sfvW
NCCL watchdog timeouts are some of the most commonly misunderstood and difficult-to-debug errors in AI training. In the past, PyTorch released Flight Recorder to make debugging these errors easier, but interpreting its outputs has been tricky. Based on an analysis of countless NCCL watchdog timeouts at Meta, the PyTorch team at Meta has put together a guide on how to effectively use Flight Recorder to debug NCCL watchdog timeouts, and to better understand why they occur (spoiler: it's usually not because of network or NCCL issues). If you've ever spent hours debugging these errors and are looking for a better solution, this blog post is for you: https://t.co/HIBdkYgpv8 β Yifei Liu, Uttam Thakore, Ph.D., Junjie W., Yongzhong Yang #PyTorch #OpenSourceAI #NCCL #FlightRecorder
so stoked to have @hrishioa speaking at @aiDotEngineer singapore! hrishi is the ceo of southbridge ai - data agents that work unsupervised across difficult, messy data across healthcare, finance, scientific research previously he was cto at greywing (@ycombinator w21). he co-authored "make smart contracts smarter" which was adopted by the @ethereum foundation. @swyx @ivanleomk @agrimsingh @unprofeshme @aimuggle
@ESYudkowsky @slatestarcodex It does include export controls. Specifically: https://t.co/iFLwBnhg5Y
Noise cancelling headphones are apparently a futile solution to office noise⦠they mostly delude us into thinking they help. This study finds we feel they improve our ability to concentrate, but found no significant improvements in actual concentration. https://t.co/KrX8JIQxuO https://t.co/EoUNyVn9cr
OpenClaw 2026.3.24 π¦ π Improved OpenAI API: talk to sub-agents with @openwebui ποΈ Skill & tool management Control UI π¨ Slack interactive reply buttons π Native Microsoft Teams π§΅ Smart Discord auto-thread naming Any client. Any model. One runtime. https://t.co/LhRgNElE0b
The background noise in open offices, in multiple experiments, decreases the ability to do analytical work, increase bugs in software, and decrease the ability to find bugs. Open floor plans are just a bad idea for any solo work. https://t.co/Y2KX5HjIEd
Tech companies pay millions of dollars for their employees and then stick them in open-plan offices that make it nearly impossible to get work done. Best strategy for poaching employees is probably to just offer them an office with a door.

What Iβm reading next: https://t.co/HxkByA4ry0

Thrilled to announce Claude Code auto-fix β in the cloud. Web/Mobile sessions can now automatically follow PRs - fixing CI failures and addressing comments so that your PR is always green. This happens remotely so you can fully walk away and come back to a ready-to-go PR. https://t.co/F41RTeymXJ
π The @GoogleDeepMind team just added Gemini 3.1 to the Live API, so we built a small demo showing how Gemini voice agents can plug directly into the document processing ecosystem powered by LlamaIndex. π₯ In this example, we integrate LiteParse to enable fast, fully-local document parsing. With our TUI-based voice assistant, you can literally talk to your terminal: - Speak commands - Trigger live document parsing via tool calls - Hear the agent read back results in real time π The assistant can extract content from single files or entire folders, leveraging the lightning-fast local parsing that LiteParse provides β‘ Take a look at the demoπ π©βπ» GitHub repo https://t.co/ySmenP2HoY π LiteParse docs https://t.co/NlpoI4CqEq
Photography is deeply personal. Itβs about people, relationships, and moments that matter β shaping how we remember the world and the people around us. So excited to finally share what weβve been working on - a personalized photo generation and editing model designed for your own photography. The idea is simple: your photos should still feel like you - preserving your identity, your expression, your story, while giving you the freedom to shape the image around it, or reimagine it. What excites me most isnβt just the ability to generate something entirely new, but the chance to fix the small, real things β the missed moment, the person who wasnβt there, the expression that didnβt quite land β and finally arrive at the photo you had in your head all along.

Thanks @_akhaliq sharing our work! πWe are exciting to introduce VideoCUA π·A Large-scale video dataset to advance human-level computer-use agent. - Total 55 hours & 6 million frames - 10,000 human-demonstrated tasks - 87 desktop apps, recorded at 30fps π€Fully open-source at huggingface https://t.co/u0mmoTw5zD
CUA-Suite Massive Human-annotated Video Demonstrations for Computer-Use Agents paper: https://t.co/hi1WnffQSY https://t.co/ACjYOPcDzH
Btw you may need to force refresh the page due to browser caching (Cmd + Shift + R or Ctrl+ Shift + R). And yes, you can now also sort by date, size, and name: https://t.co/pow8khbpr9
NEW: Elonβs attorneys are calling on Clinton-appointed Judge Charles Breyer to investigate the jury in Musk's recent Twitter acquisition case after jurors were mocking Musk in court. RIGGED https://t.co/4qxyTm9PSx
"In a sane world, what happens is the leadership of the United States sits down with the leadership in China and leadership around the world to work together so that we don't go over the edge and create a technology which could perhaps destroy humanity. " β Bernie Sanders https://t.co/mkJOJyBpKD
After @Pinterest @Airbnb @NotionHQ @cursor_ai, today itβs @eoghan @intercom publicly sharing that theyβre finding it better, cheaper, faster to use and train open models themselves rather than use APIs for many tasks. And hundreds of other companies are doing the same without sharing. Ultimately, I believe the majority of AI workflows will be in-house based on open-source (vs API). It took much more time than we anticipated but itβs happening now!
LagerNVS Latent Geometry for Fully Neural Real-time Novel View Synthesis paper: https://t.co/d8Boz9XIGR https://t.co/KWZQdHYL4L