@jerryjliu0
Last week I gave a talk at AI Dev ’26 by @DeepLearningAI on “AI can’t read PDFs, how do we fix it” . I’m sharing the slides publicly if others are interested in doing a deep dive into document understanding. AI agents are going to automate huge amounts of knowledge work, but knowledge work depends on data, a lot of that data is in documents/PDFs, and existing OCR tools suck. PDFs are a format that is inherently hard to read, and I dive into specific reasons why (tables, layouts), and why frontier VLMs and benchmarks are still insufficient. Even as agents get better and more general, they need the right tools to read and act over PDFs. They both need this at the data ingest layer, as well as tools they can call on the fly. Check out the slides: https://t.co/IyaTlQMPlW We’re building high-quality AI document processing, both with LlamaParse, along with OSS efforts like LiteParse and ParseBench. If you have a ton of PDFs that you’re hoping to unlock with AI, come talk to us! https://t.co/TsPLwqZ8yg