LlamaIndex @llamaindex.bsky.social · 12/05/2026Need document parsing that stays fully local and private? 👀 Meet liteparse-server, a self-hostable, open-source HTTP server for parsing documents and generating screenshots from PDFs, Office files, and images. ✅ 100% self-hosted ✅ Private by default ✅ Open source ✅ Built for production deployments 210
LlamaIndex @llamaindex.bsky.social · 11/05/2026Ever wished your agent could read PDFs, images, and Office documents as easily as plain text? Or combine the safety of a secure sandbox with the full power of Bash access? We built exactly that. Meet 𝘀𝗮𝗻𝗱𝗯𝗼𝘅𝗲𝗱-𝗹𝗶𝘁, a Rust 🦀 CLI agent that combines: 210
LlamaIndex @llamaindex.bsky.social · 07/05/2026A few weeks ago @simonw got Claude to port LiteParse to the browser. Today, we are launching that work as a complete guide in our docs! developers.llamaindex.ai/liteparse/g...developers.llamaindex.aiBrowser UsageRun LiteParse in the browser with Vite. 100
LlamaIndex @llamaindex.bsky.social · 06/05/2026What if you could extract text from any photo on your phone? We built LlamaParse Mobile, an @expo.dev + @reactnative.dev app for iOS & Android, powered by the LlamaParse TypeScript SDK 📱 210
LlamaIndex @llamaindex.bsky.social · 06/05/2026LlamaIndex NYC takeover, 5/13 🗽 Our CEO Jerry Liu is in town. Two events, open to every NYC builder: 🛠️ FinParse Workshop — laptops out, hands-on with @jerryjliu0 → luma.com/updli8i6 🍕 AI Engineers on Tap — happy hour w/ @tabs → luma.com/tklfgwh8luma.comNYC AI Engineer On Tap · LumaCalling all AI Engineers in NYC. What happens when an AI infrastructure powerhouse (LlamaIndex) and a fintech darling (Tabs) walk into a bar? You get the… 000
LlamaIndex @llamaindex.bsky.social · 30/04/2026Building scalable, distributed document processing pipelines isn’t easy. That’s why we teamed up with @render.com to build a system that: 121
LlamaIndex @llamaindex.bsky.social · 29/04/2026Parsing documents with AI agents just got a lot more seamless🚀 We've rebuilt the LlamaParse MCP server to handle your document processing workflows, and you can connect it today to any MCP-compatible client at mcp.llamaindex.ai/mcp 🌐 120
LlamaIndex @llamaindex.bsky.social · 27/04/2026Loan processors spend 40–60% of their time reconciling income across tax returns, pay stubs, W-2s, and bank statements. We built an end-to-end pipeline that automates it with LlamaParse + the Claude Agent SDK: 📄 Schema-driven extraction across 4 doc types with confidence scores + citations 200
LlamaIndex @llamaindex.bsky.social · 23/04/2026ParseBench is now live on @Kaggle. The first document OCR benchmark built for AI agents — 2,000 enterprise pages, 167K+ test rules, 5 dimensions that actually break downstream agents. Benchmark your parser against 14 methods including GPT-5 Mini, Gemini 3, Textract, and LlamaParse. 200
LlamaIndex @llamaindex.bsky.social · 22/04/2026LiteParse: our open-source, layout-aware PDF parser for AI agents. The secret? Grid projection. Instead of heavy ML layout models or flat text extraction, it projects text onto a monospace grid so alignment preserves structure. Full deep dive into the grid projection algorithm behind the magic ↓ 130
LlamaIndex @llamaindex.bsky.social · 10/04/2026LiteParse hit 4K+ GitHub stars in 3 weeks. ~500 pages in 2 seconds. No GPU. No API keys. 50+ file formats. Now @LoganMarkewich, our Head of Open Source, will show you how to build with it. Live workshop — April 28, 9 AM PST: Build a Financial Due Diligence Agent with LiteParse. 240
LlamaIndex @llamaindex.bsky.social · 09/04/2026Agents like OpenClaw are incredibly powerful, as long as the information they receive is clean and structured🦞 320
LlamaIndex @llamaindex.bsky.social · 07/04/2026Visually rich documents are especially challenging for agents. Tables, charts, and images often break traditional document pipelines, making complex reasoning difficult📄 So we teamed up with LanceDB to build a structure-aware PDF QA pipeline🚀 Here’s how it works: 211
LlamaIndex @llamaindex.bsky.social · 07/04/2026Open call to fintech leaders in NYC 🏦 May 13, in-person workshop with @jerryjliu0 on turning complex financial docs into LLM-ready data using agentic OCR. Build real pipelines. Hear from a Top 5 PE firm's production agent. Make sure to bring your laptops→ luma.com/updli8i6luma.comTurn Complex Financial Docs into LLM-ready Data with Jerry Liu and LlamaIndex · LumaA hands-on workshop for engineers building VLM-powered OCR that works on real-world financial documents. Most document pipelines fail quietly. They work on a… 100
LlamaIndex @llamaindex.bsky.social · 02/04/2026After the release of Parse v2, Extract is also getting an upgrade — 𝗶𝗻𝘁𝗿𝗼𝗱𝘂𝗰𝗶𝗻𝗴 𝗘𝘅𝘁𝗿𝗮𝗰𝘁 𝘃2! 🎉 We've been reworking the experience from the ground up to make document extraction more powerful and easier to use than ever. Here's what's new: 281
LlamaIndex @llamaindex.bsky.social · 31/03/2026LlamaIndex is proud to be named to the 2026 Enterprise Tech 30, #3 in the Early Stage category. The ET30 is an annual list by @Wing_VC and Eric Newcomer, voted on by 210
LlamaIndex @llamaindex.bsky.social · 30/03/2026We’ve moved to a new office and it’s time to celebrate. Swing by this Thursday to meet our team, grab a bite, and make new friends. Note: Space is limited, so please RSVP early. luma.com/mkh44c7wluma.comStartup Party Up for First Thursday · LumaWe’ve moved to the 'AI Waterfront' and it’s time to celebrate. Swing by on April 2nd to see our new office on 2nd street, meet our team, and make new… 100
LlamaIndex @llamaindex.bsky.social · 30/03/2026Our OSS engineer @cle-does-things.bsky.social recently built 𝗹𝗶𝘁𝗲𝘀𝗲𝗮𝗿𝗰𝗵, a fully local document ingestion and retrieval CLI/TUI application powered by LiteParse ⚡ litesearch demonstrates how developers can assemble a high-performance, local-first pipeline using tools from across the ecosystem: 241
LlamaIndex @llamaindex.bsky.social · 27/03/2026Transform your document processing with intelligent table extraction that goes beyond basic OCR. 200
LlamaIndex @llamaindex.bsky.social · 26/03/2026🚀 The @GoogleDeepMind team just added Gemini 3.1 to the Live API, so we built a small demo showing how Gemini voice agents can plug directly into the document processing ecosystem powered by LlamaIndex. 🔥 In this example, we integrate LiteParse to enable fast, fully-local document parsing. 200
LlamaIndex @llamaindex.bsky.social · 26/03/2026Bounding boxes are key for citations, and we just shipped a new guide showing how to use LiteParse for visual citations! developers.llamaindex.ai/liteparse/g... 100
LlamaIndex @llamaindex.bsky.social · 25/03/2026Word docs are one of the most common file formats people process in LlamaParse, and they've always been surprisingly frustrating to parse well. 221
LlamaIndex @llamaindex.bsky.social · 24/03/2026Congratulations to Zubeen, one of our LlamAgent contest winners, for building an agentic AI workflow that automates GDPR breach report structuring! 200
LlamaIndex @llamaindex.bsky.social · 23/03/2026We’ve published a new blog with @developers.google.com on how to build a smart financial assistant using LlamaParse, our state-of-the-art agentic document parser, together with Gemini 3. 120
LlamaIndex @llamaindex.bsky.social · 23/03/2026If you've ever worked in or around legal, you know that discovery is where document parsing really gets stress-tested. 110
LlamaIndex @llamaindex.bsky.social · 20/03/2026LlamaParse now has an official Agent Skill you can use across 40+ agents. With built-in instructions for parsing complex documents, including different formats, tables, charts, and images, your agents gain access to deeper document understanding, not just raw text extraction. 👇 Watch the demo 100
LlamaIndex @llamaindex.bsky.social · 20/03/2026Our new open-source LiteParse now comes with ready-to-use agent skills that work seamlessly with coding agents. `npx skills add run-llama/llamaparse-agent-skills --skill liteparse` 110
LlamaIndex @llamaindex.bsky.social · 19/03/2026We've spent years building LlamaParse into the most accurate document parser for production AI. Along the way, we learned a lot about what fast, lightweight parsing actually looks like under the hood. Today, we're open-sourcing a light-weight core of that tech as LiteParse 🦙 121
LlamaIndex @llamaindex.bsky.social · 18/03/2026Context engineering is the new prompt engineering — and if you're building AI agents, you need to understand the difference and why parsing your data correctly sits at the heart of it 220
LlamaIndex @llamaindex.bsky.social · 18/03/2026LlamaParse Agentic Plus mode now delivers precise visual grounding with bounding boxes for the most challenging document elements. Our latest update brings major improvements to how we handle complex visual content: 100
LlamaIndex @llamaindex.bsky.social · 17/03/2026One of the hardest problems with document parsing is trust. How do you know the output actually corresponds to what's in the source? LlamaParse has visual grounding with bounding box citations for outputs, and it addresses exactly this. Two ways to use it: 110
LlamaIndex @llamaindex.bsky.social · 16/03/2026Agentic AI transforms document extraction from simple text transcription into intelligent reasoning, dramatically reducing manual review queues and maintenance overhead. 100
LlamaIndex @llamaindex.bsky.social · 13/03/2026Choosing between Skills and MCP tools for your AI agents? Here's an overview from @cle-does-things.bsky.social and @tuana.dev 🔧 MCP tools offer deterministic API calls with fixed schemas - perfect for precise, predictable operations but require dev knowledge and introduce network latency 121
LlamaIndex @llamaindex.bsky.social · 12/03/2026Ever wondered what we mean by 'agentic' OCR? It's parsing that reasons about documents instead of just reading them. Agentic OCR adapts to layout changes by treating document processing as a goal-oriented task rather than simple text extraction. 110
LlamaIndex @llamaindex.bsky.social · 11/03/2026🔎 semtools v3.0.0 is out, and it's a great step forward for anyone using semantic search and document parsing from the command line. 100
LlamaIndex @llamaindex.bsky.social · 10/03/2026🚀 The team at @GoogleDeepMind just released Gemini Embedding 2, a frontier embeddings model with 3072 dimensions and state-of-the-art semantic quality. 141
LlamaIndex @llamaindex.bsky.social · 09/03/2026If you’re working with lots of slide decks and need a better way to search through them, Surreal Slides makes it simple 🌀 161
LlamaIndex @llamaindex.bsky.social · 06/03/2026PDFs are the bane of every AI agent's existence: here's why parsing them is so much harder than you think 📄 Every developer building document agents eventually hits the same wall: PDFs weren't designed to be machine-readable. They're drawing instructions from 1982, not structured data. 110
LlamaIndex @llamaindex.bsky.social · 05/03/2026"Just send the PDF to GPT-4o" Ok. We did. Here's what happened: • Reading order? Wrong. • Tables? Half missing. • Hallucinated data? Everywhere. • Bounding boxes? Nonexistent. • Cost at 100K pages? Brutal. So we're doing it live. 100
LlamaIndex @llamaindex.bsky.social · 05/03/2026Creating agent workflows and architecting the logic is one thing, making them durable and fail-safe is another👇 New integration for durable agent workflows with @dbos.dev execution - Make sure your agents survive crashes, restarts, and errors without writing any checkpoint code. 145
LlamaIndex @llamaindex.bsky.social · 04/03/2026Huge thank you to everyone who joined the Google DeepMind hackathon in NYC with us over the weekend 💛 110
LlamaIndex @llamaindex.bsky.social · 04/03/2026If you need to split complex or composite documents into structured categories or sections, LlamaSplit is built for the job ✂️ 100
LlamaIndex @llamaindex.bsky.social · 03/03/2026LlamaIndex has evolved far beyond a RAG framework - we're now focused on agentic document processing that automates knowledge work. 🚀 Agent orchestration has fundamentally changed with sophisticated reasoning loops, tool discovery through Skills/MCP, and coding agents that write Python for you 100
LlamaIndex @llamaindex.bsky.social · 02/03/2026When you parse a document with LlamaParse, you also get access to layout data for figures, charts, etc. Parse the document, specify to save layout images, and access those images on the response! Each image will be a cropped screenshot of that specific layout element. 110
LlamaIndex @llamaindex.bsky.social · 27/02/2026Turn your PDF charts into pandas DataFrames with specialized chart parsing in LlamaParse! This tutorial walks you through extracting structured data from charts and graphs in PDFs, then running data analysis with pandas - no manual data entry required. 110
LlamaIndex @llamaindex.bsky.social · 26/02/2026Build a private equity deal sourcing agent that automatically classifies investment opportunities and extracts key financial metrics using our LlamaAgents Builder. This step-by-step guide shows you how to create an agent that processes deal files like teasers and financial summaries: 100
LlamaIndex @llamaindex.bsky.social · 24/02/2026Document OCR benchmarks are hitting a ceiling - and that's a problem for real-world AI applications. 110
LlamaIndex @llamaindex.bsky.social · 23/02/2026🚀 LlamaAgents Builder just leveled up: File uploads are here! Our natural language interface for building agentic document workflows now supports file uploads. 230
LlamaIndex @llamaindex.bsky.social · 20/02/2026🚀 Big drop from Google DeepMind: Gemini 3.1 Pro is here, and we built a hands-on demo powered by LlamaCloud to put it to work and turn your receipt photos into real financial insights! 121
LlamaIndex @llamaindex.bsky.social · 19/02/2026More reasoning doesn't always mean better results - especially for document parsing. We tested GPT-5.2 at four reasoning levels on complex documents and found that higher reasoning actually hurt performance while dramatically increasing costs and latency. 100