@jerryjliu0: We've built a new feature in LlamaParse that lets you automatically extract any complex form into a structured JSON out…
Summary
LlamaIndex introduces a new LlamaParse feature that automatically extracts complex form fields into structured JSON without requiring a predefined schema, simplifying document form processing.
View Cached Full Text
Cached at: 08/07/26, 06:55 PM
We’ve built a new feature in LlamaParse that lets you automatically extract any complex form into a structured JSON output
The best part is there’s no schema needed! We will systematically detect and extract out every single form key and corresponding form value from the document (blank if not filled).
Simply set processing_options.forms='enrich'
Check it out: http://login.llamaindex.ai/sign-up API docs: https://developers.api.llamaindex.ai/api/resources/parsing/methods/get/#…
Sign in
OR
Don’t have an account?Sign up
Terms of ServiceandPrivacy Policy
LlamaIndex 🦙 (@llama_index): Parsing a W-2 into markdown was always the easy part. Getting the fields out was a second pipeline: define a schema, map the fields, handle the edge cases.
Set processing_options.forms=‘𝗲𝗻𝗿𝗶𝗰𝗵’, and LlamaParse returns a dedicated JSON with field names, values, and checkbox
Similar Articles
@jerryjliu0: we've built the world's most advanced engine for document extraction over complex documents the video below shows an ov…
LlamaIndex has launched LlamaExtract Agentic Plus, an advanced AI engine for extracting data from complex documents like long tables and forms, and is promoting their LlamaParse service for improved document processing.
@jerryjliu0: We've massively improved our document parsing capabilities across the board in the past ~3 months. Our LlamaParse cost-…
LlamaIndex has significantly improved LlamaParse over the past three months, achieving 10-20% better accuracy on complex tables, charts, and grounding while maintaining costs below 0.4 cents per page, as measured against their ParseBench benchmark.
@jerryjliu0: We built state-of-the-art models for reading forms Form documents have the following properties that trip up VLMs: They…
This blog post explains why VLMs struggle with form documents and introduces LlamaParse, a purpose-built tool for accurate and cost-effective form parsing.
@jerryjliu0: We’re excited to rollout an official batch parsing experience to LlamaParse. Instead of hitting our APIs one file at a …
LlamaIndex announced an official batch parsing experience for LlamaParse, allowing users to parse up to 10,000 files at once through a dedicated UI with batch auditing and failure inspection, removing the need for custom async scripts.
@jerryjliu0: Our "agentic plus" extractor in LlamaParse is great for extracting out massive volumes of fields (e.g. 10k-100k+ fields…
Jerry Liu introduces ExtractBench and highlights the 'agentic plus' extractor in LlamaParse for handling massive volumes of fields in long documents, with benchmark results available on ExtractBench.