document-processing

Tag

Cards List
#document-processing

Splitting Documents at Lower Cost: Multi-Split Boundary Decisions for LLM-Based Page Stream Segmentation

arXiv cs.AI · yesterday Cached

This paper introduces Multi-Split Boundary Decision (MSBD) to reduce inference costs in zero-shot page stream segmentation using large language models, demonstrating improved efficiency while maintaining accuracy for appropriate window sizes.

0 favorites 0 likes
#document-processing

Turn PDF into podcast show

Reddit r/artificial · 2d ago

PaperPod is a service that converts PDF documents into podcast episodes, generating a 20-minute two-host episode in about one minute without requiring GPU or studio equipment.

0 favorites 0 likes
#document-processing

@grgerwcwetwet: Here's a highly recommended Agent Skill for party learning that's totally worth installing: learn-from-materials. It di…

X AI KOLs Timeline · 5d ago Cached

The learn-from-materials Agent Skill converts PDFs, EPUBs, Word docs, and PPTs into interactive learning websites with source tracing and knowledge base building.

0 favorites 0 likes
#document-processing

Typst makes big strides

Lobsters Hottest · 6d ago Cached

Typst, a typesetting system positioned as a LaTeX replacement, has released version 0.15 with new features including support for variable fonts, MathML, and multiple bibliographies.

0 favorites 0 likes
#document-processing

From Pixels to Pairs: A Comprehensive Benchmark of LLM-Based Key-Value Extraction in Noisy Document Settings

arXiv cs.CL · 2026-09-17 Cached

This paper introduces a comprehensive benchmark for evaluating LLMs in key-value extraction from documents under OCR noise, revealing substantial performance degradation and emphasizing the need for joint optimization of OCR quality and LLM reasoning.

0 favorites 0 likes
#document-processing

Evidence-Aligned Local Composition of Discrete Experts for Sequence Restoration

arXiv cs.AI · 2026-09-10 Cached

The paper introduces evidence-aligned local composition, a method for restoring corrupted documents by inferring soft weightings over frozen domain experts from the marginal evidence of the corrupted observation, achieving high accuracy in tracking expert regions without labeled data.

0 favorites 0 likes
#document-processing

@GithubProjects: TextGen runs powerful AI models on your own computer, fully offline and private. - No tracking, no internet needed, no …

X AI KOLs Timeline · 2026-09-02 Cached

TextGen is a tool that enables users to run powerful AI models locally on their computers with full privacy, offline access, and support for text, images, and documents through a simple download and setup.

0 favorites 0 likes
#document-processing

@weaviate_io: We stopped asking PDFs to become text before we searched them. With late-interaction multi-vector retrieval, you can em…

X AI KOLs Following · 2026-09-01 Cached

Weaviate introduces a method to search PDFs without text extraction by embedding each page as an image using late-interaction multi-vector retrieval, demonstrated on NVIDIA investor decks.

0 favorites 0 likes
#document-processing

Building a high-accuracy semantic evidence/RAG system for financial documents — looking for feedback

Reddit r/AI_Agents · 2026-08-26

The article outlines a multi-gate architecture for a high-accuracy semantic evidence and RAG system for financial documents, emphasizing traceability, reconciliation, and hybrid retrieval, and seeks feedback on its design.

0 favorites 0 likes
#document-processing

@trueventures: Coming to Connected Stack in SF — @jerryjliu0, CEO & Co-founder of @llama_index, nominated by @GreylockVC. LlamaIndex i…

X AI KOLs Timeline · 2026-08-24 Cached

The Connected Stack Conference in San Francisco will feature AI industry leaders from companies like Anthropic and LlamaIndex, focusing on enterprise AI and document infrastructure for AI agents.

0 favorites 0 likes
#document-processing

@jerryjliu0: The latest RAG trend for the current agent harnesses (Codex, Cowork) is to do two passes of document processing to solv…

X AI KOLs Timeline · 2026-08-23 Cached

The article discusses a two-pass document processing trend for AI agents, where a fast OSS pass enables efficient retrieval and a VLM-based pass ensures accuracy, promoting Llama Index's LiteParse and LlamaParse tools to enhance cost and performance.

0 favorites 0 likes
#document-processing

@jerryjliu0: Our "agentic plus" extractor in LlamaParse is great for extracting out massive volumes of fields (e.g. 10k-100k+ fields…

X AI KOLs Following · 2026-08-15 Cached

Jerry Liu introduces ExtractBench and highlights the 'agentic plus' extractor in LlamaParse for handling massive volumes of fields in long documents, with benchmark results available on ExtractBench.

0 favorites 0 likes
#document-processing

@jerryjliu0: We're not Palantir, but we do think a lot about evals and hillclimbing w.r.t. document processing. If you have really h…

X AI KOLs Following · 2026-08-09 Cached

Jerry Liu promotes LlamaParse and LlamaAgents for large-scale document extraction, emphasizing LLM evals and hillclimbing for accuracy and cost. He also connects FDE work with evals and RL environments.

0 favorites 0 likes
#document-processing

@DanKornas: Traditional text-focused RAG systems cannot effectively process the images, tables, equations, charts, and multimedia f…

X AI KOLs Timeline · 2026-08-03 Cached

RAG-Anything is a multimodal document-processing RAG system built on LightRAG that parses documents, constructs a multimodal knowledge graph, and uses hybrid vector-graph retrieval to answer queries.

0 favorites 0 likes
#document-processing

Xberg v1 is out

Reddit r/LocalLLaMA · 2026-08-02

Xberg v1 is released as the successor to Kreuzberg, a high-performance content intelligence framework supporting 101 document formats, 367 code/data types, audio/video transcription, and URL ingestion, with pure-Rust PDF and OCR backends, multiple language bindings, and mobile/WASM support.

0 favorites 0 likes
#document-processing

IDP AutoOpt: Agent-Driven Optimization of Document Processing Pipeline Configurations

arXiv cs.AI · 2026-07-31 Cached

This paper presents IDP AutoOpt, an autonomous LLM agent that optimizes intelligent document processing pipeline configurations, matching or exceeding human-expert accuracy at lower cost and reducing configuration time from weeks to under two hours.

0 favorites 0 likes
#document-processing

@jerryjliu0: We’re excited to rollout an official batch parsing experience to LlamaParse. Instead of hitting our APIs one file at a …

X AI KOLs Following · 2026-07-30 Cached

LlamaIndex announced an official batch parsing experience for LlamaParse, allowing users to parse up to 10,000 files at once through a dedicated UI with batch auditing and failure inspection, removing the need for custom async scripts.

0 favorites 0 likes
#document-processing

@PrajwalTomar_: A fully offline AI just read over 4,000 pages of declassified UFO files and answered questions about them with citation…

X AI KOLs Timeline · 2026-07-30 Cached

A fully offline AI reads over 4,000 pages of declassified UFO files using OCR and vector database, providing cited answers locally without cloud or API keys.

0 favorites 0 likes
#document-processing

DocAnnot -- Accelerating the Creation of Key Information Extraction Datasets with GenAI-Powered Auto-annotation

arXiv cs.CL · 2026-07-29 Cached

Introduces DocAnnot, a framework that uses a large vision-language model, OCR, and a spatially informed contextual matching algorithm to automatically generate training datasets for key information extraction from documents, reducing manual annotation effort. Evaluated on CORD and SROIE benchmarks, it achieves reasonable F1-scores and enables efficient human verification.

0 favorites 0 likes
#document-processing

We Used AI to Reduce a 50–70 Hour Manual Document Sorting Process to Around 3–5 Hours

Reddit r/ArtificialInteligence · 2026-07-28

A team used AI to automate a manual document sorting process, reducing labor from 50-70 hours to 3-5 hours per month by grouping scanned pages into documents and generating PDFs.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback