knowledge-extraction

Tag

Cards List
#knowledge-extraction

NLP-Driven Knowledge Extraction and Thematic Classification of Translated Ancient Indian Medical Texts

arXiv cs.CL · 2026-09-01 Cached

This paper utilizes NLP techniques such as NER and BERTopic, along with Neo4j, to extract, classify, and visualize knowledge from translated ancient Indian medical texts, enhancing accessibility and digital preservation.

0 favorites 0 likes
#knowledge-extraction

@tom_doerr: HyperExtract converts unstructured documents into structured knowledge graphs, hypergraphs, and lists using large langu…

X AI KOLs Timeline · 2026-08-27 Cached

HyperExtract is an LLM-powered framework that converts unstructured documents into structured knowledge graphs, hypergraphs, and lists, simplifying knowledge extraction and management.

0 favorites 0 likes
#knowledge-extraction

HERMES: a multi-agent framework for structured knowledge extraction from ultra-long documents in geoscience

arXiv cs.CL · 2026-08-17 Cached

HERMES is a scalable multi-agent framework for extracting structured knowledge from ultra-long scientific documents in geoscience, achieving high accuracy and sixfold efficiency improvement over manual methods.

0 favorites 0 likes
#knowledge-extraction

I've watched the same Ray Dalio video 11 times and remember almost nothing. So I turned it into a Claude skill

Reddit r/AI_Agents · 2026-08-12

The author describes how they built a Claude skill to extract Ray Dalio's Big Cycle framework from a video transcript, using a structured extraction prompt and organizing the output into a SKILL.md. The article shares their step-by-step process for turning video knowledge into runnable AI workflows.

0 favorites 0 likes
#knowledge-extraction

@RituWithAI: Someone just built a tool that turns any book into a skill your AI agent can use forever. One command. Any PDF, ePub, o…

X AI KOLs Timeline · 2026-07-29 Cached

book-to-skill is an open-source tool that converts PDFs, ePubs, and long documents into structured SKILL.md files for AI coding agents like Claude Code and Cursor, enabling agents to permanently apply book knowledge during coding sessions.

0 favorites 0 likes
#knowledge-extraction

@tom_doerr: AI-powered page-by-page PDF knowledge extraction and summarization https://github.com/echohive42/AI-reads-books-page-by…

X AI KOLs Timeline · 2026-06-30 Cached

An AI-powered tool for extracting knowledge and generating summaries from PDF books page by page.

0 favorites 0 likes
#knowledge-extraction

@Jolyne_AI: A Python script that can automatically read PDF books: AI Reads Books. Drop a PDF in, run it to parse content page by page, capture key knowledge points, and automatically generate a well-structured Markdown summary. GitHub: https://github.com…

X AI KOLs Timeline · 2026-06-28 Cached

A Python script that can automatically parse PDF book content, extract key knowledge points, and generate Markdown-format summaries, aiming to improve reading and knowledge organization efficiency.

0 favorites 0 likes
#knowledge-extraction

@GitHub_Daily: For those in quantitative research, daily facing massive financial reports and cutting-edge papers, manually filtering valuable content is like finding a needle in a haystack. Recently discovered an open-source project called QuantMind, focused on intelligent knowledge extraction and retrieval for quantitative finance. It can automatically fetch papers, news, blogs, and turn unstructured documents into searchable...

X AI KOLs Timeline · 2026-06-12 Cached

QuantMind is an open-source framework for intelligent knowledge extraction and retrieval in quantitative finance. It can automatically fetch unstructured content like papers and news, build a queryable structured knowledge base, and support natural language retrieval.

0 favorites 0 likes
#knowledge-extraction

Automatic Construction of a Legal Citation Graph from 100 Million Ukrainian Court Decisions: Large-Scale Extraction, Topological Analysis, and Ontology-Driven Clustering

arXiv cs.CL · 2026-05-18 Cached

This paper constructs the first large-scale citation graph from 100.7 million Ukrainian court decisions, extracting over 500 million citation links. It demonstrates that the citation structure can automatically recover legal domain boundaries and predict legislative importance with near-perfect accuracy, and releases the pipeline and data as open resources.

0 favorites 0 likes
#knowledge-extraction

Experience Compression Spectrum: Unifying Memory, Skills, and Rules in LLM Agents

arXiv cs.CL · 2026-04-20 Cached

This paper proposes the Experience Compression Spectrum, a unifying framework that integrates agent memory, skill discovery, and rule-based systems along a single axis of increasing compression (5-20× for episodic memory, 50-500× for procedural skills, 1000×+ for declarative rules). The work identifies a critical gap—the 'missing diagonal'—showing that existing systems operate at fixed compression levels without adaptive cross-level support, and articulates design principles for scalable, full-spectrum agent learning systems.

0 favorites 0 likes
#knowledge-extraction

yifanfeng97/Hyper-Extract

GitHub Trending (daily) · 2026-06-18 Cached

Hyper-Extract is an open-source CLI tool that uses LLMs to extract structured knowledge from unstructured documents, supporting various output formats like knowledge graphs and hypergraphs.

0 favorites 0 likes
← Back to home

Submit Feedback