AskChem: Claim-Centered Infrastructure for Chemistry Literature Synthesis
Summary
AskChem is a claim-centered infrastructure for chemistry literature synthesis, converting papers into atomic, provenance-carrying claims indexed from 147K papers, with a web interface and API access for AI agents.
View Cached Full Text
Cached at: 07/31/26, 05:53 AM
Paper page - AskChem: Claim-Centered Infrastructure for Chemistry Literature Synthesis
Source: https://huggingface.co/papers/2607.28618
Abstract
Chemistryliteraturesynthesisoftenrequiresassemblingspecificfindingsscatteredacrossmanypublications,yetexistingliterature-searchsystemsprimarilyreturnrankeddocumentlists.Asaresult,scientistsandAIagentsneedtolocaterelevantinformation,verifytheirprovenance,andassemblecross-paperanswersmanually.WepresentAskChem,aclaim-centeredinfrastructureforcross-paperchemistrysearch.AskChemchangestheunitofretrievalfromthepapertotheprovenance-carryingclaim:eachpaperisconvertedintoatomic,typedclaims,eachgroundedbyasourceDOIandaverbatimquoteoranexplicitevidencelocator.Overthissharedclaimstore,AskChemexposescomplementarystructuresforsearchandsynthesis:astabilizedfacetedtaxonomyforhierarchicalretrievalandbrowsing,anevidencegraphlinkingclaimsthroughrelations,andanexploratorylivingtaxonomythatsituatesindexedpapersunderscientificprinciples.AskChemcurrentlyindexes2.4Mclaimsfrom147Kpapersandprovidesawebinterface,aswellasREST,SDK,andMCPaccessforAIagents.OnAskChem-Bench,groundingaGPT-5.5readerinAskChemyields100%resolvableDOIs,comparedwith88.3%withoutretrieval,andthehighestcitationdensityamongfivetestedsystems.AskChemisliveathttps://askchem.org.
View arXiv pageView PDFProject pageGitHub2Add to collection
Get this paper in your agent:
hf papers read 2607\.28618
Don’t have the latest CLI?curl \-LsSf https://hf\.co/cli/install\.sh \| bash
Models citing this paper0
No model linking this paper
Cite arxiv.org/abs/2607.28618 in a model README.md to link it from this page.
Datasets citing this paper1
#### bing-yan/askchem Updatedabout 1 hour ago • 139
Spaces citing this paper0
No Space linking this paper
Cite arxiv.org/abs/2607.28618 in a Space README.md to link it from this page.
Collections including this paper0
No Collection including this paper
Add this paper to acollectionto link it from this page.
Similar Articles
Evaluating and Guarding Citation Faithfulness in Agentic Scientific Synthesis
This paper evaluates citation faithfulness in agentic scientific synthesis systems, showing that current verifiers are unreliable with unsupported-citation rates varying from 3% to 18% depending on strictness. It proposes a gold-anchored evaluation protocol and a deployable guard that uses split-conformal prediction to provide a distribution-free bound on truly unsupported citations.
AI lets chemists design molecules by simply describing them
EPFL researchers developed Synthegy, an AI framework that uses large language models to guide chemical retrosynthesis and reaction mechanism analysis through natural language instructions, significantly improving strategic planning for chemists.
Evidence-Ledger Adjudication for Claim-Evidence Traceability
This paper introduces evidence-ledger adjudication, a workflow for claim-evidence traceability in AI-assisted writing, evaluated on a blind benchmark from AVeriTeC, CLIMATE-FEVER, and SciFact, showing agent-based methods outperform baselines.
AISE-Bench: A Full-Cycle Curated Benchmark for Information Seeking on Academic Knowledge Graphs
This paper introduces AISE-Bench, a curated benchmark with 1,133 QA pairs for evaluating LLM agents on multi-step API planning and grounded summarization for academic knowledge graphs. The benchmark reveals that even the strongest model achieves only moderate performance, highlighting challenges in stepwise correctness and traceable reasoning.
Agentic generation of verifiable rules for deterministic, self-expanding reaction classification
This paper presents a multi-agent LLM pipeline that automatically generates and verifies reaction rules for chemical synthesis, expanding a standard taxonomy from 68 to 14,073 classes without human curation, achieving 97.7% classification accuracy on unseen reactions.