evidence-chain

Tag

Cards List
#evidence-chain

Agent Safety Should Be a Runtime Contract

Hugging Face Daily Papers · 4d ago Cached

This paper argues that AI agent safety should be enforced at runtime via preventive controls and verifiable evidence, rather than relying solely on training-time alignment. It grounds the position in audits of safety incidents, false completions, trajectory schemas, and publication trends.

0 favorites 0 likes
#evidence-chain

Calibrated Selective Fact-Checking via Evidence Chain Evaluation

arXiv cs.AI · 2026-07-22 Cached

This paper introduces Evidence Chain Evaluation (ECE), a selective fact-checking framework that allows LLM-based verification agents to abstain from giving verdicts when evidence is weak, sparse, or inconsistent. On ECE-Bench, ECE achieves 97.8% selective accuracy at 93.7% coverage, demonstrating a safety-oriented trade-off for handling epistemically weak evidence.

0 favorites 0 likes
#evidence-chain

@Apodex_AI: Meet 𝗔𝗽𝗼𝗱𝗲𝘅 𝟭.𝟬 — a heavy-duty agent team for deep research, which sets the SOTA! The team searches the web, re…

X AI KOLs Timeline · 2026-06-08 Cached

Apodex 1.0 is a heavy-duty AI agent team for deep research that achieves state-of-the-art performance by searching the web, reasoning over evidence, and producing reports with verifiable evidence chains.

0 favorites 0 likes
← Back to home

Submit Feedback