fact-checking

Tag

Cards List
#fact-checking

PopUpFactCheck: It even breaks the news since airtime!!! (QUALITY BREAKTHROUGH)

Reddit r/artificial ↗ · 2d ago

PopUpFactCheck.com improved its fact-checking quality by switching the underlying GPT-OSS-120B model to high reasoning effort, enhancing performance on attribution and judgment tasks while using caching and cost-effective routing to manage expenses.

0 favorites 0 likes
#fact-checking

ARAFA: An LLM-Generated Arabic Fact-Checking Dataset

arXiv cs.CL ↗ · 3d ago Cached

This paper introduces Arafa, a large-scale Arabic fact-checking dataset generated using LLMs, aimed at addressing the scarcity of resources for automatic fact-checking in Arabic.

0 favorites 0 likes
#fact-checking

@andy_matuschak: Playing with programmable highlighters, e.g. * green highlight on a citation -> finds and prints the cited paper * oran…

X AI KOLs Following ↗ · 2026-09-15 Cached

Andy Matuschak shares an experiment with programmable highlighters that automate citation finding, fact-checking, and interaction with AI agents.

0 favorites 0 likes
#fact-checking

DARE: Dialectical Agentic Reasoning for Structured Knowledge Fact Checking

arXiv cs.CL ↗ · 2026-09-15 Cached

DARE is a dialectical agentic reasoning framework for structured knowledge fact checking that uses an iterative retrieve–reason–reflect process, achieving 88.12% accuracy with an 8B model and matching GPT-4o performance.

0 favorites 0 likes
#fact-checking

R2VC: Modular Fact-Checking with Retrieval, Verification, and Confidence Calibration

arXiv cs.CL ↗ · 2026-09-14 Cached

The article presents R2VC, a modular fact-checking system that combines retrieval, verification, and confidence calibration to improve accuracy and reliability in automated fact checking, achieving a 13.74% accuracy increase on the FEVER benchmark.

0 favorites 0 likes
#fact-checking

CWF: A Collaborative Writing Framework for Personalized and Reliable Popular Science Writing

arXiv cs.AI ↗ · 2026-09-10 Cached

This paper introduces a collaborative writing framework for personalized and reliable popular science writing, featuring a new dataset, benchmark, and multi-agent fact-checking mechanism that achieves state-of-the-art performance.

0 favorites 0 likes
#fact-checking

Chalked for Mac

Product Hunt ↗ · 2026-09-03 Cached

Chalked for Mac is a productivity tool that prepares replies based on live calendar and sourced facts, allowing users to insert them with the Tab key.

0 favorites 0 likes
#fact-checking

Full Fact analysis shows AI chatbots spouting misinformation about AI-generated images, wars and royal fall outs

Reddit r/ArtificialInteligence ↗ · 2026-09-03 Cached

Full Fact's analysis shows AI chatbots like ChatGPT, Gemini, and Grok frequently generate misinformation when responding to false claims, including about AI-generated images and wars, highlighting their unreliability for fact-checking.

0 favorites 0 likes
#fact-checking

A third of Perplexity's citations don't contain the number they're cited for

Hacker News Top ↗ · 2026-09-02 Cached

An audit of Perplexity's AI search models found that over a third of citations for specific figures do not contain those numbers on the linked pages, with failures including inaccessible or irrelevant sources.

0 favorites 0 likes
#fact-checking

CoVer: Conflict-Aware Claim Verification

arXiv cs.AI ↗ · 2026-09-02 Cached

CoVer is a factual adjudication framework for addressing evidence-level and aggregation-level conflicts in claim verification, with strong performance evaluated on the ContraNote dataset from X's Community Notes system.

0 favorites 0 likes
#fact-checking

Built an open-source fact-checker for AI agents, it won't let a claim through unless it can actually back it up

Reddit r/AI_Agents ↗ · 2026-09-01

Built an open-source fact-checker for AI agents that verifies claims by fetching real-time sources and providing truth and confidence scores, useful for both public and internal documents.

0 favorites 0 likes
#fact-checking

A hallucination class that passes fact-checking: the claim is true and the quotation marks are fabricated

Reddit r/ArtificialInteligence ↗ · 2026-08-30

The article describes a type of AI hallucination where claims are accurate but quotations are fabricated, evading standard fact-checking, and discusses implementation challenges in detecting such errors.

0 favorites 0 likes
#fact-checking

ElementCheck: Complexity-Aware Long-Form Text Factuality Evaluation via Sentence Elements

arXiv cs.CL ↗ · 2026-08-28 Cached

ElementCheck is a complexity-aware framework that improves long-form text factuality evaluation by extracting sentence elements and organizing them into an element graph, accompanied by the new benchmark FastFact-Sent for enhanced verification accuracy.

0 favorites 0 likes
#fact-checking

Generating Biomedical Fact-Checking Reports with RL-Enhanced Agentic Search

arXiv cs.AI ↗ · 2026-08-26 Cached

This paper introduces BioCheck Agent, an LLM-based agent that generates structured biomedical fact-checking reports using RL-enhanced agentic search, showing improved accuracy and reduced hallucinations compared to base models.

0 favorites 0 likes
#fact-checking

Yesterday I put ChatGPT, Claude and Gemini in a group chat. Now I want Reddit to break it

Reddit r/artificial ↗ · 2026-08-25

A user tested ChatGPT, Claude, and Gemini in a group chat for mutual fact-checking to catch hallucinations, and invites Reddit to provide challenging prompts to find shared blind spots.

0 favorites 0 likes
#fact-checking

NepOOC-M: Bilingual Nepali-English Benchmark and Comparative Analysis of Multimodal Architectures for OOC Detection

arXiv cs.CL ↗ · 2026-08-21 Cached

The paper introduces NepOOC, the first public Nepali-dominant multilingual benchmark for out-of-context misinformation detection, and evaluates multimodal architectures, finding that text-only models achieve strong performance.

0 favorites 0 likes
#fact-checking

Whether LLMs Can Navigate Beliefs and Facts Depends on How You Phrase It

arXiv cs.CL ↗ · 2026-08-19 Cached

Research shows that large language models' ability to confirm user beliefs depends on phrasing, with accuracy varying across epistemic expressions due to task confusion where models default to fact-checking.

0 favorites 0 likes
#fact-checking

Hallucination Span Detection with Input-Side Evidence Alignment

arXiv cs.CL ↗ · 2026-08-18 Cached

This paper introduces a task for hallucination span detection in LLMs by aligning output tokens with input evidence, proposing an encoder-based model that uses prediction confidence to detect hallucinations without manual alignments.

0 favorites 0 likes
#fact-checking

ReflectFact: Self-Reflective Agents for Improving Comprehension and Reasoning in Multi-Hop Fact Verification

arXiv cs.AI ↗ · 2026-08-14 Cached

ReflectFact is a self-reflective agent framework for multi-hop fact verification that addresses objective and knowledge conflicts via reasoning path planning, evidence-drift verification, and reasoning reflection, achieving state-of-the-art results on HOVER and EX-FEVER.

0 favorites 0 likes
#fact-checking

Decomposition-Induced Context-Memory Conflict: When Fact-Checking Pipelines Contradict Their Own Source Text

arXiv cs.CL ↗ · 2026-08-12 Cached

This paper identifies and characterizes Decomposition-Induced Context-Memory Conflict (DI-CC), a failure mode in decompose-then-verify pipelines where decomposition substitutes the model's parametric beliefs for the source text. The authors show it is mechanistically related to classical context-memory conflict, that SelfCheckGPT fails to detect it, and that context-aware decoding suppresses it but introduces severe parsing failures.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback