Tag
An analysis of the top 50 most-cited AI researchers, highlighting how landmark papers like Attention is All You Need have shaped the field and whose work drives AI's influence.
该推文盘点了 10 个用于让 AI 智能体为其输出提供可验证证据的开源 GitHub 项目,涵盖科学文献检索(OpenScholar、PaperQA2)、文档解析(Docling)、网络爬取(Crawl4AI)、RAG(RAGFlow、GraphRAG、LightRAG)、深度研究(Open Deep Research)以及事实核查与评估(DeepEval、Ragas),并重点展示了 Allen AI 的 OpenScholar 代码库。
LlamaIndex introduces Extract v2.5, a set of frontier-agent document extraction models (Cost Effective, Agentic, Agentic Plus) that now claim SOTA price-performance on ExtractBench, outperforming Opus 5.5 and GPT-6 Sol at 30%-4x lower cost, with major accuracy gains on long lists, multi-page records and scanned forms, plus new Advanced Citations and Structural Reasoning features available on LlamaParse.
This paper investigates the role of citations in human and LLM preferences for scientific question answering, finding that humans prefer diverse citations but fewer overall, while LLMs exhibit stronger citation-related preferences despite lacking source access.
The article introduces a website called Second City Citation Source that allows users to search for and view detailed information about parking enforcement officers in Chicago.
This paper models the strategic interaction between content providers and generative search engines as a repeated Stackelberg game, showing how GEO can escalate into citation wars, and proposes a verifiable-content reward mechanism (VCR) to align incentives and achieve win-win outcomes.
An experiment found that Perplexity AI search frequently cited low-view YouTube videos (some with affiliate links) in product answers, suggesting that AI visibility does not equal popularity and raising concerns about source quality.
A US federal appeals court sanctioned two lawyers for filing briefs with AI-made-up cases, underscoring the persistent issue of AI hallucinations in legal practice and the necessity of independent citation verification.
Tips on how to get cited in Google AI Overviews by analyzing keyword gaps using Outrank's free AI Overview Tracker.
OpenBioRQ is a new benchmark of 12,553 unsolved biomedical research questions that tests agentic models' ability to verify sources and avoid false citations. It reveals that current models often link to wrong papers and suffer from agentic collapse on hard questions.
This project adds an auditable academic research pipeline to Claude Code, including checkpoints such as citation verification and experiment claim alignment, ensuring the credibility of research outputs.
The author is building CLYCITE, a search engine that grounds answers in retrieved sources, provides citations, and publicly publishes its accuracy rates by category. They seek community feedback on whether a public accuracy dashboard and an ad-free subscription model would be valuable.
Tim O'Reilly discusses the challenges of integrating AI into scientific publishing, including hallucinated citations, propagation of retracted papers, and training on compromised literature, and calls for adapting existing scientific infrastructure for AI use.
Google unveils an autonomous research system that can search the web, reason across sources, use tools via MCP, generate charts, and produce cited reports.