Tag
Hermes Agent v0.20.0 introduces grounded citations that link every research claim to a specific source and passage, plus a fact-checking mode that categorizes claims as confirmed, unverified, or contradicted.
This paper benchmarks 8 LLM judges for citation quality in deep-research systems, finding that cheaper models remain competitive with frontier models on source relevance and factual support, but differ in directional bias which matters for RL training.
A product manager shares hands-on testing of Minimax M3's 1M context window on a real Q3 strategic brief, noting strong source attribution up to ~200K tokens but synthesis degradation beyond that.