novelty-assessment

Tag

Cards List
#novelty-assessment

Do Methods Support the Claims? Intra-Paper Verification for Peer Review

arXiv cs.CL · 2026-07-30 Cached

This paper introduces intra-paper claim verification, a framework that uses LLMs to evaluate whether novelty claims in a paper are supported by its methodological evidence, addressing a gap in existing automated peer review systems. Human evaluation shows significant alignment with human reviewer concerns, especially for novelty-related issues.

0 favorites 0 likes
#novelty-assessment

On the Limits of LLM-as-Judge for Scientific Novelty Assessment

Hugging Face Daily Papers · 2026-06-10 Cached

This paper introduces RQ-Bench, a benchmark to evaluate LLMs' ability to assess the novelty of scientific research questions. It finds that LLM judges consistently rate generated questions as more novel than human experts do, raising concerns about the reliability of using LLMs for scientific novelty evaluation.

0 favorites 0 likes
← Back to home

Submit Feedback