Tag
This paper evaluates large language models' ability to verify citation support in legal documents, finding that while models detect wrong-case citations effectively, they struggle with pinpoint page references, often confusing topical relevance with precise support.