We’ve been analyzing how people are using LLMs for legal and compliance tasks (GDPR, AI Act, etc.).

Reddit r/ArtificialInteligence News

Summary

Analysis of LLM usage in legal and compliance tasks reveals that models often produce confident but unverifiable citations, raising questions about reliable legal grounding for AI outputs.

One thing keeps coming up: Models often produce very confident legal answers, but the citations are either missing, outdated, or not traceable to official consolidated sources. Even when they cite something, it’s not always clear if it reflects the current version of the law. I’m curious how others handle this in practice: Do you rely on manual verification every time? Or is there already a reliable way teams are solving the “verifiable legal grounding” problem for AI outputs?
Original Article

Similar Articles

DLawBench: Evaluating LLMs Through Multi-Turn Legal Consultation

arXiv cs.CL

DLawBench is a new benchmark for evaluating large language models in multi-turn legal consultation, covering Chinese and US law with four client types. Experiments show significant room for improvement, with the best model achieving only 0.562 on legal reasoning.