legal-nlp

Tag

Cards List
#legal-nlp

LegalPincite: Multi-level Legal Information Retrieval Dataset

Hugging Face Daily Papers · 5d ago Cached

Introduces LegalPincite, a large-scale legal information retrieval dataset built from CJEU judgments, featuring masked queries, full corpora, and paragraph-level citation annotations to enable multi-level retrieval evaluation.

0 favorites 0 likes
#legal-nlp

Semantics of Subterfuge: Benchmarking Legal Deception Detection Against General-domain State-of-the-Art

arXiv cs.CL · 6d ago Cached

This paper surveys and benchmarks NLP-based automatic deception detection in legal contexts, comparing fine-tuned transformers and seven LLMs with various prompting strategies across seven datasets. Results show domain sensitivity, with fine-tuned models excelling in general domains and few-shot LLMs competitive in low-resource legal settings.

0 favorites 0 likes
#legal-nlp

BLAD: A Historically Contextualized, Multilingual Dataset of Bangladeshi Legal Acts (1799 to 2025)

arXiv cs.CL · 2026-07-21 Cached

This paper introduces BLAD, a curated multilingual dataset of 1,484 Bangladeshi legal acts spanning 1799 to 2025, with structured metadata for temporal and cross-lingual legal NLP research.

0 favorites 0 likes
#legal-nlp

DECODEM: Data Extraction from Corporate Organizational Documents via Enhanced Methods

arXiv cs.CL · 2026-07-20 Cached

Introduces DECODEM, benchmark datasets for evaluating automated extraction of corporate governance variables from legal documents using large language models, showing high accuracy for many provisions.

0 favorites 0 likes
#legal-nlp

When Reasoning Hurts Legal Drafting: The Verbalization Bottleneck in Patent Claim Generation

arXiv cs.CL · 2026-07-14 Cached

This paper investigates whether Chain-of-Thought (CoT) prompting benefits patent claim generation, finding that implicit CoT (where reasoning is internal) consistently outperforms explicit CoT, which can introduce a verbalization bottleneck that compromises output quality through abstraction of details, disruption of patterns, and error propagation.

0 favorites 0 likes
#legal-nlp

PRecG: Legal Precedent Retrieval with Graph Neural Networks and Rhetorical Role Segmentation

arXiv cs.CL · 2026-07-13 Cached

PRecG proposes a novel pipeline for legal precedent retrieval by decomposing documents into rhetorical segments, building knowledge graphs for each segment, and learning hierarchical representations using graph neural networks. Experiments on Indian legal data show effectiveness over baselines.

0 favorites 0 likes
#legal-nlp

From Judgments to Issues: Structured Extraction of Legal Reasoning with Citation-Hallucination Control

arXiv cs.CL · 2026-07-07 Cached

This paper presents an automated pipeline that uses the DeepSeek V3 model to decompose Italian tax-court judgments into individual legal issues structured in XML following the IRAC framework, and includes a hallucination-detection filter using the Linkoln parser to validate citations, validated by expert annotators.

0 favorites 0 likes
#legal-nlp

A Tree-of-Thoughts Inspired Hybrid Approach for Legal Case Judgement Summarization using LLMs

arXiv cs.CL · 2026-06-29 Cached

Proposes a tree-of-thoughts inspired extractive-abstractive approach for legal case judgement summarization using LLMs, with experiments on DeepSeek and LLama showing improved summaries over extractive or abstractive methods alone.

0 favorites 0 likes
#legal-nlp

LAUKIN: A Multi-jurisdictional Common Law Contract Dataset

arXiv cs.CL · 2026-06-12 Cached

Introduces LAUKIN, a dataset of clause pairs from Australia, UK, and India contracts labeled for legal equivalence, and evaluates 12 models achieving 65.11% macro-F1, establishing a challenging benchmark.

0 favorites 0 likes
#legal-nlp

HKJudge: A Legal Discourse-Annotated Corpus for Interpreting What Courts Find, How They Reason, and What They Rule

arXiv cs.CL · 2026-06-08 Cached

HKJudge is the first sentence-level expert-annotated legal discourse corpus for Hong Kong criminal judgments, featuring a two-tier discourse schema and benchmark evaluations of BERT-based and LLM models.

0 favorites 0 likes
#legal-nlp

EURO-5K: When Does Domain Pretraining Matter? Benchmarking Transformers for EU Reporting Obligation Extraction

arXiv cs.CL · 2026-06-03 Cached

This paper introduces EURO-5K, a sentence-level dataset for extracting reporting obligations from EU legislation, and benchmarks discriminative and generative transformer models under full fine-tuning and parameter-efficient QLoRA. Results show that legal pretraining primarily benefits models with limited adaptation capacity, and all approaches converge around 3K samples.

0 favorites 0 likes
#legal-nlp

Enhancing BiGRU with a KAN Block for Legal Document Classification and Summarization

arXiv cs.CL · 2026-06-02 Cached

This paper introduces a KAN-enhanced BiGRU architecture for classifying and summarizing multilingual legal documents from Bangladesh, achieving modest accuracy and ROUGE scores and demonstrating that the KAN block improves classification accuracy over the baseline BiGRU.

0 favorites 0 likes
#legal-nlp

UA-Legal-Bench: A Benchmark for Evaluating Large Language Models on Ukrainian Legal Reasoning

arXiv cs.CL · 2026-05-29 Cached

Introduces UA-Legal-Bench, a five-task benchmark for evaluating large language models on Ukrainian legal reasoning, built from the Unified State Register of Court Decisions. Evaluates 11 LLMs, revealing task-dependent few-shot effects and the misleading nature of accuracy on imbalanced legal tasks.

0 favorites 0 likes
#legal-nlp

Generating Legal Commentaries from Case Databases via Retrieval, Clustering, and Generation

arXiv cs.CL · 2026-05-26 Cached

This paper presents a fully automated pipeline that transforms court decisions into legal commentaries by extracting, clustering, and summarizing paragraph-level chunks using LLMs, evaluated on German civil code cases.

0 favorites 0 likes
#legal-nlp

Temporal Concept Drift in Legal Judgment Prediction: Neural Baselines Across Three Epochs of Ukrainian Court Decisions

arXiv cs.CL · 2026-05-26 Cached

This paper investigates temporal concept drift in legal judgment prediction by fine-tuning transformer models on Ukrainian court decisions from three epochs defined by geopolitical disruptions. Findings show severe forward degradation, asymmetry in backward transfer, and that chronological continual learning effectively mitigates forgetting while domain pretraining reduces degradation magnitude.

0 favorites 0 likes
#legal-nlp

LP-Eval: Rubric and Dataset for Measuring the Quality of Legal Proposition Generation

arXiv cs.CL · 2026-05-20 Cached

This paper introduces LP-Eval, a rubric and dataset for evaluating legal proposition generation by large language models, with annotations by legal experts. Results show that rubric-guided LLM evaluations align more closely with expert assessments than direct scoring.

0 favorites 0 likes
#legal-nlp

IMLJD: A Computational Dataset for Indian Matrimonial Litigation Analysis

arXiv cs.CL · 2026-05-20 Cached

The paper introduces IMLJD, a computational dataset designed for analyzing Indian matrimonial litigation, supporting natural language processing and legal analytics research.

0 favorites 0 likes
#legal-nlp

Automatic Construction of a Legal Citation Graph from 100 Million Ukrainian Court Decisions: Large-Scale Extraction, Topological Analysis, and Ontology-Driven Clustering

arXiv cs.CL · 2026-05-18 Cached

This paper constructs the first large-scale citation graph from 100.7 million Ukrainian court decisions, extracting over 500 million citation links. It demonstrates that the citation structure can automatically recover legal domain boundaries and predict legislative importance with near-perfect accuracy, and releases the pipeline and data as open resources.

0 favorites 0 likes
#legal-nlp

A Few Good Clauses: Comparing LLMs vs Domain-Trained Small Language Models on Structured Contract Extraction

arXiv cs.CL · 2026-05-08 Cached

This paper compares a domain-trained small language model (Olava Extract) against frontier LLMs for structured contract extraction, showing that the specialized model achieves higher F1 scores and dramatically lower cost.

1 favorites 1 likes
← Back to home

Submit Feedback