legal-ai

Tag

Cards List
#legal-ai

Harvey turns legal context into stronger drafts with GPT-6 Astra

OpenAI Blog ↗ · 5d ago Cached

Harvey integrates GPT-6 Astra to enhance AI-driven legal document drafting, providing more context and structured outputs for law firms.

0 favorites 0 likes
#legal-ai

IntLawNER: A Named Entity Recognition Dataset and Benchmark in International Law

arXiv cs.AI ↗ · 5d ago Cached

IntLawNER is a new named entity recognition dataset and benchmark for international law, covering gold-annotated sentences from legal texts and evaluating model performance with few-shot improvements.

0 favorites 0 likes
#legal-ai

Beyond Accuracy and Surface Fluency: Risk-Sensitive Evaluation of LLMs for Legal Clause Generation

arXiv cs.CL ↗ · 6d ago Cached

This paper presents a risk-sensitive evaluation framework for LLM-generated contract clauses, focusing on legal failure modes and quality dimensions to assess risks beyond accuracy or fluency.

0 favorites 0 likes
#legal-ai

@rohanpaul_ai: Legal work saw one of Grok 4.7’s biggest jumps. On the Harvey Legal Agent Benchmark, Grok 4.7 scored 19.6%, while GPT-5…

X AI KOLs Following ↗ · 6d ago Cached

Grok 4.7 shows significant improvements in legal work and terminal tasks, outperforming competitors on benchmarks like Harvey Legal Agent Benchmark and Terminal-Bench 4.0, while keeping token prices unchanged.

0 favorites 0 likes
#legal-ai

@xeophon: Pacing the frontier

X AI KOLs Timeline ↗ · 6d ago Cached

Vals AI evaluated Grok 4.7, finding it ranks #24 on the Vals Index with a score of 54.2%, down from Grok 4.6, but shows improvements in legal and medical domains.

0 favorites 0 likes
#legal-ai

Before the Arrest: Benchmarking LLMs on Criminal Profiling from Incomplete Evidence

arXiv cs.CL ↗ · 2026-09-18 Cached

The paper introduces the PIJ benchmark for evaluating large language models on criminal profiling tasks from incomplete evidence, highlighting performance gaps and biases in inferential reasoning.

0 favorites 0 likes
#legal-ai

ChatGPT vs Gemini

Reddit r/ArtificialInteligence ↗ · 2026-09-18

A user compares ChatGPT and Gemini, finding Gemini superior in providing accurate legal reasoning and acknowledging logical flaws, while ChatGPT relies on circular arguments and avoids direct answers.

0 favorites 0 likes
#legal-ai

Legal LLM Hallucination Should Be Evaluated as Failure of Legal Warrant

arXiv cs.CL ↗ · 2026-09-17 Cached

This position paper argues that legal LLM hallucinations should be evaluated as failures of legal warrant rather than factual inaccuracies, proposing a new benchmark framework for assessing legal AI systems.

0 favorites 0 likes
#legal-ai

LexAgentHallu: A Hierarchical Benchmark for Profiling Hallucinations in Legal Agents

arXiv cs.AI ↗ · 2026-09-11 Cached

This paper introduces LexAgentHallu, a hierarchical benchmark for profiling hallucinations in legal AI agents, evaluating multi-step trajectories with fine-grained metrics.

0 favorites 0 likes
#legal-ai

LexIssue: Benchmarking Legal Issue Identification in Chinese Civil Litigation

arXiv cs.CL ↗ · 2026-09-04 Cached

This paper introduces LexIssue, a benchmark for identifying disputed legal issues in Chinese civil litigation, constructed from real cases and expert annotations, and shows that retrieval-augmented generation enhances performance.

0 favorites 0 likes
#legal-ai

OBJECTION! Lawyer Agents Mitigate Guilty Bias in Legal Judgment Prediction

arXiv cs.CL ↗ · 2026-09-03 Cached

The paper introduces OBJECTION, an inference-time pipeline using adversarial lawyer agents to mitigate guilty bias in legal judgment prediction models, demonstrating a significant reduction in false guilty rates and releasing a new 'Natural Innocent' dataset.

0 favorites 0 likes
#legal-ai

@grapeot: Same type of company, three years apart, opposite results. In 2023, Bloomberg trained a 50B model from scratch, fed with 363B tokens of private financial data, official conclusion: not used in any products. In 2026, Thomson Reuters spent $40 million, used...

X AI KOLs Timeline ↗ · 2026-08-29 Cached

This article compares the different strategies of Bloomberg and Thomson Reuters in AI model development, analyzes the shift from training large models from scratch to fine-tuning on open bases, and the impact of this trend on vertical AI applications.

0 favorites 0 likes
#legal-ai

Referent

Product Hunt ↗ · 2026-08-26 Cached

Referent is an AI-native legal practice management platform that combines AI, CRM, and workflow tools to help law firms streamline operations.

0 favorites 0 likes
#legal-ai

OpenAI-backed legal tech firm pivots to Chinese Kimi K3 open-weight model

Reddit r/artificial ↗ · 2026-08-21 Cached

Harvey, an OpenAI-backed legal tech firm, has pivoted to using Moonshot AI's Kimi K3 open-weight model to develop its first in-house model, Harvey Tenet, marking a shift towards Chinese open-weight AI systems in Western tech.

0 favorites 0 likes
#legal-ai

@FinanceYF5: How to do world-class AI research without the budget of a major model lab? Harvey shares his 'Moneyball' approach: let domain experts guide synthetic data, build legal evaluation sets; collaborate with multiple new labs, complete post-training internally, and through model routing, automatic switching, and SLA, provide stable service to 60 countries…

X AI KOLs Following ↗ · 2026-08-21 Cached

Harvey shares strategies for achieving world-class AI research on a limited budget through domain experts guiding synthetic data, establishing evaluation sets, collaborating with labs, and employing model routing.

0 favorites 0 likes
#legal-ai

ContractScrub: A benchmark for final review of legal contracts

arXiv cs.AI ↗ · 2026-08-21 Cached

Introduces ContractScrub, a benchmark for evaluating LLMs on legal contract scrubbing tasks, revealing that current frontier models perform poorly on this domain-specific challenge.

0 favorites 0 likes
#legal-ai

@lqiao: We are excited to drive the research work of Tenet with Harvey, delivering frontier quality across 24 areas of corporat…

X AI KOLs Following ↗ · 2026-08-21 Cached

Tenet AI system demonstrates superior performance in corporate law tasks, outperforming Opus5 and Fable across multiple dimensions.

0 favorites 0 likes
#legal-ai

Harvey post-trains Kimi K3 for long-horizon legal work (10 minute read)

TLDR AI ↗ · 2026-08-21 Cached

Harvey has post-trained the Kimi K3 model to create Harvey Tenet for long-horizon legal work, achieving improved performance on legal benchmarks and cost-efficiency.

0 favorites 0 likes
#legal-ai

@LangChain: LangSmith for Startups Spotlight: Vector Legal Vector Legal is the premier AI-native law firm for startups, venture cap…

X AI KOLs Following ↗ · 2026-08-20 Cached

Vector Legal, an AI-native law firm for startups and venture capital, uses LangSmith for agent deployment and fleet management in their legal workflows, having recently raised a seed round.

0 favorites 0 likes
#legal-ai

Thomson Reuters launches CoCounsel Legal with a Westlaw-backed agentic workflow

Reddit r/artificial ↗ · 2026-08-20

Thomson Reuters launched the next generation of CoCounsel Legal, an AI-powered product integrating legal research, drafting, and workflows with Westlaw and Practical Law, including the Westlaw Brief Builder for automated drafting while maintaining lawyer control.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback