natural-language-processing

Tag

Cards List
#natural-language-processing

To What Extent Do Large Language Models Understand Bangla Idioms?

arXiv cs.CL · 5d ago Cached

This paper introduces a benchmark dataset for Bangla idioms and evaluates recent large language models on idiom-related tasks, revealing substantial variability in performance across models.

0 favorites 0 likes
#natural-language-processing

FrameBench:A Language Understanding Benchmark Based on Frame Semantics

arXiv cs.CL · 5d ago Cached

FrameBench is a new benchmark for evaluating whether large language models can distinguish context-dependent frame-semantic interpretations of verbs, constructed for English and Japanese using FrameNet resources and released with code.

0 favorites 0 likes
#natural-language-processing

MineTRACE: An Evidence-Grounded Interactive Reasoning System for Mineral Prospectivity

arXiv cs.AI · 6d ago Cached

MineTRACE is a web-based evidence-grounded interactive reasoning system that integrates geochemical, geophysical, and geological data to provide transparent mineral prospectivity scores and natural language interaction, supporting efficient and verifiable mineral exploration.

0 favorites 0 likes
#natural-language-processing

When Persona Attributes Improve Population Alignment in Large Language Models

arXiv cs.CL · 6d ago Cached

This paper explores how persona prompting with different attribute selection methods affects the alignment of large language models with human responses in social surveys. It finds that effectiveness depends on human response variation and the choice of attributes.

0 favorites 0 likes
#natural-language-processing

MultiGhostBench: A Multilingual Benchmark for Long-Form LLM-Generated Text Attribution under Distribution Shifts

arXiv cs.CL · 6d ago Cached

MultiGhostBench is a multilingual benchmark for long-form LLM-generated text attribution under distribution shifts, featuring 928 books in six languages and highlighting performance degrades and no single method consistently best across settings.

0 favorites 0 likes
#natural-language-processing

Candidate Generation and Definition-Guided Verification for Sentence-Level Depression Symptom Recognition

arXiv cs.CL · 6d ago Cached

This study proposes a two-stage framework for sentence-level depression symptom recognition using candidate generation and definition-guided verification, achieving best accuracy and F1 scores among evaluated methods.

0 favorites 0 likes
#natural-language-processing

Pangram Has Emerged as the Gold Standard of AI Detection. Should You Trust It?

Wired · 2026-09-02 Cached

The article explores Pangram, an AI startup that detects AI-generated text, its role in publishing scandals, and questions about the trustworthiness of its detection model.

0 favorites 0 likes
#natural-language-processing

@xuanyuanzhifeng: https://x.com/xuanyuanzhifeng/status/2095044737531306357

X AI KOLs Timeline · 2026-09-02 Cached

This article provides a detailed explanation of the Transformer architecture, covering attention mechanisms, QKV, and residual connections, while tracing its historical development from N-gram to LSTM.

0 favorites 0 likes
#natural-language-processing

The Curse of Multilinguality in Lexical Normalization

arXiv cs.CL · 2026-09-02 Cached

This paper investigates the curse of multilinguality in lexical normalization, finding that training a single model on multiple languages leads to decreased per-language accuracy, with optimal performance when languages are trained in small groups.

0 favorites 0 likes
#natural-language-processing

CUDA-Harness: Harnessing Agentic CUDA Kernel Generation and Optimization from Natural Language

arXiv cs.CL · 2026-09-02 Cached

CUDA-Harness is a framework that uses agentic techniques to generate and optimize CUDA kernels from natural language descriptions, addressing challenges in Text2CUDA by connecting high-level semantics with low-level implementation and verification.

0 favorites 0 likes
#natural-language-processing

AI Historian: Helping historians organize and verify person-centred temporal clues from dispersed historical narratives

arXiv cs.CL · 2026-09-01 Cached

AI Historian is an AI agent system that helps historians organize and verify person-centred temporal clues from dispersed historical narratives, reducing the cost of historical research while achieving high accuracy in temporal localization.

0 favorites 0 likes
#natural-language-processing

I built a natural language IVR that routes callers without phone trees

Reddit r/ArtificialInteligence · 2026-08-28

The article demonstrates a Python/Flask example using Telnyx Call Control and AI inference to create a natural language IVR system, allowing callers to verbally state their needs instead of navigating fixed menus.

0 favorites 0 likes
#natural-language-processing

Data Science Approaches to Evaluating Honours Candidates

arXiv cs.CL · 2026-08-28 Cached

This paper presents the first application of data science to evaluate the UK Honours system using natural language processing, introducing a novel sentiment analysis algorithm called Minos to assess public opinion on honours recipients.

0 favorites 0 likes
#natural-language-processing

Natural-Language Policies to Executable Decisions: An Interpretable Large Language Model Framework

arXiv cs.CL · 2026-08-28 Cached

This paper presents a production-grade framework that uses large language models to convert natural-language pricing policies into executable decisions for tourism pricing, achieving significant efficiency gains and auditability in real-world deployment.

0 favorites 0 likes
#natural-language-processing

Behind the [MASK]: Disentangling Representation and Faithfulness in DAPF-Based Dementia Detection

arXiv cs.CL · 2026-08-27 Cached

This paper investigates the interpretability of DAPF-based models for dementia detection, revealing that while DAPF achieves strong performance, its token-level explanations lack faithfulness.

0 favorites 0 likes
#natural-language-processing

From Association to Causation: Improving Retrieval Precision of Retrieval-Augmented Generation via Causal Relations and an Attention Mechanism

arXiv cs.AI · 2026-08-25 Cached

This paper introduces a causal graph-based attention mechanism to enhance retrieval precision in Retrieval-Augmented Generation (RAG) systems, showing improvements in keyword-stuffing regimes of proprietary knowledge bases.

0 favorites 0 likes
#natural-language-processing

Do Large Language Models Perform Well on Comprehending Poetic Logic in Modern Chinese Poetry?

arXiv cs.CL · 2026-08-25 Cached

The paper introduces Peony, the first benchmark for evaluating large language models on comprehending poetic logic in modern Chinese poetry, and evaluates six mainstream LLMs, revealing their limitations in this specialized task.

0 favorites 0 likes
#natural-language-processing

Ontology-Driven Structural Regularization for Document-Level Relation Extraction

arXiv cs.CL · 2026-08-24 Cached

The paper introduces an ontology-driven framework to quantify and enforce structural consistency in document-level relation extraction datasets, reducing logical contradictions and improving model generalization when using distant supervision data.

0 favorites 0 likes
#natural-language-processing

ImmigrationReason: A Structured Dataset of U.S. Immigration Appeals for Legal Reasoning Research

arXiv cs.CL · 2026-08-24 Cached

This paper introduces ImmigrationReason, a large-scale structured dataset of U.S. immigration appeals for legal reasoning research, addressing the gap in administrative adjudication data for NLP studies.

0 favorites 0 likes
#natural-language-processing

An ambiguity taxonomy for evaluating large language model performance on clinical registry abstraction: a multi-site prospective study

arXiv cs.CL · 2026-08-24 Cached

This study develops an ambiguity taxonomy to evaluate large language model performance on clinical registry abstraction from unprocessed EMR data, finding that LLM accuracy is significantly lower than human abstractors and declines as task ambiguity increases.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback