natural-language-processing

Tag

Cards List
#natural-language-processing

Do we need to answer that question? Salience and Answerability of Potential Questions in Naturalistic Dialogue

arXiv cs.CL ↗ · yesterday Cached

This paper empirically investigates the relationship between salience and answerability of questions in naturalistic dialogue, finding a robust but low positive correlation that is weaker than in monologic text, indicating that conversational structure is less predictable.

0 favorites 0 likes
#natural-language-processing

Coupled Usage-Sense Processes: Temporal and Attributable Lexical Semantic Change

arXiv cs.CL ↗ · yesterday Cached

This paper introduces Coupled Usage–Sense Processes (CUSP), a hierarchical model for analyzing lexical semantic change by quantifying timing, mechanisms, and attributions in word usage over time.

0 favorites 0 likes
#natural-language-processing

Do LLMs Understand Context? A Knowledge Graph-Based Evaluation Framework

arXiv cs.AI ↗ · yesterday Cached

This paper proposes a knowledge graph-based evaluation framework for assessing the contextual understanding of large language models in question answering, introducing a new similarity measure called S3KG that achieves superior performance over existing baselines.

0 favorites 0 likes
#natural-language-processing

Automatic Rank Allocation for Low-Rank Adaptation in Large Language Models via lp Regularization

arXiv cs.LG ↗ · 4d ago Cached

This paper introduces ℓp-LoRA, a principled method for automatic rank allocation in low-rank adaptation using ℓp regularization, demonstrating competitive performance on NLP tasks.

0 favorites 0 likes
#natural-language-processing

When Explanations Cannot Be Read: Measuring and Correcting SHAP and LIME Rendering for Right-to-Left Languages

arXiv cs.LG ↗ · 4d ago Cached

The study introduces SHAP-RTL, a rendering layer that corrects the visualization of SHAP and LIME explanations for right-to-left languages, addressing issues like token sequence and script shaping while preserving original attribution values.

0 favorites 0 likes
#natural-language-processing

Can Classical Semantic-Extractive Summarization Be Evaluated in Hindi? A Replication Study

arXiv cs.CL ↗ · 4d ago Cached

This paper replicates a distributional-semantics extractive summarization method for Hindi and evaluates it on standard corpora, finding that sentence position is the only contributing feature and current Hindi benchmarks fail to incentivize advanced content selection.

0 favorites 0 likes
#natural-language-processing

Contrastive Language Models

Hacker News Top ↗ · 5d ago

The article likely presents research on contrastive language models, exploring the use of contrastive learning techniques in language model development.

0 favorites 0 likes
#natural-language-processing

Improving LLM-based Autonomous Web Agents with Filtering

arXiv cs.CL ↗ · 5d ago Cached

The paper proposes retrieval strategies to filter irrelevant HTML content for LLM-based autonomous web agents, improving performance on benchmarks like WebArena.

0 favorites 0 likes
#natural-language-processing

Not What You Meant: Can LLMs Follow a Specified Negation Semantics?

arXiv cs.AI ↗ · 5d ago Cached

This paper introduces NAF-Bench to study how large language models adhere to specified negation semantics, finding that frontier models like o4-mini perform well while open-source models lag, and suggesting improvements via solver delegation or fine-tuning.

0 favorites 0 likes
#natural-language-processing

Phonemizing User-Generated Text: A Benchmark, Taxonomy, and Compositional Approach

arXiv cs.CL ↗ · 5d ago Cached

This paper introduces UGTPhon, a benchmark for grapheme-to-phoneme conversion in user-generated text, and presents a compositional approach that improves performance by leveraging canonical forms.

0 favorites 0 likes
#natural-language-processing

Building Socio-Affective Artificial Intelligence for Interactive Multi-Agent Simulations

arXiv cs.AI ↗ · 5d ago Cached

This paper presents design principles and the 'AGIMUD' software for enabling socio-affective interactions between humans and multiple AI agents in simulated dynamic worlds, leveraging generative AI and distributed processing.

0 favorites 0 likes
#natural-language-processing

ARAFA: An LLM-Generated Arabic Fact-Checking Dataset

arXiv cs.CL ↗ · 6d ago Cached

This paper introduces Arafa, a large-scale Arabic fact-checking dataset generated using LLMs, aimed at addressing the scarcity of resources for automatic fact-checking in Arabic.

0 favorites 0 likes
#natural-language-processing

Compressing Long Context into Answer-Aligned Memory Embeddings for LLM Inference

arXiv cs.CL ↗ · 6d ago Cached

The paper proposes a Context-to-Answer-Aligned Memory Compression (CMC) framework that compresses long input contexts into compact memory embeddings to reduce LLM inference costs without modifying decoder weights, achieving significant performance and efficiency gains.

0 favorites 0 likes
#natural-language-processing

Peerify: Benchmarking Peer-Review Claim Verification

arXiv cs.CL ↗ · 6d ago Cached

Peerify is a pipeline for automatically verifying peer-review claims against manuscript evidence, using a benchmark of 800 claims from NeurIPS 2024 and ICLR 2024, demonstrating that retrieval-centered verification outperforms entailment baselines.

0 favorites 0 likes
#natural-language-processing

@VraserX: What a beautiful day for AI. GPT-6 Sol and Luna are finally here, bringing Astra-level advances into models that are fa…

X AI KOLs Timeline ↗ · 6d ago Cached

OpenAI has released GPT-6 Sol and GPT-6 Luna, which are faster and more affordable models based on GPT-6 Astra's advances, aimed at scalable use.

0 favorites 0 likes
#natural-language-processing

Beyond the Stitching Assumption: A Unified Framework for Multimodal Synthetic Data Evaluation via Semantic Quantization

arXiv cs.CL ↗ · 2026-09-22 Cached

This paper presents a unified framework for evaluating multimodal synthetic data using semantic quantization and cross-modal metrics, emphasizing the need for explicit evaluation with permutation baselines and coverage reporting.

0 favorites 0 likes
#natural-language-processing

Hemmingway-1: An AI that writes like a person (Qwen3.8-27B finetune)

Reddit r/LocalLLaMA ↗ · 2026-09-21

Hemmingway-1 is a finetuned version of the Qwen3.8-27B AI model, designed to generate text in a more human-like writing style.

0 favorites 0 likes
#natural-language-processing

What's the best way to get an agent to turn meeting notes into action items reliably?

Reddit r/AI_Agents ↗ · 2026-09-21

The user describes challenges and partial solutions for making an AI agent reliably convert raw meeting notes into structured action items with owners and due dates, highlighting issues like hallucination and missed context.

0 favorites 0 likes
#natural-language-processing

CaLR: Causal Latent Revision for Robust Diffusion Reasoning

arXiv cs.AI ↗ · 2026-09-21 Cached

The paper introduces CaLR, a framework that reformulates reasoning as constrained latent optimization using causal topology to enhance diffusion language models, achieving state-of-the-art performance on complex benchmarks.

0 favorites 0 likes
#natural-language-processing

When Does Reasoning Help in Machine Translation? A Hierarchical Analysis of LRM Reasoning Traces

arXiv cs.CL ↗ · 2026-09-21 Cached

This paper analyzes reasoning traces in large reasoning models for machine translation, finding that reasoning benefits are conditional and introducing Hierarchical Meta-Summarization to understand trace patterns.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback