attribution

Tag

Cards List
#attribution

OpenAI’s feud with mathematicians is only escalating

TechCrunch AI · yesterday Cached

A group of 25 Fields Medal-winning mathematicians has signed an open letter accusing AI labs, particularly OpenAI, of threatening mathematical research through unverified proofs and poor attribution practices, escalating tensions and raising concerns about the integrity of open science.

0 favorites 0 likes
#attribution

Through the Looking Glass: Directly Reading and Writing Transformers

arXiv cs.CL · 2d ago Cached

The paper reveals that only a small fraction of transformer components are necessary for token predictions and introduces direct methods to read and write to these components with minimal changes.

0 favorites 0 likes
#attribution

On the Role of Citations in Preference Data

arXiv cs.CL · 2026-08-25 Cached

This paper investigates the role of citations in human and LLM preferences for scientific question answering, finding that humans prefer diverse citations but fewer overall, while LLMs exhibit stronger citation-related preferences despite lacking source access.

0 favorites 0 likes
#attribution

@RayFernando1337: Can my fam at X & Space xAI help? This guy is stealing original content instead of giving attribution to the creators. …

X AI KOLs Following · 2026-08-24 Cached

A user on X seeks assistance from X and Space xAI to address content theft by another user who fails to provide attribution.

0 favorites 0 likes
#attribution

An agent ran a full git workflow autonomously this week: init, commit, push, in seconds. The interesting question is whose name is on the commit.

Reddit r/AI_Agents · 2026-08-16

The article discusses a demo of an AI agent autonomously executing git workflows, highlighting critical issues with attribution, identity, and compliance, and proposes cryptographic agent identities while questioning accountability models.

0 favorites 0 likes
#attribution

LongRCA Bench: Diagnosing Responsible Roles and Root Causes in Long-Horizon Agent Failures

Hugging Face Daily Papers · 2026-08-15 Cached

LongRCA Bench introduces a benchmark for diagnosing failures in long-horizon agent trajectories, and the RCTA method improves responsible role and root-cause step attribution.

0 favorites 0 likes
#attribution

Beyond Attention: Signed Integrated Gradients Attribution in a BiomeGPT-Style Microbiome Transformer

arXiv cs.LG · 2026-08-10 Cached

This paper proposes using signed, fusion-aware Integrated Gradients for attributing predictions in feature-tokenized transformers like BiomeGPT, overcoming limitations of CLS attention weights and revealing disease-supporting versus protective microbial signals.

0 favorites 0 likes
#attribution

Interpreting Black-Box Large Language Models with Sentence-Level Energy Landscapes

arXiv cs.AI · 2026-08-05 Cached

This paper proposes a model-agnostic, post-hoc interpretation framework for black-box LLMs using sentence-level energy landscapes. A surrogate Energy-Based Model simulates the target LLM, and a lightweight interpreter network identifies which prompt sentences most influence a given output, without requiring further API calls.

0 favorites 0 likes
#attribution

Token-Level Diagnosis of Sycophancy in LLMs with Attribution-Guided Steering

arXiv cs.CL · 2026-08-03 Cached

This paper introduces an Integrated Gradients-based token attribution method to diagnose sycophancy in LLMs at the token level, and proposes attribution-guided contrastive activation steering to reduce sycophantic behavior during inference without retraining.

0 favorites 0 likes
#attribution

@jerryjliu0: One of the hardest parts of document parsing is getting granular attribution and bounding boxes. This lets you ground e…

X AI KOLs Following · 2026-07-22 Cached

LlamaParse now supports multi-layered bounding boxes (region, line, word) for granular document text attribution, improving auditability in invoices, research reports, and more.

0 favorites 0 likes
#attribution

LAPO: Leave-One-Turn Attribution for Self-Generated Process Rewards in Multi-Turn Search Reasoning

arXiv cs.AI · 2026-07-16 Cached

LAPO proposes a leave-one-turn attribution method for self-generated process rewards in multi-turn search reasoning, enabling fine-grained credit assignment without external reward models. It achieves state-of-the-art results across seven datasets.

0 favorites 0 likes
#attribution

Sparse Inter-Layer Dependencies of Transformer FFN Neurons

arXiv cs.LG · 2026-07-15 Cached

This paper introduces a training-free attribution method to identify sparse inter-layer dependencies in Transformer FFN neurons, showing that small subsets of preceding activations suffice to preserve neuron activations with high fidelity.

0 favorites 0 likes
#attribution

Nonlinear Axiomatic Attribution for Cooperative Games

arXiv cs.LG · 2026-07-14 Cached

This paper introduces a class of nonlinear axiomatic attribution methods for cooperative games to overcome the limitations of the linear Shapley value, which has an excessively large null space. Experimental results demonstrate the potential effectiveness of these methods in terms of inclusion AUC metric compared to Shapley value variants.

0 favorites 0 likes
#attribution

One aspect that might be missing in the AI agent monetization process is attribution.

Reddit r/AI_Agents · 2026-07-13

The article discusses the challenge of attribution in AI agent monetization, where determining credit for conversions becomes complex when agents recommend products and influence users before clicks.

0 favorites 0 likes
#attribution

MultAttnAttrib: Training-Free Multimodal Attribution in Long Document Question Answering

arXiv cs.CL · 2026-07-03 Cached

Introduces MultAttnAttrib, a training-free method for multimodal attribution in long document QA, along with the MultAttrEval benchmark. It outperforms prompting-based methods and matches frontier models like GPT-5.4.

0 favorites 0 likes
#attribution

Turn-Averaged SAEs for Feature Discovery and Long-Context Attribution

arXiv cs.CL · 2026-06-30 Cached

This paper introduces turn-averaged sparse autoencoders (SAEs) that operate on average activations across conversational turns, enabling efficient feature discovery and attribution graphs for long contexts. It also proposes a nested architecture for joint training with per-token features.

0 favorites 0 likes
#attribution

GBC: Gradient-Based Connections for Optimizing Multi-Agent Systems

Hugging Face Daily Papers · 2026-06-26 Cached

Proposes Gradient-Based Connections (GBC), a method that models multi-agent LLM systems as computational graphs and uses gradient signals to attribute errors to specific agents, enabling better system-level optimization.

0 favorites 0 likes
#attribution

MGI: Member vs Generated Inference

arXiv cs.LG · 2026-06-24 Cached

Introduces the Member vs Generated Inference (MGI) task to distinguish training members from generated outputs in generative models, and proposes Data Circuit Breaker (DCB), a three-stage method combining autoencoder and latent generator signals, which outperforms existing methods across autoregressive and diffusion models.

0 favorites 0 likes
#attribution

Faithful by Construction: Claim-Anchored Attribution for Multi-Document Summarization

arXiv cs.CL · 2026-06-24 Cached

This paper introduces CAMS, a modular multi-document summarization framework that extracts atomic claims with token-level provenance, clusters equivalent claims, and rewrites them into summaries with fine-grained, multi-source traceability, significantly improving faithfulness and citation precision.

0 favorites 0 likes
#attribution

Who Drifted: the System or the Judge? Anytime-Valid Attribution in LLM Evaluation Pipelines

arXiv cs.AI · 2026-06-16 Cached

Proposes an anytime-valid attribution method that uses a human-labeled anchor set and a betting e-process to distinguish whether score drift in LLM evaluation pipelines comes from the system or the judge, resolving the ambiguity caused by silent judge changes.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback