negative-result

Tag

Cards List
#negative-result

Prompt-Space Meta-Learning Does Not Transfer Across Users: A Frozen-LLM Negative Result

arXiv cs.LG · 2026-09-03 Cached

The paper demonstrates that prompt-space meta-learning for personalizing frozen large language models does not transfer across users, as the meta-validation objective is statistically invariant to user-support correspondence, leading to no significant improvement over seed prompts or controls.

0 favorites 0 likes
#negative-result

On Scope Classification and Current Knowledge-Editing Benchmarks: A Negative Result, with INLAY as a Gradient-Free Case Study

arXiv cs.CL · 2026-08-28 Cached

This paper presents a negative result showing that current knowledge-editing benchmarks cannot effectively evaluate scope classifiers, using INLAY, a gradient-free editor, to demonstrate that no per-query routing method can improve performance due to structural limitations in the benchmarks.

0 favorites 0 likes
#negative-result

Cacheable by Design? Training Mixture-of-Experts Routers for Locality Against the Edge Memory-Bandwidth Wall: A Pre-Registered Negative Result with a Systems Measurement Study

arXiv cs.AI · 2026-08-20 Cached

This paper presents a pre-registered negative result on training mixture-of-experts routers for cache locality against memory-bandwidth walls, showing that miss reduction trades off with language modeling quality despite training mechanisms.

0 favorites 0 likes
#negative-result

VectraYX-Vision-1B: A Sub-2B Spanish/LATAM Cybersecurity Vision-Language Model with Structured Visual Reasoning and Native Tool Use

arXiv cs.CL · 2026-08-11 Cached

This paper introduces VectraYX-Vision-1B, a sub-2B Spanish/LATAM cybersecurity vision-language model, but reports a negative visual-grounding result, raising architectural questions about NoPE layers and releasing code, benchmarks, and checkpoints.

0 favorites 0 likes
#negative-result

Accepted Prefixes Are Not All You Need: A Negative Result on PEFT-Based Block-Diffusion Drafting

arXiv cs.AI · 2026-07-15 Cached

This paper studies PEFT-BD, a speculative decoding method using a LoRA-like adapter as a block-diffusion drafter, and finds that despite nontrivial accepted prefixes, it does not yield speedup because the drafter still requires a full-backbone pass, making it not compute-efficient.

0 favorites 0 likes
#negative-result

Phase-Localized Curation Does Not Help: A Negative Result on Per-Phase Metric Selection for Demonstration Filtering

arXiv cs.LG · 2026-06-16 Cached

This paper investigates whether per-phase metric selection improves demonstration curation for behavior cloning policies. The authors find that phase-gated curation never outperforms global or uniform metric application, and the dilution of defect signals across phases explains the failure.

0 favorites 0 likes
#negative-result

A Negative Result on Cross-Model Activation Transfer in a Pythia Multi-Hop Setting

arXiv cs.AI · 2026-06-03 Cached

This paper investigates whether direct activation transfer between language models can improve reasoning, using a linear translation layer from Pythia-160M to Pythia-410M. Despite achieving high representational alignment, the transferred activations do not improve multi-hop question answering, yielding a negative result.

0 favorites 0 likes
← Back to home

Submit Feedback