perturbation

Tag

Cards List
#perturbation

Plato’s Cave has a problem: telling someone they’re seeing shadows just puts another shadow on the wall

Reddit r/artificial · 2026-08-24

The article explores the philosophical problem of Plato's Cave in the context of LLMs, proposing an experiment to compare how different conversational regimes—reconstructive versus perturbation-sensitive—might yield measurable differences in interaction behavior.

0 favorites 0 likes
#perturbation

Perturbation-based Regional Interpretability through Subtraction Mapping (PRISM): naming-error dissociations in language models and post-stroke aphasia

arXiv cs.LG · 2026-08-14 Cached

This paper introduces PRISM, a perturbation-based method for spatially resolved interpretability of large language models, adapting neuroimaging subtraction analysis to transformers and applying it in parallel to post-stroke aphasia patients to recover shared phonemic-favoring dissociations.

0 favorites 0 likes
#perturbation

Decomposition of Evidence, Contradiction, and Fragility in Perturbation Responses

arXiv cs.AI · 2026-08-14 Cached

Introduces DECAF, a method that decomposes perturbation responses into evidence, contradiction, and fragility components, improving interpretability over raw response magnitude and achieving strong results across vision benchmarks.

0 favorites 0 likes
#perturbation

AndroidReality: How Far Are Mobile Agents from the Real World?

arXiv cs.AI · 2026-08-11 Cached

Introduces AndroidReality, a perturbation-based framework for evaluating and improving the robustness of mobile agents, with a taxonomy of real-world interface perturbations and a training-free Test-Time Introspective Recovery (TTIR) mechanism.

0 favorites 0 likes
#perturbation

Different Perturbations, Different Mechanisms: Understanding Continued Pre-training for Zero-Shot Dialect Robustness

arXiv cs.CL · 2026-08-07 Cached

This paper systematically studies perturbation-based continued pre-training (CPT) for improving zero-shot dialect robustness in multilingual LLMs, comparing six training conditions across German, Italian, and Arabic. It finds that character-noised CPT is the most effective general strategy and reveals that different perturbation methods induce distinct robustness mechanisms.

0 favorites 0 likes
#perturbation

Perturbation Sensitivity at Convergence: A Simple Signal for Identifying Spuriously Correlated Samples

arXiv cs.LG · 2026-08-07 Cached

This paper proposes a method to identify spuriously correlated samples after model convergence by measuring prediction fragility under input perturbation, requiring no group labels or early-stopping epochs. Rebalancing training with detected samples improves worst-group accuracy on Waterbirds from 57.3% to 80.8%.

0 favorites 0 likes
#perturbation

Inpainting Insights: Elevating Visual XAI with Photorealistic Perturbations

arXiv cs.LG · 2026-07-20 Cached

This paper proposes using generative inpainting to create photorealistic perturbations for LIME, improving the quality of explanations by avoiding out-of-distribution artifacts common in traditional occlusion methods.

0 favorites 0 likes
#perturbation

@QuixiAI: https://arxiv.org/abs/2509.21401 this is the coolest thing I've seen in at least an hour @TroyDoesAI @elder_plinius @ma…

X AI KOLs Following · 2026-06-28 Cached

Proposes JaiLIP, a method that jailbreaks vision-language models by generating imperceptible adversarial images using loss-guided perturbation, achieving high toxicity and outperforming existing methods.

0 favorites 0 likes
#perturbation

@arcinstitute: Because PerturbSpace uses standard single-cell sequencing, it's compatible with any single-cell readout. In one day, th…

X AI KOLs Timeline · 2026-05-26 Cached

Arc Institute's PerturbSpace enables high-throughput single-cell profiling of transcriptome, location, CRISPR guides, clonal relationships, and surface proteins from many samples in one day, using standard single-cell sequencing.

0 favorites 0 likes
#perturbation

How Do Document Parsers Break? Auditing Structural Vulnerability in Document Intelligence

arXiv cs.CL · 2026-05-20 Cached

This paper identifies Footprint Bias in document layout analysis robustness evaluation and proposes a structure-aware auditing framework that decouples probe construction and pathway attribution, showing that small structurally targeted probes cause comparable downstream degradation to larger perturbations.

0 favorites 0 likes
#perturbation

Geometric coherence of single-cell CRISPR perturbations reveals regulatory architecture and predicts cellular stress

Hugging Face Daily Papers · 2026-04-17 Cached

This paper introduces Shesha, a geometric stability metric that quantifies directional coherence of single-cell CRISPR perturbation responses using mean cosine similarity, revealing regulatory architecture and predicting cellular stress across 2,200+ perturbations in five CRISPR datasets.

0 favorites 0 likes
← Back to home

Submit Feedback