icml

Tag

Cards List
#icml

CaLR: Causal Latent Revision for Robust Diffusion Reasoning

arXiv cs.AI ↗ · 2026-09-21 Cached

The paper introduces CaLR, a framework that reformulates reasoning as constrained latent optimization using causal topology to enhance diffusion language models, achieving state-of-the-art performance on complex benchmarks.

0 favorites 0 likes
#icml

What We Learned by Reproducing 2,200 papers from ICML

Hugging Face Blog ↗ · 2026-08-13 Cached

Hugging Face shares findings from a community hackathon where 1,200 participants used coding agents to reproduce 2,226 ICML papers, highlighting issues in review rigor and the potential for AI agents to scale verification.

0 favorites 0 likes
#icml

Hot Take: LLM can 'jump'

Reddit r/singularity ↗ · 2026-08-11 Cached

A blog post arguing that LLMs could potentially achieve scientific breakthroughs like General Relativity through deductive, non-abductive routes, countering Tom Zahavy's 'LLMs can't jump' position paper.

0 favorites 0 likes
#icml

@huggingface: How AI agents reproduced ICML 2026 papers

X AI KOLs Timeline ↗ · 2026-08-07 Cached

Hugging Face hosted a live broadcast discussing how AI agents reproduced ICML 2026 papers.

0 favorites 0 likes
#icml

Subliminal Learning is Non-Semantic Distillation

arXiv cs.AI ↗ · 2026-08-07 Cached

This paper investigates subliminal learning in language models, showing that biases can transfer from teacher to student via seemingly random synthetic data. The authors find that adding Gaussian noise to weights increases transfer, and that students inherit not just the semantic bias but also the type of intervention used, with implications for training safety and data auditing.

0 favorites 0 likes
#icml

Benchmarks Are Not Monolithic: Sample-Level Auditing and Orchestration for LLM Evaluation

arXiv cs.CL ↗ · 2026-08-03 Cached

This paper proposes a dataset-centric meta-evaluation framework that audits LLM benchmarks at the sample level across five latent dimensions, exposing internal heterogeneity and enabling criterion-driven composition of benchmark subsets for targeted model evaluation.

0 favorites 0 likes
#icml

A fundamental flaw leaves LLMs strikingly vulnerable to attack

MIT Technology Review ↗ · 2026-07-30 Cached

Researchers present a paper at ICML arguing that a fundamental flaw in how LLMs identify instructions makes them impossible to fully secure against attacks, demonstrating successful exploits against models from OpenAI, Anthropic, Alibaba, and DeepSeek.

0 favorites 0 likes
#icml

VeriSimpl: Robust Optimization Modeling from Natural Language using Simplification-based Verification

arXiv cs.AI ↗ · 2026-07-24 Cached

VeriSimpl introduces a solver-LLM framework that uses simplification-based verification to ensure correct translation of natural language optimization problems into solver formulations, achieving improved accuracy over existing methods.

0 favorites 0 likes
#icml

DecodeShare: Tracing the Shared Subspace of LLM Decode-Time Decisions

arXiv cs.AI ↗ · 2026-07-24 Cached

DecodeShare proposes a method to identify a low-dimensional subspace consistently shared across tasks in LLM decode-time hidden states and shows that disturbing this subspace degrades decision performance more than random or prefill-derived subspaces, with implications for activation steering.

0 favorites 0 likes
#icml

Training Continuous Chain of Thought Models: A Tale of Two Regimes

arXiv cs.AI ↗ · 2026-07-21 Cached

This paper introduces C-MTP, a direct supervision method for training continuous chain-of-thought models that compresses reasoning traces into latent representations. The method performs competitively on simple tasks but reveals that both direct and indirect supervision methods struggle with complex long reasoning traces, showing about 65% performance drop.

0 favorites 0 likes
#icml

RELIC: Revealed Principles for Learning Interpretable Composable Skills in Multi-Agent Planning

arXiv cs.AI ↗ · 2026-07-21 Cached

Introduces RELIC, a framework for learning interpretable and composable skills in multi-agent planning via revealed principles, enabling privacy-preserving coordination and cross-agent skill transfer without sharing code.

0 favorites 0 likes
#icml

AI is more likely than humans to form biases when hiring

MIT Technology Review ↗ · 2026-07-20 Cached

New research shows that LLMs can develop their own biases from experience and stereotype job applicants more than humans, raising concerns about AI in hiring.

0 favorites 0 likes
#icml

@AnimaAnandkumar: Excited to share our @icmlconf paper: M+Adam: Low-Precision Training via Additive–Multiplicative Optimization We introd…

X AI KOLs Timeline ↗ · 2026-07-14 Cached

Introduces M+Adam, a novel optimizer combining additive and multiplicative updates to enable effective low-precision training of LLMs using BF16, FP8, or FP4 master weights, avoiding failure modes of standard optimizers.

0 favorites 0 likes
#icml

What will be left for us to work on?

Hacker News Top ↗ · 2026-07-14 Cached

Arvind Narayanan's ICML 2026 keynote argues that the 'AI as Normal Technology' framework is useful for understanding AI's impacts, rejects the notion of sudden job loss from AI, and envisions a future of human-AI co-superintelligence requiring significant adaptation.

0 favorites 0 likes
#icml

Prompt-engineering paper accepted to ICML [R]

Reddit r/MachineLearning ↗ · 2026-07-13

A paper titled 'Verbalized Sampling: How to Mitigate Mode Collapse and Unlock LLM Diversity' has been accepted to ICML. It proposes a simple prompt-engineering trick for more diverse sampling, sparking debate over whether such work belongs at a top-tier ML conference.

0 favorites 0 likes
#icml

@AnimaAnandkumar: Congratulations!

X AI KOLs Following ↗ · 2026-07-10 Cached

Pengrui Han's paper received the Best Paper Award at the ICML Combining Theory and Benchmarks Workshop, with congratulations from Anima Anandkumar.

0 favorites 0 likes
#icml

Contrastive Order Learning: A General Framework for Ordinal Regression

arXiv cs.LG ↗ · 2026-07-10 Cached

ConOrd proposes a contrastive learning framework for ordinal regression that integrates contrastive learning and order learning, achieving state-of-the-art performance on facial age estimation, image quality assessment, and video quality assessment.

0 favorites 0 likes
#icml

Stochastic Order Learning: An Approach to Rank Estimation Using Noisy Data

arXiv cs.LG ↗ · 2026-07-10 Cached

This paper reformulates rank estimation with noisy ordinal labels as a stochastic ordering problem and proposes a learning framework (SOL) that captures ordinal label uncertainty through discriminative and stochastic order losses, achieving reliable rank estimation under various noise types.

0 favorites 0 likes
#icml

Towards the Explainability of Temporal Graph Networks via Memory Backtracking and Topological Attribution

arXiv cs.LG ↗ · 2026-07-10 Cached

This paper introduces MemExplainer, a method to explain predictions of Temporal Graph Networks (TGNs) by attributing contributions through topology attribution trees and memory backtracking trees, using Layer-wise Relevance Propagation (LRP) for faithful explanations.

0 favorites 0 likes
#icml

Flexible Video Diffusion (3 minute read)

TLDR AI ↗ · 2026-07-10 Cached

Flex-Forcing introduces a unified framework for video diffusion that supports both autoregressive and bidirectional generation modes, offering flexible control for video generation tasks.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback