entropy

Tag

Cards List
#entropy

What Is Entropy, Really?

Wired · 2d ago Cached

An insightful explainer on the true meaning of entropy, contrasting the common 'disorder' metaphor with a probabilistic interpretation using dice rolling analogies.

0 favorites 0 likes
#entropy

@_yusufknl: In 1948, Claude Shannon invented the math behind every LLM you use today. He tested it by making his wife guess the nex…

X AI KOLs Timeline · 5d ago Cached

A detailed walkthrough explains how Claude Shannon's 1948 information theory underlies LLMs and shows that the 'next-token prediction' story is misleading, linking compression and prediction mathematically.

0 favorites 0 likes
#entropy

How Much Human Label Variation Does Formal Semantic Structure Explain?: Group-Level Effects and Item-Level Ceilings in NLI

arXiv cs.CL · 6d ago Cached

This paper measures how much formal semantic structure explains human label variation in natural language inference (NLI) using ChaosNLI data, finding group-level effects on entropy but item-level ceilings and null composition effects.

0 favorites 0 likes
#entropy

Comparing Semantic Navigation in Humans and Large Language Models using Natural Language Processing

arXiv cs.CL · 2026-07-15 Cached

This paper compares semantic search dynamics between humans and LLMs using verbal fluency data, finding that humans exhibit more variable and exploratory search patterns that current models fail to reproduce.

0 favorites 0 likes
#entropy

Entropy in Semantic Memory Navigation in Blind and Sighted Individuals: The Effect of Visual Experience

arXiv cs.CL · 2026-07-15 Cached

This study uses semantic entropy, an NLP embedding-based metric, to compare semantic memory navigation between blind and sighted individuals. Results show that visual experience influences entropy patterns, with sighted individuals having higher entropy for abstract concepts while blind individuals exhibit higher entropy for visually salient concrete concepts.

0 favorites 0 likes
#entropy

Diagnosing and Mitigating Thinking Collapse in On-Policy Self-Distillation

arXiv cs.CL · 2026-07-14 Cached

The paper identifies 'Thinking Collapse' in on-policy self-distillation for large language models, characterized by a decline in intermediate reasoning steps, and proposes AD-OPSD, a control framework that mitigates this collapse by anchoring high-suppression-risk tokens to a reference prior. The method achieves up to +4.1% absolute average accuracy improvement on mathematical benchmarks.

0 favorites 0 likes
#entropy

Accelerating Large Language Model Inference with Self-Supervised Early Exits

arXiv cs.CL · 2026-07-13 Cached

This paper introduces a self-supervised early exit method for LLMs, allowing computation to stop early at intermediate layers when confidence is high, thereby reducing inference cost. It also presents Dynamic Self-Speculative Decoding (DSSD) which achieves higher token acceptance than existing baselines.

0 favorites 0 likes
#entropy

I mapped Anthropic’s J-Space Hallucination signal across 7 datasets on Qwen3-4B to find out where it works and where it breaks

Reddit r/LocalLLaMA · 2026-07-12

This article evaluates Anthropic's J-Space hallucination detection method across 7 datasets on Qwen3-4B, finding it effective for catching high-confidence errors in factual retrieval but blind to internalized myths and failing on math tasks where thresholds don't transfer.

0 favorites 0 likes
#entropy

@omarsar0: A Visual Introduction to Information Theory (bookmark it) Information Theory is such an beautiful and powerful subject.…

X AI KOLs Following · 2026-07-09 Cached

An intuitive, visual introduction to information theory covering entropy, mutual information, and channel capacity, assuming only basic probability. The paper explains fundamental limits of compression and transmission.

0 favorites 0 likes
#entropy

Time was speeding up, slowing down, or even stopping

Reddit r/singularity · 2026-07-08

Researchers used a Bose-Einstein condensate as a mini universe to experimentally demonstrate that time can emerge from entropy exchange, supporting the theory of relational or emergent time in quantum physics.

0 favorites 0 likes
#entropy

Connections in Math: the two kinds of random

Hacker News Top · 2026-07-05 Cached

An exploration of why two statistically identical sequences—pure noise and the digits of π—differ in compressibility, distinguishing between statistical redundancy and process-based compression.

0 favorites 0 likes
#entropy

Making LLMs Better at Creative Writing using Entropy

Reddit r/LocalLLaMA · 2026-07-02 Cached

This article presents a technique to improve LLM creative writing by modifying the sampling process using entropy, aiming to reduce the generic 'LLM feel' in generated text.

0 favorites 0 likes
#entropy

@snowboat84: I've been pondering this for years: the relationship between statistical mechanics and AI. Statistical mechanics, using a statistical approach to molecular dynamics, reproduces the elegant fundamental theorems of thermodynamics, especially the beautiful relationships between macroscopic quantities like entropy, free energy, and of course temperature and pressure. The question is, does AI have these thermodynamic mac...

X AI KOLs Timeline · 2026-06-25 Cached

This tweet explores the relationship between statistical mechanics and artificial intelligence, citing a paper that proposes a thermodynamic theory for machine learning systems, introducing concepts like temperature, entropy, and energy, and treating the training process as a phase transition.

0 favorites 0 likes
#entropy

Beyond Entropy: Learning from Token-Level Distributional Deviations for LLM Reasoning

arXiv cs.AI · 2026-06-20 Cached

Introduces Independent Combinatorial Tokens (ICT) framework that uses Jensen-Shannon divergence between token logit distributions to identify critical branching points, preventing entropy collapse and explosion in RLVR for LLM reasoning. Achieves up to 14.9% pass@4 improvement on Qwen models.

0 favorites 0 likes
#entropy

@johnschulman2: PPO had a second wave in the LLM era for reasons unanticipated by the original paper - the importance-ratio objective f…

X AI KOLs Following · 2026-06-18 Cached

This paper reveals that the clipping mechanism in PPO and GRPO biases entropy in RLVR for LLMs: clip-low increases entropy, clip-high decreases it. The authors prove that standard clipping reduces entropy even with random rewards, and show that adjusting clip-low can prevent entropy collapse and promote exploration.

0 favorites 0 likes
#entropy

@snowboat84: Several years ago, dissipative systems and nonlinear complex systems were extremely popular in academic and cultural circles. To fully review dissipative systems, one must start with non-dissipative thermodynamics. The second law of thermodynamics (entropy law) states that everything should move towards chaos and stillness. But life grows, forests succeed, and even the large models in data centers are constantly "learning" order. …

X AI KOLs Timeline · 2026-06-18 Cached

This is a popular science article of over 25,000 characters, starting from the origin of entropy, reviewing the development of dissipative system theory, and exploring a three-level analysis of whether AI belongs to dissipative systems (hardware level, training level, static model).

0 favorites 0 likes
#entropy

Integrating Local and Global Entropy for Uncertainty Quantification in LLMs

arXiv cs.LG · 2026-06-10 Cached

This paper proposes Global-Local Uncertainty (GLU), an unsupervised single-pass score that fuses token-level local entropy with hidden-state geometric global entropy for uncertainty quantification in LLMs, showing that the two are near-orthogonal and together capture confident-but-wrong failures.

0 favorites 0 likes
#entropy

Breaking Entropy Bounds: Accelerating RL Training via MTP with Rejection Sampling

Hugging Face Daily Papers · 2026-06-10 Cached

Bebop proposes entropy-aware multi-token prediction with rejection sampling and a novel TV loss to accelerate RL training of LLMs, achieving up to 1.8x speedup. The method addresses the degradation of acceptance rates during RL by optimizing training objectives.

0 favorites 0 likes
#entropy

When Does Multi-Agent Collaboration Help? An Entropy Perspective

arXiv cs.AI · 2026-06-08 Cached

This paper examines multi-agent systems (MAS) from an entropy perspective, analyzing intra- and inter-agent dynamics. It finds that single agents often outperform MAS and introduces the Entropy Judger algorithm to improve MAS performance.

0 favorites 0 likes
#entropy

Entropy

Lobsters Hottest · 2026-06-07 Cached

A technical blog post exploring randomness, Linux entropy, and building a tool called morerandom that uses WASM plugins to feed the system entropy pool.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback