HuggingFace

Articles from HuggingFace

Cards List

LimiX-2: A Contextual Mechanism Network Towards General Structured-Data Intelligence

Hugging Face Daily Papers · 2026-09-15 Cached

LimiX-2 is a pretrained foundation model for structured data that uses contextual mechanism networks to achieve #1 on major tabular benchmarks, supporting multiple tasks without task-specific parameter updates.

0 favorites 0 likes

EventEgoHands++: Event-based Egocentric 3D Hand Mesh Reconstruction with Real Dataset

Hugging Face Daily Papers · 2026-09-15 Cached

This paper proposes EventEgoHands++, a framework for event-based 3D hand mesh reconstruction from an egocentric viewpoint, incorporating hand detection and adaptive attention, and introduces a new real-world dataset EEH-R for training and evaluation.

0 favorites 0 likes

FLAT: Resampling Image and Text into 1D Flexible-Length Aligned Transmodal Tokens for Retrieval and Generation

Hugging Face Daily Papers · 2026-09-15 Cached

FLAT is a representation pre-training framework that jointly optimizes a shared multimodal encoder with downstream decoders for text-to-image and image-to-text tasks, using flexible-length aligned 1D sequences to enable cross-modal retrieval and generation with state-of-the-art results.

0 favorites 0 likes

ImpossibleRubrics: Stress-Testing Generated Rubrics as Reward Signals

Hugging Face Daily Papers · 2026-09-15 Cached

This paper introduces ImpossibleRubrics, an open-source evaluation framework that uses fine-grained, adversarial rubrics to stress-test large language models, aiming to reduce evaluation bias and expose hidden failure modes.

0 favorites 0 likes

ScienceBuddy: Recursive-in-Recursive Self-Improvement for Interactive Scientific Agents

Hugging Face Daily Papers · 2026-09-15 Cached

ScienceBuddy introduces a recursive-in-recursive self-improvement paradigm for interactive scientific agents, enabling continual evolution through researcher collaboration and feedback.

0 favorites 0 likes

AI for Games in the Foundation Model Era

Hugging Face Daily Papers · 2026-09-15 Cached

This paper surveys the use of foundation models in game AI across roles like playing, modeling, design, and evaluation, highlighting transferability challenges and the need for game-specific validation.

0 favorites 0 likes

Qwen/Qwen-Image-2.1

Hugging Face Models Trending · 2026-09-14 Cached

Qwen-Image-2.1 is an open-source unified text-to-image and image editing model with 7B parameters, featuring efficient architecture, transparency support, and versatile editing capabilities.

0 favorites 0 likes

Bellman Policy Optimization

Hugging Face Daily Papers · 2026-09-14 Cached

Bellman Policy Optimization (BPO) is a critic-free reinforcement learning method that reformulates Policy Mirror Descent using the Bellman equation for autoregressive generation with terminal rewards, improving mathematical reasoning in large language models.

0 favorites 0 likes

MoME: Mixture-of-Memory Embeddings for Context-Aware Sparse Lookup

Hugging Face Daily Papers · 2026-09-14 Cached

Introduces MoME, a context-aware memory mechanism for LLMs that uses a mixture of slots to handle token polysemy, improving over baselines in pretraining experiments.

0 favorites 0 likes

EvoOntology: A Self-Evolving Ontology Layer for Data Agents

Hugging Face Daily Papers · 2026-09-14 Cached

EvoOntology introduces a self-evolving ontology layer for data agents, encapsulated as an MCP server, to bridge the agent-data gap and improve performance on heterogeneous data tasks as shown in benchmarks.

0 favorites 0 likes

Assessing nnU-Net Generalization across Brain Tumor Populations in BraTS-GoAT 2026

Hugging Face Daily Papers · 2026-09-14 Cached

This paper assesses the generalization of nnU-Net for brain tumor segmentation in the BraTS-GoAT 2026 challenge, reporting performance metrics and analyzing failure cases.

0 favorites 0 likes

HypoEvolve: Genetic Algorithms Enable Multi-Agent LLMs to Discover Scientific Hypotheses

Hugging Face Daily Papers · 2026-09-14 Cached

HypoEvolve introduces a framework using genetic algorithms to coordinate multi-agent LLMs for scientific hypothesis discovery, demonstrated through drug repurposing in cancer research with improved performance over baselines.

0 favorites 0 likes

HarnessVLN: Unifying Training-Free Embodied Navigation through an Agent Harness

Hugging Face Daily Papers · 2026-09-14 Cached

HarnessVLN is a zero-shot, training-free framework for embodied navigation that unifies perception, retrieval, grounding, navigation, recovery, and termination through a unified tool interface, achieving state-of-the-art results on benchmarks like R2R and RxR.

0 favorites 0 likes

ModularRSI: Modular and Generalizable Recursive Harness Self-Improvement

Hugging Face Daily Papers · 2026-09-14 Cached

ModularRSI introduces a modular and generalizable framework for recursive self-improvement in AI agent harnesses, using contrastive learning across tasks to evolve modules independently and enhance performance on unseen tasks.

0 favorites 0 likes

The Router Within: Eliciting Native Skill Routing from a Frozen LLM

Hugging Face Daily Papers · 2026-09-14 Cached

The paper introduces Gavel, a method that elicits native skill routing from frozen LLMs via linear projections, enabling efficient tool selection without context overload and outperforming existing pipelines on benchmarks.

0 favorites 0 likes

Disentangling Representation Evolution in Transformers through Directional Decomposition

Hugging Face Daily Papers · 2026-09-14 Cached

This paper decomposes transformer representation updates into parallel and perpendicular components to study evolution geometry, linking it to editing robustness, compression diagnosis, and training improvements.

0 favorites 0 likes

ModaLens: Measuring Image Sensitivity in Report-Conditioned Medical VLMs

Hugging Face Daily Papers · 2026-09-14 Cached

ModaLens is a paired image-swap audit that measures how report availability reduces image sensitivity in medical vision-language models, demonstrated using MedGemma-27B on the MIMIC-CXR dataset.

0 favorites 0 likes

How Lossless Is Lossless Speculative Decoding? The Role of Numerical Precision in Orthrus

Hugging Face Daily Papers · 2026-09-14 Cached

This paper investigates the losslessness of Orthrus, a hybrid autoregressive-diffusion model for inference acceleration, finding that it requires high numerical precision (FP32) for exact trajectory matching, while BF16 divergence does not impair downstream performance.

0 favorites 0 likes

LynnReal-Omni: Native multi-modal Video Generation for Agentic Visual Workflows

Hugging Face Daily Papers · 2026-09-14 Cached

LynnReal-Omni is a unified multimodal video diffusion framework that integrates agentic visual controls with high-fidelity generation and real-time acceleration for stable, controllable video creation.

0 favorites 0 likes

Not All Prompts Are Equal: Exploration-Guided Prompt Scaffolding for Multimodal Reinforcement Post-Training

Hugging Face Daily Papers · 2026-09-14 Cached

This paper proposes an exploration-guided prompt scaffolding framework that dynamically adjusts training prompts for multimodal reinforcement learning, achieving up to 9.7% relative improvement in performance on benchmarks.

0 favorites 0 likes
← Previous
Next →
← Back to home

Submit Feedback