arXiv

Articles from arXiv

Cards List

Learning Where to Look: A Shared Relative-Alignment Module for Time-Series Forecasting and PPG-to-Vital-Sign Reconstruction

arXiv cs.LG ↗ · 23h ago Cached

ROOSTER is a shared module that learns alignment between condition and target sequences for time-series forecasting and PPG-to-vital-sign reconstruction, achieving superior performance across multiple benchmarks.

0 favorites 0 likes

Quantum Reinforcement Learning for Cost and Delay Tradeoffs in Quantum Cloud Orchestration

arXiv cs.LG ↗ · 23h ago Cached

This paper proposes QRLQ, a quantum reinforcement learning framework that integrates parameterised quantum circuits with dueling double deep Q-networks to optimize cost and delay tradeoffs in quantum cloud orchestration, achieving lower costs and delays than heuristic baselines while using fewer parameters than classical DRL.

0 favorites 0 likes

Stable Neural Decoding Across Sessions via Task-Conditioned Latent Alignment for Brain-Machine Interfaces

arXiv cs.LG ↗ · 23h ago Cached

This paper proposes Task-Conditioned Latent Alignment (TCLA) to stabilize neural decoding across sessions in brain-machine interfaces by learning a shared latent space. Evaluated on nonhuman primate data, TCLA shows improved robustness compared to existing methods.

0 favorites 0 likes

Counterfactual Constraint-Conditioned On-Policy Distillation for Multi-Constraint Instruction Following

arXiv cs.LG ↗ · 23h ago Cached

This paper proposes CC-OPD, a novel on-policy distillation method for multi-constraint instruction following that uses counterfactual ablations to enhance training signals, achieving superior performance where a 1.5B model surpasses its 7B teacher on benchmarks.

0 favorites 0 likes

When Labels Are Scarce: An Oscillatory State Space Model for Vibration Diagnosis

arXiv cs.LG ↗ · 23h ago Cached

DualRes is a compact oscillatory state-space model for machine fault diagnosis from vibration data, achieving state-of-the-art performance with limited labels and reduced computational requirements for edge deployment.

0 favorites 0 likes

Active Learning for Biodiversity Monitoring: From Label Efficiency to Reliable Ecological Inference

arXiv cs.LG ↗ · 23h ago Cached

This review article synthesizes active learning research for biodiversity monitoring, addressing label efficiency and highlighting the need for methods that support validation and reliable ecological inference.

0 favorites 0 likes

Forecast Workflow Bench: Evaluating Language-Model Decisions with Budgeted Forecast Tools

arXiv cs.LG ↗ · 23h ago Cached

FWBench introduces a benchmark for evaluating how language models select and use time-series forecasts to make cost-constrained decisions, comparing hosted and local configurations on electricity and cycle-hire datasets with efficient budget usage by GPT-6 Astra.

0 favorites 0 likes

Anomaly-Free Self-Optimization via AUC Bounds

arXiv cs.LG ↗ · 23h ago Cached

This paper introduces a framework using AUC bounds as a differentiable objective for anomaly-free self-optimization of anomaly detection systems, achieving performance gains over conventional model selection methods.

0 favorites 0 likes

Quantization-Robust Unlearning through the Lens of Retain-Forget Loss Landscapes Interaction

arXiv cs.LG ↗ · 23h ago Cached

This paper proposes a quantization-robust unlearning framework for large language models, using loss landscape analysis to ensure effective forgetting while maintaining model utility after compression.

0 favorites 0 likes

Discrete Diffusion Models via Evolving Variational Autoregressive Networks

arXiv cs.LG ↗ · 23h ago Cached

This paper introduces a discrete diffusion model using variational autoregressive networks to parameterize normalized probability distributions, applied to Ising models for accurate thermodynamic computations and enhanced Monte Carlo sampling.

0 favorites 0 likes

KITE: KV-Invariant Transformer Expansion for Efficient Agentic LLM Scaling

arXiv cs.LG ↗ · 23h ago Cached

This paper introduces KITE, a KV-invariant transformer expansion method that efficiently scales LLMs by reducing inference costs while maintaining performance. It presents the SST model that achieves lower training loss and reduced inference cost compared to baselines.

0 favorites 0 likes

NGN: Learning Neural Network Size as a Differentiable Count

arXiv cs.LG ↗ · 23h ago Cached

The paper presents Neurogenesis Network (NGN), a differentiable parameterization for learning the optimal size of neural networks during training, applicable to various architectures like MLPs, CNNs, and Transformers.

0 favorites 0 likes

SR-Fraud: An Outcome-Supervised Reflective LLM Agent Framework for Non-Stationary Payment Fraud Detection

arXiv cs.LG ↗ · 23h ago Cached

SR-Fraud is an outcome-supervised reflective LLM agent framework for non-stationary payment fraud detection, improving detection metrics over traditional methods on a production benchmark.

0 favorites 0 likes

Graph Learning with Spectral Connectivity Priors for Scarce Data

arXiv cs.LG ↗ · 23h ago Cached

The paper proposes a spectral connectivity-regularized graph learning framework (SCoGL) that incorporates Laplacian spectral priors to improve graph recovery and downstream tasks like graph signal denoising when data is scarce.

0 favorites 0 likes

What Converges in the Platonic Representation Hypothesis? Structure over Geometry

arXiv cs.LG ↗ · 23h ago Cached

This paper challenges the interpretation of the Platonic Representation Hypothesis by distinguishing between relational structure and metric geometry, showing that relational convergence is robust while metric geometry convergence is weaker in various models after calibration.

0 favorites 0 likes

Repurposing Pre-trained LLMs as High Fidelity Continuous Text Autoencoders

arXiv cs.LG ↗ · 23h ago Cached

The paper proposes LLMAE, a method to repurpose pre-trained decoder-only LLMs as continuous text autoencoders using a latent bottleneck, achieving high-fidelity reconstruction and enabling downstream tasks like image captioning.

0 favorites 0 likes

Full-Covariance Smoothing of Bayesian Neural Networks for Online Adaptation

arXiv cs.LG ↗ · 23h ago Cached

This paper proposes a full-covariance smoothing technique for Bayesian neural networks to enable efficient online adaptation by propagating correlations through nonlinear activations, demonstrated in tasks like classification and control.

0 favorites 0 likes

Discover, Falsify, Revise: Auditing Input-Use Claims from Source Code to Predictive Contribution in Agent-Discovered Cell Models

arXiv cs.LG ↗ · 23h ago Cached

Introduces CellAudit, a method to audit input-use claims in AI virtual cells by examining source code and predictive contributions, using falsification to bridge the prediction–claim gap in agentic model discovery.

0 favorites 0 likes

A Scaling Study for fMRI Foundation Models

arXiv cs.LG ↗ · 23h ago Cached

This paper conducts a scaling study for fMRI foundation models, revealing that performance depends on the combination of pretraining data size, model size, and training duration, not just compute.

0 favorites 0 likes

Tail-Aware Geometry Learning for Conformal Ellipsoids

arXiv cs.LG ↗ · 23h ago Cached

This paper proposes a tail-aware geometry learning framework for conformal ellipsoids that decouples tail sensitivity from coverage guarantees, improving uncertainty quantification in multivariate settings.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback