Articles from arXiv
ROOSTER is a shared module that learns alignment between condition and target sequences for time-series forecasting and PPG-to-vital-sign reconstruction, achieving superior performance across multiple benchmarks.
This paper proposes QRLQ, a quantum reinforcement learning framework that integrates parameterised quantum circuits with dueling double deep Q-networks to optimize cost and delay tradeoffs in quantum cloud orchestration, achieving lower costs and delays than heuristic baselines while using fewer parameters than classical DRL.
This paper proposes Task-Conditioned Latent Alignment (TCLA) to stabilize neural decoding across sessions in brain-machine interfaces by learning a shared latent space. Evaluated on nonhuman primate data, TCLA shows improved robustness compared to existing methods.
This paper proposes CC-OPD, a novel on-policy distillation method for multi-constraint instruction following that uses counterfactual ablations to enhance training signals, achieving superior performance where a 1.5B model surpasses its 7B teacher on benchmarks.
DualRes is a compact oscillatory state-space model for machine fault diagnosis from vibration data, achieving state-of-the-art performance with limited labels and reduced computational requirements for edge deployment.
This review article synthesizes active learning research for biodiversity monitoring, addressing label efficiency and highlighting the need for methods that support validation and reliable ecological inference.
FWBench introduces a benchmark for evaluating how language models select and use time-series forecasts to make cost-constrained decisions, comparing hosted and local configurations on electricity and cycle-hire datasets with efficient budget usage by GPT-6 Astra.
This paper introduces a framework using AUC bounds as a differentiable objective for anomaly-free self-optimization of anomaly detection systems, achieving performance gains over conventional model selection methods.
This paper proposes a quantization-robust unlearning framework for large language models, using loss landscape analysis to ensure effective forgetting while maintaining model utility after compression.
This paper introduces a discrete diffusion model using variational autoregressive networks to parameterize normalized probability distributions, applied to Ising models for accurate thermodynamic computations and enhanced Monte Carlo sampling.
This paper introduces KITE, a KV-invariant transformer expansion method that efficiently scales LLMs by reducing inference costs while maintaining performance. It presents the SST model that achieves lower training loss and reduced inference cost compared to baselines.
The paper presents Neurogenesis Network (NGN), a differentiable parameterization for learning the optimal size of neural networks during training, applicable to various architectures like MLPs, CNNs, and Transformers.
SR-Fraud is an outcome-supervised reflective LLM agent framework for non-stationary payment fraud detection, improving detection metrics over traditional methods on a production benchmark.
The paper proposes a spectral connectivity-regularized graph learning framework (SCoGL) that incorporates Laplacian spectral priors to improve graph recovery and downstream tasks like graph signal denoising when data is scarce.
This paper challenges the interpretation of the Platonic Representation Hypothesis by distinguishing between relational structure and metric geometry, showing that relational convergence is robust while metric geometry convergence is weaker in various models after calibration.
The paper proposes LLMAE, a method to repurpose pre-trained decoder-only LLMs as continuous text autoencoders using a latent bottleneck, achieving high-fidelity reconstruction and enabling downstream tasks like image captioning.
This paper proposes a full-covariance smoothing technique for Bayesian neural networks to enable efficient online adaptation by propagating correlations through nonlinear activations, demonstrated in tasks like classification and control.
Introduces CellAudit, a method to audit input-use claims in AI virtual cells by examining source code and predictive contributions, using falsification to bridge the prediction–claim gap in agentic model discovery.
This paper conducts a scaling study for fMRI foundation models, revealing that performance depends on the combination of pretraining data size, model size, and training duration, not just compute.
This paper proposes a tail-aware geometry learning framework for conformal ellipsoids that decouples tail sensitivity from coverage guarantees, improving uncertainty quantification in multivariate settings.