hyperparameter-optimization

Tag

Cards List
#hyperparameter-optimization

AgentHPOBench: A Benchmark For Evaluating LLM Agents as Sequential Hyperparameter Optimizers

arXiv cs.AI · yesterday Cached

Introduces AgentHPOBench, a sequential benchmark for evaluating LLM agents as hyperparameter optimizers across 30 machine learning tasks, showing current agents have measurable but limited iterative refinement abilities.

0 favorites 0 likes
#hyperparameter-optimization

Efficient Heteroscedastic Bayesian Optimization for Risk-Aware AutoRL

arXiv cs.LG · 5d ago Cached

Proposes ERAHBO, an efficient heteroscedastic Bayesian optimization method for risk-aware hyperparameter optimization in reinforcement learning, using adaptive re-sampling to improve sample efficiency over fixed-budget approaches.

0 favorites 0 likes
#hyperparameter-optimization

HOBA: Hierarchical On-Policy Bidding Agents for Adaptive Online Advertising

arXiv cs.AI · 6d ago Cached

HOBA proposes a hierarchical reinforcement learning framework for online advertising that uses a large language model for hyperparameter inference, a SARSA agent for expert model selection, and a dynamic expert pool for bid execution, achieving a +3.6% improvement in a large-scale A/B test.

0 favorites 0 likes
#hyperparameter-optimization

Optimizing Large Language Models for Causality Assessment in Pharmacovigilance: Developing a Performance Metric as Objective for Bayesian Hyperparameter Optimization

arXiv cs.CL · 2026-07-07 Cached

This paper presents a method to optimize GPT-5.2 temperature for Naranjo causality assessment in pharmacovigilance, achieving significant agreement improvements via Bayesian hyperparameter optimization with a novel composite metric (EWACS).

0 favorites 0 likes
#hyperparameter-optimization

SemiScope: Disentangling Classifier Tuning and Joint Optimization in Semi-Supervised Security Classification

arXiv cs.LG · 2026-07-02 Cached

This paper introduces SemiScope, an analysis tool designed to disentangle the effects of classifier tuning from joint SSL and classifier optimization in semi-supervised security classification. Results show that most performance gains attributed to joint optimization can be recovered by simply tuning the classifier and its decision threshold with Bayesian optimization.

0 favorites 0 likes
#hyperparameter-optimization

How Good Can Linear Models Be for Time-Series Forecasting?

Hugging Face Daily Papers · 2026-06-25 Cached

This paper demonstrates that careful preprocessing—especially context length selection, normalization, and regularization—can make simple linear models like Ridge regression competitive with or superior to large Transformer, MLP, and CNN models on time-series forecasting benchmarks.

0 favorites 0 likes
#hyperparameter-optimization

LLMZero: Discovering Adaptive Training Strategies for RL Post-Training via LLM Agents

arXiv cs.LG · 2026-06-18 Cached

LLMZero uses LLM agents to search over training trajectories via tree search, discovering adaptive multi-parameter transitions for RL post-training that outperform fixed schedules and grid search across diverse tasks.

0 favorites 0 likes
#hyperparameter-optimization

Optuna Constrained Tree-Structured Parzen Estimator Is a Joint Density Generalization of c-TPE

arXiv cs.LG · 2026-06-10 Cached

This paper demonstrates that Optuna's constrained Tree-Structured Parzen Estimator (TPE) is a joint density generalization of the c-TPE algorithm, showing its invariance to constraint duplication while independent c-TPE degrades. The authors outline practical tradeoffs and directions for future study.

0 favorites 0 likes
#hyperparameter-optimization

Hyperparameter Learning for Latent Factorization of Tensors for Representation Learning to Large-scale Dynamic Weighted Directed Network

arXiv cs.LG · 2026-06-10 Cached

This paper proposes an automated hyperparameter optimization framework based on Differential Evolution for Latent Factorization of Tensors (LFT) to improve prediction accuracy on large-scale dynamic weighted directed networks, reducing the need for manual tuning.

0 favorites 0 likes
#hyperparameter-optimization

Staged Factorial Screening for Budget-Constrained Micro-Pretraining

arXiv cs.LG · 2026-06-05 Cached

This paper proposes a staged factorial screening workflow for budget-constrained micro-pretraining, demonstrating that short designed experiments can identify stable hyperparameter penalty directions and support a screen-then-refine strategy.

0 favorites 0 likes
#hyperparameter-optimization

Auto Research with Specialist Agents Develops Effective and Non-Trivial Training Recipes

Hugging Face Daily Papers · 2026-05-07 Cached

This paper introduces an auto-research framework using specialist agents to iteratively refine training recipes through an empirical loop of code execution and feedback. The system autonomously improves performance on tasks like Parameter Golf and NanoChat without human intervention by leveraging lineage feedback.

0 favorites 0 likes
← Back to home

Submit Feedback