multi-objective

Tag

Cards List
#multi-objective

Efficient Online Lexicographic Generalized Low-Rank Matrix Bandits

arXiv cs.LG · 2026-08-06 Cached

This paper introduces Lexi-LowGLM, an efficient algorithm for generalized low-rank matrix bandits with multiple prioritized objectives, using online Newton updates to reduce computational complexity and achieve regret bounds depending on effective low-rank dimensions.

0 favorites 0 likes
#multi-objective

Top-$k$ Pareto Bandits: Hypervolume Regret for Multi-Objective Slate Selection

arXiv cs.LG · 2026-07-30 Cached

This paper introduces THV-UCB, an algorithm for multi-objective bandit problems with slate selection, and establishes gap-free and gap-dependent regret bounds for hypervolume regret.

0 favorites 0 likes
#multi-objective

Prompt Optimization for User Simulation in Conversational Recommender Systems: A Multi-Objective Framework

arXiv cs.AI · 2026-07-02 Cached

This paper proposes a framework to automatically optimize prompts for LLM-based user simulators in conversational recommender systems, addressing issues like positive bias and limited behavioral diversity.

0 favorites 0 likes
#multi-objective

Deterministic Pareto-Optimal Policy Synthesis for Multi-Objective Reinforcement Learning

arXiv cs.LG · 2026-06-26 Cached

This paper introduces a novel preference-conditioned Bellman operator based on Chebyshev scalarization to compute deterministic Pareto-optimal policies for Multi-Objective Markov Decision Processes, proving its convergence and effectiveness in capturing the entire Pareto frontier.

0 favorites 0 likes
#multi-objective

Breaking the Filter Bubble: A Semantic Pareto-DQN Framework for Multi-Objective Recommendation

arXiv cs.AI · 2026-06-24 Cached

Proposes a multi-objective reinforcement learning framework combining semantic embeddings with Pareto-DQN to balance engagement, diversity, and fairness in recommendations, mitigating filter bubbles.

0 favorites 0 likes
#multi-objective

Holistic Data Scheduler for LLM Pre-training via Multi-Objective Reinforcement Learning

Hugging Face Daily Papers · 2026-06-23 Cached

Introduces Holistic Data Scheduler (HDS), a reinforcement learning-based framework that dynamically adjusts data mixtures during LLM pre-training using a multi-objective reward function, achieving 44% fewer iterations to reach target perplexity and a 7.2% improvement on MMLU.

0 favorites 0 likes
#multi-objective

Optimizing Lithium Production Decisions under Geological, Demand, and Pricing Uncertainties: A POMDP Framework for Multi-Objective Decision Making

arXiv cs.AI · 2026-06-18 Cached

This paper proposes a POMDP framework for multi-objective decision making in lithium production, addressing geological, demand, and pricing uncertainties to optimize mine opening and extraction method selection. The approach outperforms human-inspired heuristics by dynamically adapting to shifting price regimes through belief state planning.

0 favorites 0 likes
#multi-objective

CRAFT: Cost-aware Refinement And Front-aware Tuning of Prompts

arXiv cs.CL · 2026-06-04 Cached

CRAFT is a Pareto-front prompt optimizer that jointly optimizes for accuracy and token cost, avoiding the 'scalarization collapse' of weighted-sum approaches by maintaining a diverse population of prompts across the accuracy-cost trade-off frontier using NSGA-II and budget-aware validation.

0 favorites 0 likes
#multi-objective

WeCon: An Efficient Weight-Conditioned Neural Solver for Multi-Objective Combinatorial Optimization Problems

arXiv cs.LG · 2026-05-25 Cached

Presents WeCon, a weight-conditioned neural solver for multi-objective combinatorial optimization problems that achieves comparable hypervolume to the state-of-the-art while reducing inference time by 40%.

0 favorites 0 likes
#multi-objective

Multi-Objective Constraint Inference using Inverse reinforcement learning

arXiv cs.AI · 2026-05-11 Cached

This paper introduces MOCI, a novel framework for inferring shared constraints and individual preferences from heterogeneous expert demonstrations in reinforcement learning, outperforming existing baselines in predictive performance and computational efficiency.

0 favorites 0 likes
← Back to home

Submit Feedback