PMGDA: A Preference-based Multiple Gradient Descent Algorithm

It is desirable in many multi-objective machine learning applications, such as multi-task learning with conflicting objectives, multi-objective reinforcement learning, to find a Pareto solution that can match a given preference of a decision maker. These problems are often large-scale with available gradient information but cannot be handled very well by the existing algorithms. To tackle this issue, this paper proposes a novel predict-and-correct framework for locating a Pareto solution that fits the preference of a decision maker. In the proposed framework, a constraint function is introduced in the search progress to align the solution with a user-specific preference, which can be optimized simultaneously with multiple objective functions. Experimental results show that our proposed method can efficiently find a particular Pareto solution under the demand of a decision maker for standard multiobjective benchmark, multi-task learning, multi-objective reinforcement learning problems with more than thousands of decision variables.

Paper

References (64)

Scroll for more · 38 remaining

Similar papers

© 2026 NYSGPT2525 LLC