Not All Objectives Are Born Equal: Priority-Constrained Descent for Hierarchical Multi-Objective Optimization
Abstract
Deep learning problems rarely involve objectives that are equal in importance. A primary objective defines the goal, whilst secondary objectives, such as sparsity, compression, or robustness constrain the solution. While existing multi-objective methods have proven effective in practice, they have a clear symmetry problem and neglect the inherent objective hierarchy built into these objective spaces. We introduce Priority-Constrained Descent (PCD), a gradient-based optimization framework designed to explicitly exploit hierarchical objective structures. PCD preserves the direction of primary descent whilst allowing for the minimal distortion necessary to guarantee progress on secondary objectives, controlled by a single τin [0, 1] that dictates the strength of the distortion. The resulting formulation is invariant to objective scaling and admits exact closed-form solutions for problems with two and three objectives. We evaluate PCD within structured network compression settings, unstructured sparsity and low-rankness, and across a variety of synthetic experiments, showing Pareto dominance and better per-objective performance with secondary progress guarantees over existing methods, further exhibiting the interpretable trade-off that τ provides.
Community
We propose Priority-Constrained Descent (PCD) for training deep learning problems where a hierarchy of objectives exists. One objective serves as our primary (e.g., accuracy) while others, such as regularizers, etc... "shape" the solution. At every step, PCD follows the primary gradient as closely as possible, while guaranteeing each secondary objective at least a fraction τ of progress.
Code (PyTorch, drop-in replacement for loss.backward()): https://github.com/DaraVaram/priority-constrained-descent
Colab tutorial: https://colab.research.google.com/github/DaraVaram/priority-constrained-descent/blob/main/notebooks/pcd_tutorial.ipynb
Project page with an 8-minute explainer video: https://daravaram.github.io/PCD/
Published in Transactions on Machine Learning Research (TMLR), 2026
Get this paper in your agent:
hf papers read 2606.29521 Don't have the latest CLI?
curl -LsSf https://hf.co/cli/install.sh | bash Models citing this paper 0
No model linking this paper
Datasets citing this paper 0
No dataset linking this paper
Spaces citing this paper 0
No Space linking this paper
Collections including this paper 0
No Collection including this paper