TensorX
返回文献探索

Paper · arXiv 2501.14912

Feasible Learning

Juan Ramirez, Ignacio Hounie, Juan Elenter, Jose Gallego-Posada, Meraj Hashemizadeh, Alejandro Ribeiro, Simon Lacoste-Julien

5 upvotesJanuary 24, 2025arXiv 预印本
AI 摘要

Feasible Learning trains models by ensuring satisfactory performance on each sample, using a primal-dual approach and minimal norm slack variables, improving tail behavior with slight impact on average performance.

Feasible Learningsample-centric learningfeasibility problemEmpirical Risk Minimizationprimal-dual approachslack variables

Abstract

We introduce Feasible Learning (FL), a sample-centric learning paradigm where models are trained by solving a feasibility problem that bounds the loss for each training sample. In contrast to the ubiquitous Empirical Risk Minimization (ERM) framework, which optimizes for average performance, FL demands satisfactory performance on every individual data point. Since any model that meets the prescribed performance threshold is a valid FL solution, the choice of optimization algorithm and its dynamics play a crucial role in shaping the properties of the resulting solutions. In particular, we study a primal-dual approach which dynamically re-weights the importance of each sample during training. To address the challenge of setting a meaningful threshold in practice, we introduce a relaxation of FL that incorporates slack variables of minimal norm. Our empirical analysis, spanning image classification, age regression, and preference optimization in large language models, demonstrates that models trained via FL can learn from data while displaying improved tail behavior compared to ERM, with only a marginal impact on average performance.

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号