Skip to content
arXiv cs.LG · Papers

Smooth Learning with Hard Constraints via Legendre-Regularized Policies

arXiv:2607.24007v1 Announce Type: cross Abstract: We revisit contextual optimization from the perspective of policy class design. A desirable policy class should be expressive enough to learn rich context-decision relationships, should enforce hard feasibility constraints rather than soft penalty terms, and should rema