arXiv cs.LG
· Papers
Smooth Learning with Hard Constraints via Legendre-Regularized Policies
arXiv:2607.24007v1 Announce Type: cross Abstract: We revisit contextual optimization from the perspective of policy class design. A desirable policy class should be expressive enough to learn rich context-decision relationships, should enforce hard feasibility constraints rather than soft penalty terms, and should rema