A Hierarchy of Policy Learning Problems
arXiv:2607.03385v1 Announce Type: new Abstract: Policy learning has received substantial attention with the goal of learning policies from observational data for decision-making. A majority of work…
arXiv:2607.03385v1 Announce Type: new Abstract: Policy learning has received substantial attention with the goal of learning policies from observational data for decision-making. A majority of work…
arXiv:2607.03161v1 Announce Type: new Abstract: In selective deployment, practitioners act only on a model-chosen subset of individuals based on predicted conditional average treatment effects, but marginal…
arXiv:2607.02681v1 Announce Type: new Abstract: Integrating information across related tasks can improve estimation and prediction in transfer, multi-task, and federated learning, but contamination and heterogeneity make…
arXiv:2303.08777v3 Announce Type: replace Abstract: Cross-validation is one of the most widely used tools for risk estimation and model selection in statistics and machine learning, yet…
arXiv:2607.02247v1 Announce Type: cross Abstract: The aggregation with exponential weights (AEW) estimator is not fully understood in the basic setting of model selection aggregation with squared…
arXiv:2607.01749v1 Announce Type: cross Abstract: Despite increasing scale and resolution, many biological measurements remain destructive, revealing only spatial information rather than the dynamics it encodes. By…
arXiv:2607.01741v1 Announce Type: new Abstract: Reinforcement Learning (RL) is a sequential decision-making framework in which an agent learns optimal policies through interaction with an environment by…
arXiv:2607.02212v1 Announce Type: new Abstract: Aqueous solubility is a key property in early-stage drug discovery, but most predictive models merge physicochemical descriptors and molecular graph information…
arXiv:2607.02003v1 Announce Type: new Abstract: Although neural networks are remarkably effective, their underlying optimization principles remain theoretically elusive, often characterized by non-convex landscapes and stochastic heuristics.…
arXiv:2607.01959v1 Announce Type: new Abstract: We propose a model agnostic methodology to measure lag relevance in machine learning forecasting models applied to univariate time series. Particularly,…