Federal grant · project grant (b)
Trustworthy Reinforcement Learning for Online Decision Making -this Research Project Will Advance the Frontiers of Modern Reinforcement Learning for Online Decision-making Problems. Reinforcement Learning Deals With How Intelligent Agents Ought to Take Actions in an Uncertain Environment in Order to Maximize the Cumulative Reward. It Has Achieved Phenomenal Success in Diverse Business and Scientific Fields. However, Less Attention Has Been Paid to the Trustworthy Aspects of Reinforcement Learning. This Project Will Investigate Different Aspects of Trustworthy Issues Like Robustness, Fairness, Causality, and Explainability in Important Online Decision-making Tasks. the Results of This Research Will Benefit Many Different Fields Such as Statistics, Machine Learning, Operations Research, Marketing, Economics, and Finance. Open-source Software Will Be Developed to Provide Applied Researchers With Cutting-edge Tools. the Project Will Recruit Students, Especially Those From Unrepresented Groups, to Be Involved in the Research and Will Develop New Courses on Statistical Reinforcement Learning and Decision Making. This Research Project Will Focus on Three Interconnected Trustworthy Reinforcement Learning Methods for Online Decision-making Problems: Dynamic Pricing, Dynamic Assortment Selection, and Matching in Two-sided Markets. Issues of Robustness, Fairness, Causality, and Explainability Will Be Addressed in These Decision-making Tasks, Which Will Advance the Exploration Techniques Used in Existing Reinforcement Learning Algorithms. the Project Will Develop New Theoretical Tools to Analyze the Statistical Properties of These Modern Reinforcement Learning Algorithms. Regret Upper Bounds and Matching Lower Bound Will Be Thoroughly Investigated. One Important Goal of Online Decision Making Is to Identify an Optimal Policy That Maximizes the Overall Gain, Based on the Contextual Information and Historical Interactions With the Environment. Due to the Complex Nature of Such Problems, There Is a High Demand for Trustworthy Tools for Learning Optimal Personalized Policy in Various Settings. the Knowledge Gained From This Research Will Benefit Learning in Online Auctions and Other Complex Market Design Problems. This Award Reflects NSF'S Statutory Mission and Has Been Deemed Worthy of Support Through Evaluation Using the Foundation's Intellectual Merit and Broader Impacts Review Criteria.
Committed
$450,000
Paid out
$416.8K
93%
Committed, not yet paid
$33.2K
7%
Loading…
Everything here is this single award's whole record — signed, amended, paid — not a fiscal-year slice. The by-year charts elsewhere split an award across the years it was committed; this page keeps it whole.
Committed is what the government has legally promised on this award so far. Contracts can also carry a ceiling — the maximum if every option is exercised. Unspent ceiling is headroom, not money owed.
The cash actually disbursed against this award. The gap from committed is the disbursement pipeline: promised, not yet cashed.
Each transaction is a signing event — an action that created or changed the award, dated the day it was signed — not a payment. Negative amounts are real: money de-committed at closeout or renegotiation.
One bar, the award’s whole arithmetic: paid out, then committed, not yet paid, then unspent ceiling.