Reward Shaping for (Inference-Time) Alignment: A Stackelberg Game Perspective
Published in ICML 2026, 2026
The contents above will be part of a list of publications, if the user clicks the link for the publication than the contents of section will be rendered as a full page, allowing you to provide more information about the paper for the reader. When publications are displayed as a single page, the contents of the above “citation” field will automatically be included below this section in a smaller font.
Recommended citation: Wang, H., Lin, T., Kong, L., Li, C., Jiang, H., & Tambe, M. Reward Shaping for (Inference-Time) Alignment: A Stackelberg Game Perspective. ICML 2026.
Download Paper
