Publications
I am broadly interested in both the theoretical limits and empirical applications of reinforcement learning and online learning (e.g., multi-armed bandits), with a current focus on designing practical algorithms with provable guarantees for LLMs. Feel free to reach out if you share similar interests!
Conferences
SP²ec: Adaptive Self-Speculative Decoding for Vision-Language Models
