Undergraduate researcher working on post-training, reinforcement learning, and optimization.
This is a page not in the menu. You can use markdown in this page.