LLM/VLM Reasoning and Post-training
I am highly interested in post-training for LLMs and VLMs. My focus is on developing more effective reinforcement learning algorithms, as well as leveraging simple and interpretable methods, to improve reasoning capabilities through post-training.




