news

Sep 02, 2026 TAing 18.404 Theory of Computation this fall.
Sep 01, 2026 Submitted our prefix-advantage study of self-distillation to the NeurIPS 2026 workshop on Transitioning from Pre-Training to Post-Training.
Aug 01, 2026 Tailored primitive initialization is the secret key to reinforcement learning was published at ACL 2026.
Jun 01, 2026 Started as a research intern at IBM Research, working on RL and self-distillation for LLM reasoning.
Jan 05, 2026 Spent January at Hudson River Trading as a WITTI wintern.
Aug 18, 2025 Our paper on a multi-agent framework for aesthetic and aerodynamic car design appeared at ASME IDETC-CIE 2025.
Jun 02, 2025 Joined Talkdesk as a software engineering intern working on multi-agent systems.
May 15, 2025 Started as a researcher at the MIT-IBM Watson AI Lab, working on RL for generative models. :tada: