Harper Hua
My Chinese name is 花硕.
I am a M.S. student in Computer Science at Stanford University, where I am fortunate to work with Prof. Jure Leskovec and Stefano Ermon. Previously, I earned my B.S. from Tsinghua University. My primary research interests include post-training for LLM agents, reinforcement learning, and adapting/customizing agents for novel tasks and complex systems. I’m also broadly interested in long-context LLMs, privacy-preserving AI, and multimodal diffusion models.
In summer 2025, I was an Applied Scientist Intern at Amazon Web Service Bedrock, where I worked on multi-turn RL post-training for code generation.
I’m always happy to chat or collaborate. Feel free to reach out!
Email: shuohua [at] stanford.edu
selected publications
- ACLSQL-Trail: Multi-Turn Reinforcement Learning with Interleaved Feedback for Text-to-SQLIn ACL Main Conference, 2026
- ICML
- NeurIPSResearchCodeBench: Benchmarking LLMs on Implementing Novel Machine Learning Research CodeIn NeurIPS Dataset & Benchmark Track (Spotlight), 2025
- ICLR
teaching
Winter 2026 - Stanford University, CS246: Mining Massive Data Sets
Fall 2025, Fall 2024 - Stanford University, CS224W: Machine Learning with Graphs
Fall 2025, Fall 2024 - Stanford University, CS224W: Machine Learning with Graphs