I’m Lily Zhang, a Tech Lead and Research Scientist in the Bay Area, California. My research focus: agentic post-training and AI training data, including RL environments. I’ve published at CVPR, ICCV, NeurIPS, RAL [1] [2] [3] [4] [5], and gave keynotes at NVIDIA GTC and IEEE IROS.
I like turning research into products people use, open-sourced: SofaGenius (Anthropic hackathon finalist, top 15/13k), Frontend Slides (25K+ GitHub stars), and the Deep-Seek research agent (500+ GitHub stars). At Latitude AI, I build LLMs, multimodal LLMs, and AI agents for physical AI (Modal GTC panel on agentic post-training).
Recent work: C-Guard, constitution-grid data-efficient RL alignment (COLM 2026 ER); AURA, RLVR reward shaping from a hierarchical failure taxonomy; Eureka: feature engineering as agentic code generation w Alibaba Cloud, SFT + RL post-trained AI-infra agent deployed in production (NeurIPS-W 2025, DASFAA 2026). Publishing as Xianling Zhang.
Find all published research here.