Thoughts on LLM systems, tutorials, and project updates.
Focused on making large language models faster, cheaper, and more accessible.
Building serving frameworks for TTS / Omni model architectures
Runtime optimization and GPU performance profiling
Created and presented official SGLang tutorial videos (Diffusion, Cookbook). Expanded test coverage for OpenAI-compatible API endpoints across multiple PRs.
Develop the official SGLang course and write technical content for the SGLang blog. Support the SGLang developer community and deliver technical talks at industry conferences.
Bellevue, WA
Led Tableau Mobile end-to-end feature efforts. Delivered TabAgent, an embedded AI assistant for Tableau serving millions of users. Built a LangGraph AI agent automating bug-blitz processes, improving UX validation efficiency by 50%+.
Seattle, WA
Implemented Tableau-Pulse features (React Native + Redux) shipping to 100k+ customers.
Santa Clara, CA
Built an AI content assistant (ChatGPT APIs) generating social posts from artist prompts, reducing content-creation time by 80% and serving 10k+ artists.
Shandong, China
Deployed production-grade extraction models on cloud inference servers. Built a LangChain + Qwen agent to normalize heterogeneous EMR formats.
University of Virginia (UVA)
Graduated with High Distinction.
LLM Inference & Systems
Infra & Tools
Programming Languages