Post Job Free
Sign in

AI Agent & LLM Developer (Data Scientist)

Location:
United States
Posted:
September 07, 2026

Contact this candidate

Resume:

Haoyang Yu

La Jolla, CA • 909-***-**** • **********@*****.***

EDUCATION

Harvard University, Cambridge, MA Sep 2026 - Expected Dec 2027 Master of Science in Data Science

University of California, San Diego, La Jolla, CA Sep 2022 - Mar 2026 Bachelor of Science in Data Science Minor in Cognitive Science & Mathematics GPA: 3.87/4.00 Achievement: Alice C. Tyler Scholarship (24-25); UCSD TOWN & GOWN (24-25, 25-26); Provost’s Honors for every quarter. Relevant Courses: Systems for Scalable Analytics, Data Visualization & Management, Probability & Statistics TECHNICAL SKILLS

Programming: Python, JavaScript, MATLAB, SQL

Machine Learning: PyTorch, scikit-learn, LLMs, Deep Learning, Transformer, Fine-tuning, Post-training, RLHF, SFT LLM / Agent: OpenClaw, Codex, Prompt Engineering, Context Engineering, Agent Workflow EXPERIENCE

PANGO VALVE LLC Remote - Houston, TX July 2026 – Present AI Application Assistant

● Developed an enterprise AI Agent platform for expense management using ERPNext, Frappe, Python, and LLMs, supporting expense applications, reimbursements, payment requests, multi-step approvals, and intelligent workflow automation.

● Designed intelligent workflow orchestration, role-based permission systems, and financial knowledge models supporting AI-assisted approval decisions for 40–80 enterprise users.

● Built LLM-powered document understanding pipelines integrating OCR, invoice parsing, document completeness verification, duplicate detection, policy compliance reasoning, and approval recommendation generation.

● Prototyped conversational AI interfaces and Agent-based task execution pipelines for enterprise expense workflows, enabling natural language interaction and future integration with external Agent frameworks. Hao AI Lab La Jolla, CA May 2024 – June 2026

Research Assistant Supervised by: Dr. Hao Zhang

● Built scalable evaluation pipelines for LLM Agents across 20+ foundation models and 10+ interactive environments, supporting 10,000+ GPU evaluation runs.

● Developed an open benchmarking platform, LMGame Bench, for evaluating LLM reasoning, decision-making, and multi-turn agent performance on Hugging Face.

● Designed modular multi-agent evaluation harnesses supporting reproducible interaction, tool execution, and multi-turn reasoning experiments.

● Co-developed Video ScienceBench, a benchmark for evaluating scientific understanding in video generation models, covering causal reasoning, temporal consistency, and physics-based scenarios. PARAMETRICS AI Remote — New York Aug 2025 – Nov 2025 Data Analyst

● Built ML pipelines for a healthcare data client, integrating clinical trial records, patient behavior logs, and genomic datasets to enable predictive modeling across 50K+ samples.

● Developed survival and logistic regression models to forecast treatment responses and adverse events, improving baseline accuracy by ~15% and supporting data-driven clinical decisions. Scripps Research - The Zorrilla Laboratory La Jolla, CA July 2024 – Dec 2024, June 2025 – Sep 2025 Research Intern Supervised by: Dr. Eric Zorrilla

● Extracted and standardized 354,000+ alcohol-related EHR records from the All of Us database by mapping ICD-9/10 codes to SNOMED CT via OMOP, aligned with the lab’s curated set of ~100 diagnoses; built SQL pipelines to merge datasets.

● Analyzed >200 behavioral assay datasets from murine stress models and hippocampal/amygdalar proteomics, applying ANOVA and regression to quantify stress-induced behavioral variance and molecular shifts.

● Designed reproducible workflows integrating data cleaning, table joins, and temporal aggregation, reducing missing/noisy entries by 30% and improving the reliability of downstream predictive models. PUBLICATIONS

Hu, L.; Huo, M.; Zhang, Y.; Yu, H.; Xing, E. P.; Stoica, I.; Rosing, T.; Jin, H.; Zhang, H.

“lmgame Bench: How Good Are LLMs at Playing Games?” arXiv preprint arXiv:2505.15146, May 2025. Accepted at ICLR 2026.

Zhang, Y.; Yu, H.; Hu, L.; Jin, H.; Zhang, H. “General Modular Harness for LLM Agents in Multi Turn Gaming Environments.” arXiv preprint arXiv:2507.11633, July 2025. Accepted at ICML 2025 Multi Agent Systems (MAS) Workshop. Yu, H.; Zhang, Y.; Hu, L.; Jin, H.; Zhang, H.

“VideoScience-Bench: Benchmarking Scientific Understanding and Reasoning for Video Generation.” arXiv preprint arXiv:2512.02942, Dec 2025. Accepted at ECCV 2026.



Contact this candidate