Haozhe Zhang

Modeling Engineer at Tesla · LLMs & Physics-based ML · creator of BenchCAD

profile_pic.jpg

Palo Alto, CA

USA

Hey, welcome to my page! 👋

I’m a modeling engineer at Tesla, working at the intersection of physics-based modeling and machine learning for next-generation battery cells — theoretical and computational modeling, reinforcement learning for automation, and data-driven optimization at scale-up.

My background is heavy on physics. I’ve spent years working out closed-form theoretical solutions across mechanics and multi-physics fields — soft-hard material integration, stretchable metasurfaces, mechanical Janus structures, nanoconfined fluid mechanics. I pair the theory with hands-on engineering, so most of what I’ve built has gone end-to-end from the math to a working system.

Outside the day job, what pulls me from one thing to the next is curiosity and a long-running urge to build new things. I keep gravitating toward LLMs, agentic systems, RL training infrastructure, voice agents for narrow domains, and AI for hardware. One of those turned into a project worth pointing at:

BenchCAD — a multimodal benchmark for programmatic CAD, the entrance to AI for hardware. A model sees four orthographic views of an industrial part and has to write the CadQuery program that rebuilds it; we score the re-executed geometry. 17,900 execution-verified parts across 106 industrial families, half of them anchored to real ISO / DIN / EN / ASME / IEC specification tables rather than free-form shapes. Both frontier labs now report it on their flagship models — Anthropic in three Claude system cards (Fable 5 / Mythos 5, Sonnet 5, Opus 5) and OpenAI in the GPT-5.6 launch table, next to OSWorld and BrowseComp. When both labs race the same benchmark and publish the gains themselves, AI for hardware has become a strategic front. benchcad.com · arXiv:2605.10865 Before Tesla, I earned my PhD in mechanical engineering from the University of Virginia and a BS from the University of Science and Technology of China. My doctoral work centered on theoretical and computational modeling of multi-physics fields.

In my spare time, I enjoy traveling and playing Leagues.

news

Jul 24, 2026 BenchCAD appears again in Anthropic’s Claude Opus 5 System Card, which reports Opus 5 on the Vision2Code subset — 0.366 voxel IoU without tools, 0.821 with — and upstreams two fixes to our reference implementation. BenchCAD Vision2Code subset scores for GPT-5.6 Sol and four Claude models, with and without tools From Anthropic’s Claude Opus 5 System Card.
Jul 9, 2026 BenchCAD is now reported by a second frontier lab: OpenAI lists it in the GPT-5.6 launch table, next to OSWorld and BrowseComp — with Anthropic’s Claude system cards, that makes BenchCAD a benchmark both labs measure themselves against. Benchmark: benchcad.com. BenchCAD scores for GPT-5.6 Sol, Luna, Terra and GPT-5.5, with and without a Python tool, as published in OpenAI's GPT-5.6 launch table BenchCAD rows from OpenAI’s GPT-5.6 launch table.
Jun 30, 2026 BenchCAD is featured again as a multimodal benchmark in Anthropic’s official Claude Sonnet 5 System Card (§8.10.3, June 30 2026), evaluating Sonnet 5 on the Vision2Code task. Benchmark: benchcad.com. BenchCAD Vision2Code subset scores in the Claude Sonnet 5 System Card From Anthropic’s Claude Sonnet 5 System Card (§8.10.3).
Jun 9, 2026 Our benchmark BenchCAD is featured as a multimodal benchmark in Anthropic’s official Claude Fable 5 & Claude Mythos 5 System Card: a dedicated section (§8.16.4, pp. 282–283) evaluates their frontier models on BenchCAD’s Vision2Code task — with two figures and a Python-tools ablation. BenchCAD Vision2Code scores in Anthropic's system card From Anthropic’s Claude Fable 5 & Claude Mythos 5 System Card (§8.16.4).
May 11, 2026 Our new paper “BenchCAD: A Comprehensive, Industry-Standard Benchmark for Programmatic CAD” is now on arXiv. arXiv:2605.10865.
Mar 24, 2026 Our paper “CADLoop: An Equivariant-Aware Skill-Grounded Loop for CAD Data Curation” has been accepted to the CVPR 2026 NeXD Workshop (Exploring the Next Generation of Data).
Nov 30, 2025 Voice agent for stocks — prototype is up.
Nov 10, 2025 Started building a personal voice agent for stock trading and news.
Aug 15, 2025 Our paper “A buckling mechanics model for pattern transformation of lattice superstructures assembled by soft-hard materials integrated units” has been published in Mechanics of Materials (Vol. 202, art. 105253).

featured

  1. arXiv
    BenchCAD: A Comprehensive, Industry-Standard Benchmark for Programmatic CAD
    Haozhe Zhang†, Kaichen Liu, Miaomiao Chen, Lei Li, Shaojie Yang, Cheng Peng, and Hanjie Chen†
    arXiv preprint arXiv:2605.10865 2026
    † Corresponding authors
    ★ Reported by frontier labs — Anthropic’s Claude Fable 5 / Mythos 5, Sonnet 5 and Opus 5 System Cards, and OpenAI’s GPT-5.6 launch