Hi there Welcome to my Homepage!

I am an undergraduate (2022-2026) at Huazhong University of Science and Technology and an incoming PhD at AutoMan@NTU with Prof. Chen Lyu, passionate about Autonomous Driving and Computer Vision.

Previously I worked at AIR@THU with Prof. Hao Zhao and Autolab@WLU with Prof. Kaicheng Yu.

Currently I conduct the WM and E2E research at Neolix.

News

Experience

Nanyang Technological University
Aug 2026 –
Ph.D at AutoMan@NTU
Neolix
Feb 2026 – Present
Research Intern at Neolix-AD
Westlake University
Jun 2025 – Jan 2026
Research Assistant at AutoLab
Tsinghua University
Jun 2024 – Nov 2025
Research Assistant at AIR
Huazhong Univ of Sci and Tech
Sep 2022 – Jul 2026
Research Assistant at XWCV

Projects

YouDrive
YOUDrive: Driving Style as a Steerable Axis for Personalized End-to-End Driving
Guantian Zheng, Jiashu Li, Jiaxing Chen, Tianyu Gao, Lidong Yu
YOUDrive treats driving style as a continuous, composable axis on a VLA backbone: a flow-matching decoder commits to feasible trajectories instead of averaging, and each persona is a low-rank task vector scaled by one coefficient. The Style Alignment Score (SAS) reports style independently of safety; on NAVSIM, a single coefficient traces a controllable style path while keeping PDMS above 0.90 (up to 0.954).
CVPR 2027 submission   [arxiv] [code] [dataset]
ATLAS
ATLAS: Large-Scale Multimodal Autonomous-Driving Backbone Pre-training
Guantian Zheng, Lidong Yu
ATLAS is a ~7.5B parameter multimodal visual backbone for autonomous driving, pretrained under a multi-teacher distillation framework combining DINOv2/VGGT (geometry & semantics), Cosmos Tokenizer (visual-generation supervision), and a 7B VLM (semantic alignment). I owned the Video Head / Render Decoder, realizing visual-token distillation via the Cosmos CI Tokenizer and systematically analyzing token-space alignment.
StyleShield
StyleShield: Exposing the Fragility of AIGC Detectors through Continuous Controllable Style Transfer
Guantian Zheng
StyleShield, the first flow matching framework for conditional text style transfer in continuous token embedding space. A single parameter γ provides smooth, continuous control over the evasion--preservation trade-off, fundamentally inaccessible to discrete-token methods.
EACL 2026 submission   [arxiv] [code] [dataset]
OpenlaneV3
Openlane-V3
Guantian Zheng, Zongzheng Zhang, Jijun Wang, Hao Zhao
OpenLane-V3 is an extended version of the OpenLaneV2 benchmark, integrating additional modalities including 3D traffic light and traffic sign annotations with semantic and positional information.
PointHypE
Enhanced Point Cloud Reconstruction with PTv3 and Dual Hyper in SVDFormer
Guantian Zheng, Boran Zhang
Proposed a HyperChamfer Embedding with a dual-hypernetwork architecture to inject global geometric structure into refinement, and integrated PTv3 backbone for efficient acceleration.
Brain
Brain-Controlled Robotic Arm
Guantian Zheng, Jincheng Yang, Dawei Ye
Achieved real-time recognition and control of a single hand with five degrees of freedom, with future plans to enable assisting paralyzed individuals in daily tasks such as eating, gripping, and writing.

Honors & Awards

  • Outstanding Graduate of HUST (2026)
  • Academic Excellence Scholarship (2025)
  • Self-Motivation and Diligence Scholarship (2024)
  • Academic Excellence Scholarship (2023)