Lead cross-functional programs to build and operate an autonomous-driving training data lake: define roadmap, coordinate data/ML/infrastructure/annotation teams, set KPIs for data quality and pipeline reliability, manage dataset generation/versioning/lineage, identify and resolve technical bottlenecks, and improve engineering and operational processes to deliver high-quality training data at scale.
XPENG is a leading smart technology company at the forefront of innovation, integrating advanced AI and autonomous driving technologies into its vehicles, including electric vehicles (EVs), electric vertical take-off and landing (eVTOL) aircraft, and robotics. With a strong focus on intelligent mobility, XPENG is dedicated to reshaping the future of transportation through cutting-edge R&D in AI, machine learning, and smart connectivity.
As a Staff Technical Program Manager focused on the Autonomous Driving Training Data Lake, you will lead cross-functional programs that enable large-scale data collection, processing, governance, and delivery for autonomous-driving model development. You will work closely with data infrastructure, machine learning, simulation, vehicle engineering, annotation, and algorithm teams to ensure that high-quality training data is available reliably, efficiently, and at scale.
You will work with best-in-class machine learning, software, data, and infrastructure engineers to build the foundation for XPENG’s autonomous-driving development lifecycle. Your work will directly improve the quality, velocity, reliability, and cost efficiency of the data used to train and validate mass-market autonomous-driving systems.
Job Responsibilities:
- Define and manage the roadmap for the autonomous-driving training data lake, covering data collection, ingestion, processing, storage, curation, and delivery.
- Coordinate execution across data infrastructure, machine learning, annotation, simulation, and autonomous-driving algorithm teams.
- Establish and track KPIs for data quality, pipeline reliability, processing latency, dataset readiness, storage efficiency, and infrastructure cost.
- Drive programs for dataset generation, versioning, metadata management, lineage, and reproducibility.
- Identify technical risks and bottlenecks in data pipelines, infrastructure capacity, and training-data delivery, and lead cross-functional resolution.
- Improve engineering planning, agile development practices, dependency management, and operational processes across data platform teams.
Basic Qualifications:
- Bachelors or Masters in computer science or other related engineering field.
- 5 years + of technical program management experience in engineering or related fields.
- Experience leading complex cross-functional data, infrastructure, software, or machine learning programs.
- Strong understanding of large-scale data systems, cloud infrastructure, and data pipelines.
- Mandarin speaking is required.
- Excellent problem-solving skills.
- Strong understanding of large-scale data systems, cloud infrastructure, and data pipelines.
- Experience in establishing strong software development processes.
Preferred Qualifications:
- Experience with autonomous-driving data platforms, machine learning infrastructure, or large-scale training-data systems.
- Experience managing large-scale data lakes, distributed data pipelines, or cloud infrastructure programs.
- Familiarity with data technologies such as object storage, Spark, Ray, Kafka, Iceberg, Parquet, Kubernetes, or similar systems.
- Understanding of machine learning data workflows, including data curation, annotation, training, evaluation, and closed-loop improvement.
- Experience managing complex cross-functional programs across multiple teams, locations, or time zones.
What do we provide:
- A fun, supportive and engaging environment.
- Opportunity to make a significant impact on the transportation revolution by the means of advancing autonomous driving.
- Opportunity to work on cutting edge technologies with the top talent in the field.
- Competitive compensation package.
- Snacks, lunches, dinners and fun activities.
The base salary range for this full-time position is $125,580 - $212,520, in addition to bonus, equity and benefits. Our salary ranges are determined by role, level, and location. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position across all US locations. Within the range, individual pay is determined by work location and additional factors, including job-related skills, experience, and relevant education or training.
We are an Equal Opportunity Employer. It is our policy to provide equal employment opportunities to all qualified persons without regard to race, age, color, sex, sexual orientation, religion, national origin, disability, veteran status or marital status or any other prescribed category set forth in federal or state regulations.
XPeng Motors Palo Alto, California, USA Office
Palo Alto, CA, United States, 94301
XPeng Motors San Jose, California, USA Office
San Jose, United States
Similar Jobs
Artificial Intelligence • Cloud • Machine Learning • Mobile • Software • Virtual Reality • App development
Leads large-scale, ambiguous ML platform and infrastructure programs across engineering organizations. Owns program lifecycles, dependencies, roadmaps, technical decisions, reliability, performance, cost optimization, and operational excellence. Uses Python, SQL, dashboards, and automation to analyze systems and guide decisions. Partners with senior engineering and product leaders, drives accountability, mentors junior TPMs and engineers, and improves platform-wide systems and engineering velocity.
Top Skills:
AirflowAWSDashboardsData VisualizationDockerFlinkGCPGitGrafanaHiveJIRAKafkaKubernetesLookerMl PlatformsNotebooksPythonSparkSQLTableau
Artificial Intelligence • Machine Learning • Robotics • Software • Transportation • Design • Manufacturing
Own the Primary Compute Unit program end to end, coordinating hardware development, software integration, validation, manufacturing, suppliers, quality, safety, and fleet issue resolution. Build integrated schedules, roadmaps, budgets, milestones, risk mitigation plans, and communication processes. Lead cross-functional teams, manage contract manufacturers and supplier activities, track work in JIRA, ensure assembly readiness and on-time parts delivery, and communicate program status to senior stakeholders while influencing without formal authority.
Top Skills:
Autonomous Vehicle HardwareCompute SystemsHardware/Software IntegrationJIRASmartsheet
Cloud • Information Technology • Security • Software • Cybersecurity
Lead end-to-end technical programs across multiple engineering orgs: design a centralized intake, run execution and release reviews, coordinate cross-functional release readiness, replace manual tracking with JIRA-based SDLC structures, and drive delivery predictability and process maturity.
Top Skills:
Ai/MlJIRAJira StructureNetwork ConfigurationNetworking ProtocolsRoutingSdlc
What you need to know about the San Francisco Tech Scene
San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.
Key Facts About San Francisco Tech
- Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
- Major Tech Employers: Google, Apple, Salesforce, Meta
- Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
- Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
- Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
- Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine



