Pony.AI Logo

Pony.AI

(Senior) Software Engineer, Deep Learning

Reposted One Month Ago
In-Office
Fremont, CA, USA
140K-280K Annually
Mid level
In-Office
Fremont, CA, USA
140K-280K Annually
Mid level
Design and develop software architecture for self-driving vehicles, focusing on deep learning models and real-time ML solutions, while optimizing for performance and resource constraints.
The summary above was generated by AI

Founded in 2016 in Silicon Valley, Pony.ai has quickly become a global leader in autonomous mobility and is a pioneer in extending autonomous mobility technologies and services at a rapidly expanding footprint of sites around the world. Operating Robotaxi, Robotruck and Personally Owned Vehicles (POV) business units, Pony.ai is an industry leader in the commercialization of autonomous driving and is committed to developing the safest autonomous driving capabilities on a global scale. Pony.ai’s leading position has been recognized, with CNBC ranking Pony.ai #10 on its CNBC Disruptor list of the 50 most innovative and disruptive tech companies of 2022. In June 2023, Pony.ai was recognized on the XPRIZE and Bessemer Venture Partners inaugural “XB100” 2023 list of the world’s top 100 private deep tech companies, ranking #12 globally. As of August 2023, Pony.ai has accumulated nearly 21 million miles of autonomous driving globally. Pony.ai went public at NASDAQ in November 2024.

Responsibility
  • Work with experts in the field of self-driving vehicles on software architecture and design, system and module design, evaluation metrics, specification and implementation of test and regression frameworks.
  • Design and develop large-scale foundation models trained on vast of real world data
  • Frame the open-ended real-world problems into well-defined ML problems; develop and apply cutting-edge ML approaches (deep learning, reinforcement learning, imitation learning, etc) to these problems; scale them to data pipelines; and streamline them to run in real-time on the cars.
  • Develop and deploy deep learning models, including vision language models (VLMs) and Large Language Models (LLMs)
  • Optimize deep learning models to run robustly under tight run-time constraints.

Requirements
  • Master in Computer Science, or at least 2 years of equivalent industry experience in similar technical fields.
  • Solid understanding of data structures, algorithms, parallel computing, code optimization and large scale data processing.
  • Experience in applied machine learning including data collection and analysis, evaluation and feature engineering.
  • Expertise in C++/Python.
  • Strong communication skills and team spirit.
Preferred Experience
  • PhD in Deep Learning, Machine Learning, Robotics, Natural Language Processing, or similar technical field of study.
  • Publications on top-tier conferences like CVPR/ICCV/ECCV/ICLR/ICML/NeurIPS/ICLR/AAAI/IJCV/PAMI
  • Experience in applying ML/DL for behavior prediction, imitation learning, motion planning.
  • Experience in deploying deep learning algorithms for real time applications, with limited computing resources.
  • Experience in convex optimization, computational geometry or linear algebra.
  • Experience in GPU/CUDA/TensorRT
Compensation and Benefits

Base Salary Range: $140,000 - $280,000 Annually

Compensation may vary outside of this range depending on many factors, including the candidate’s qualifications, skills, competencies, experience, and location. Base pay is one part of the Total Compensation and this role may be eligible for bonuses/incentives and restricted stock units.

Also, we provide the following benefits to the eligible employees:

  • Health Care Plan (Medical, Dental & Vision)
  • Retirement Plan (Traditional and Roth 401k)
  • Life Insurance (Basic, Voluntary & AD&D)
  • Paid Time Off (Vacation & Public Holidays)
  • Family Leave (Maternity, Paternity)
  • Short Term & Long Term Disability
  • Free Food & Snacks
HQ

Pony.AI Fremont, California, USA Office

3501 Gateway Blvd, Fremont, CA, United States, 94538

Similar Jobs

12 Days Ago
In-Office or Remote
Santa Clara, CA, USA
184K-357K Annually
Senior level
184K-357K Annually
Senior level
Artificial Intelligence • Computer Vision • Hardware • Robotics • Metaverse
Develop and optimize GPU-accelerated deep learning inference software for LLMs, multimodal, and generative AI models. Contribute features to vLLM, SGLang, FlashInfer, and NVIDIA inference libraries; implement algorithms, tune performance across GPU architectures, and improve large-scale model serving. The role requires strong C/C++ software engineering, collaboration across framework teams, and expertise in profiling, debugging, performance modeling, GPU programming, or multi-GPU communication.
Top Skills: CC++CpuCudaCutlassDeep Learning FrameworksFlashinferGenerative AiGpuLarge Language ModelsMultimodal AiNcclNvidia AcceleratorsNvshmemOai TritonPythonPyTorchSglangVllm
29 Days Ago
In-Office
Santa Clara, CA, USA
152K-288K Annually
Senior level
152K-288K Annually
Senior level
Artificial Intelligence • Computer Vision • Hardware • Robotics • Metaverse
Develop simulation infrastructure for evaluating deep learning workloads across NVIDIA GPUs and compiler stacks. Optimize compiler kernel code generation, computational graphs, and datacenter-scale AI deployments using performance modeling. Collaborate with architecture, software, product, and research teams to assess future GPU features and influence silicon and system design. The role requires strong MLIR, C/C++, and Python expertise, along with experience in compiler optimization, architectural simulation, or related areas.
Top Skills: Architectural SimulationCC++Compiler InfrastructureDeep Learning CompilersGpusMlirNvidia Compiler StacksPython
One Month Ago
In-Office or Remote
2 Locations
184K-357K Annually
Senior level
184K-357K Annually
Senior level
Artificial Intelligence • Computer Vision • Hardware • Robotics • Metaverse
Design and optimize CUDA kernels and cluster-scale distributed systems for deep learning. Prototype system and compiler optimizations, profile hardware-software interactions, collaborate with researchers and architects, and deliver high-performance runtime tools and maintainable code for training and inference at scale.
Top Skills: C++CudaCuda DriverFp8Int8JaxMegatronMpiMxfp4NcclNemoNvfp4PythonPyTorchSglangTensorrtTorch.CompileTritonUcxVllmXla

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account