XPENG logo
XPENG
Posted 91 days agoVerified live 2d ago

AI Infra Onboard Performance Intern

Brief overview

Santa Clara, CAIn-person
MastersOr in progress

About the company

XPENG is a leading Chinese Smart EV company that designs, develops, manufactures, and markets Smart EVs that appeal to the large and growing base of technology-savvy middle-class consumers.

Job description

Summary

XPENG is a leading smart technology company at the forefront of innovation, integrating advanced AI and autonomous driving technologies into its vehicles. They are seeking an AI Infra Onboard Performance Intern to design and optimize software for deploying large-scale AI models in production vehicles and improve inference performance across various systems.

Responsibilities

  • Design and optimize software for deploying large-scale AI models in production vehicles
  • Profile and improve inference performance across compute, memory, and I/O systems
  • Reduce latency and improve power efficiency on embedded automotive platforms
  • Deliver production-ready optimizations that scale reliably across the vehicle fleet

Skills

  • Master's or PhD in CS/CE/EE or equivalent, with relevant industry or research experience
  • Strong expertise in C++ and Python, including performance-sensitive, production-quality code
  • In-depth understanding of computer architecture and high-performance computing (memory hierarchy, parallelism, vectorization, scheduling)
  • Proven experience in application performance analysis and optimization, using profiling tools to diagnose and resolve bottlenecks
  • Experience writing and optimizing CPU or CUDA kernels
  • Experience developing performance tooling and instrumentation (e.g., eBPF, perf, custom tracing/profiling frameworks) for production or embedded systems
  • Familiarity with embedded or automotive compute platforms (e.g., NVIDIA Orin, Drive) and their power/thermal constraints
  • Previous experience in the autonomous driving or robotics industry
  • Effective at solving complex problems collaboratively within larger cross-functional teams

Qualifications

Must Haves

  • Master's or PhD in CS/CE/EE or equivalent, with relevant industry or research experience
  • Strong expertise in C++ and Python, including performance-sensitive, production-quality code
  • In-depth understanding of computer architecture and high-performance computing (memory hierarchy, parallelism, vectorization, scheduling)
  • Proven experience in application performance analysis and optimization, using profiling tools to diagnose and resolve bottlenecks

Nice to Haves

  • Experience writing and optimizing CPU or CUDA kernels
  • Experience developing performance tooling and instrumentation (e.g., eBPF, perf, custom tracing/profiling frameworks) for production or embedded systems
  • Familiarity with embedded or automotive compute platforms (e.g., NVIDIA Orin, Drive) and their power/thermal constraints
  • Previous experience in the autonomous driving or robotics industry
  • Effective at solving complex problems collaboratively within larger cross-functional teams

Benefits

  • A fun, supportive and engaging environment.
  • Infrastructures and computational resources to support your work.
  • Opportunity to work on cutting edge technologies with the top talents in the field.
  • Opportunity to make significant impact on the transportation revolution by the means of advancing autonomous driving.
  • Competitive compensation package.
  • Snacks, lunches, dinners, and fun activities.

More jobs like this