Senior Accelerator Engineer II, Compute Kernels [Remote Friendly]
Cruise LLCAbout the role
We're Cruise, a self-driving service designed for the cities we love.
We’re building the world’s most advanced self-driving vehicles to safely connect people to the places, things, and experiences they care about. We believe self-driving vehicles will help save lives, reshape cities, give back time in transit, and restore freedom of movement for many.
In our cars, you’re free to be yourself. It’s the same here at Cruise. We’re creating a culture that values the experiences and contributions of all of the unique individuals who collectively make up Cruise, so that every employee can do their best work.
Cruise is committed to building a diverse, equitable, and inclusive environment, both in our workplace and in our products. If you are looking to play a part in making a positive impact in the world by advancing the revolutionary work of self-driving cars, come join us. Even if you might not meet every requirement, we strongly encourage you to apply. You might just be the right candidate for us.
The Autonomous Vehicle (AV) software stack is a sophisticated workload that includes compute and data intensive functions within perception, prediction, planning and control. This role is within the ML Accelerators (MLA) team which is responsible for maintaining and optimizing the software stack for real-time, on-road performance. This means efficient workload mapping onto a sophisticated, resource-constrained hardware, co-architecting the HW platform with HW engineers and building platforms and tools for AI deployment and performance profiling, diagnosis and optimizations that can be employed by AV engineers (from perception, prediction, planning and controls). MLA engineers work in close collaboration with AV engineers throughout the lifecycle of crafting new AV features which is essential to the success of our mission.
As a Senior Software Engineer in Cruise’s Machine Learning Accelerator’s (MLA) Kernels and Libraries team, you will be responsible for the design and development of high-performance, safety-critical reliable accelerator solutions for Cruise’s AV workloads across our compute platforms. Your solutions will be used to implement the core AV functions across ML and non-ML workloads spanning multiple GPU architectures. As Cruise expands our vehicle offerings on our path to productization, your libraries will not only enable Cruise’s low cost custom silicon critical to our long term business needs, but also provide the needed performance portability to ensure our successful scaling.
What you’ll be doing:
- Design and develop accelerator libraries that are optimized across multiple architectures (custom accelerator inside AV and GPGPUs in the cloud).
- Build optimized ML inference primitives and custom application workloads that support multiple accelerators.
- Collaborate with compiler and runtime teams to build scalable interfaces for core primitive compute libraries, ML teams to guide the architecture of models and application teams to design the right algorithms that make efficient use of compute resources.
- Own the performance of accelerator kernels and libraries used throughout the AV stack by benchmarking, profiling and optimizing workloads.
- Follow and improve coding standards, methodologies, processes, and guidelines.
What you must have:
- Excellent GPU programming skills in CUDA or OpenCL with a thorough understanding of parallel programming patterns and GPU architecture preferably across multiple vendors.
- Strong background in software architecture, library design and design patterns.
- Strong C++ programming skills with the ability to feel comfortable in large codebases.
- Hands-on experience benchmarking, profiling, debugging and optimizing accelerator libraries and kernels to extract optimal performance.
- Solid background in system performance, high performance computing and architecture-aware optimizations.
- Strong communication skills and ability to problem solve in the presence of ambiguity.
Bonus Points!
- MS or PhD in CS, or related technical field or equivalent experience
- Experience with GPU programming models like SYCL, Triton
- Experience with ML frameworks (e.g. Pytorch, JAX, Tensorflow)
- Experience with ARM NEON, SVE, AVS or SSE-style SIMD intrinsics programming
- Experience with DSP programming
The salary range for this position is $163,200 - 240,000. Compensation will vary depending on location, job-related knowledge, skills, and experience. You may also be offered a bonus, restricted stock units, and benefits. These ranges are subject to change.
Why Cruise?
-
Our benefits are here to support the whole you:
- Competitive sa
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s