Jobs and Careers
SE

Software Engineer, ML Infrastructure

Sesame
San Francisco, United Statesfull_timeVerifiedPosted 24 Feb 2025
💰 $280,000/yr($175,000/yr$280,000/yr)

About the role

Software Engineer, ML Infrastructure

About Sesame

Sesame believes in a future where computers are lifelike, optimized for consumer-focused AI experiences, and with the ability to see, hear, and collaborate with us in ways that feel natural and human. With this vision, we're designing a new kind of computer. Our team brings together visionary founders from Oculus and Ubiquity6, alongside proven leaders across Meta and Google with deep expertise spanning hardware and software. We are building a transformative platform to unlock the next era of human-computer interaction. Join us in shaping a future where computers truly come alive.

About the role

As a Software Engineer, ML Infrastructure, you will build and scale the foundational infrastructure that powers Sesame’s AI-driven computing experiences. You will work closely with ML researchers, software engineers, and hardware teams to design robust, high-performance infrastructure that enables cutting-edge AI models to run efficiently in real-time environments.

Responsibilities:

  • Design, build, and optimize scalable ML infrastructure to support training, evaluation, and deployment of AI models.

  • Develop and maintain data pipelines for large-scale machine learning workflows.

  • Implement efficient model-serving architectures for real-time inference.

  • Collaborate with ML engineers and researchers to improve model performance and deployment efficiency.

  • Build monitoring and observability tools to ensure system reliability and performance.

Required qualifications:

  • Bachelor's or Master's degree in Computer Science, Engineering, or a related field.

  • 5+ years of experience in software engineering, focusing on ML infrastructure, distributed systems, or backend engineering.

  • Proficiency in Python, C++, or another systems programming language.

  • Experience with ML frameworks such as TensorFlow, PyTorch, or JAX.

  • Hands-on experience with cloud platforms (AWS, GCP, or Azure) and containerization (Docker, Kubernetes).

  • Knowledge of hardware acceleration (TPUs, GPUs) and efficient model deployment strategies.

  • Strong understanding of distributed computing, data pipelines, and model-serving architectures.

Preferred qualifications:

  • Experience optimizing ML model inference for real-time applications and edge devices

  • Experience writing and optimizing ML kernels in CUDA or Triton

  • Familiarity with large-scale training pipelines and model orchestration tools.

  • Experience with profiling tools such as Nvidia Nsight or Pytorch Profiler

  • Prior work in AI-driven consumer applications or human-computer interaction.

Sesame is committed to a workplace where everyone feels valued, respected, and empowered. We welcome all qualified applicants, embracing diversity in race, gender, identity, orientation, ability, and more. We provide reasonable accommodations for applicants with disabilities—contact careers@sesame.com for assistance.

Apply for this role

Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.

Apply Now →Generate Application Kit

Free account required — sign up in 30s

Company

Sesame

View company profile →