Principal Engineering Manager
MicrosoftAbout the role
Azure Edge + Platform (AE+P) brings together Edge platforms, devices, and services to deliver Edge solutions, operating systems, and engineering systems. Driven by its customers’ needs, Azure Edge + Platform seeks to accelerate growth for Azure, Edge & Devices (E&D), and Microsoft’s customers worldwide.
The organization’s portfolio spans the Cloud Edge Stack, Azure Engineering Systems, Azure Media Services - for end-to-end media workflow and analytics - and Microsoft’s Operating Systems including the Azure Host OS (Operating System) and Windows. This portfolio impressively powers the world with more than one billion monthly active devices.
As part of AE+P, the Microsoft Observability Platform and Experience engineering team is responsible for both the external and internal observability product offerings. These services enable customers to monitor, detect, troubleshoot, and mitigate issues with their services. Azure Monitor is our product family and includes Log Analytics, Application Insights, Container Insights, Managed Prometheus, VM insights and more. Azure Monitor is a greater than $1B dollar business that is growing rapidly. We also run some of the world’s highest scale observability services for Microsoft internally, processing over 600PB of logs daily and tracking over 100 billion active metric time series. Our mission is to provide end-to-end observability for cloud and edge, enabling all customers to efficiently operate and provide a high-quality experience with their apps, services, and infrastructure. We are looking for an Engineering Manager to lead our cloud-native observability efforts, with a particular focus on Kubernetes observability. Most new applications, including cutting-edge AI and machine learning workloads, are now developed and run on Kubernetes. Kubernetes and AI workload observability present unique challenges and exciting opportunities. We are dealing with distributed systems, microservices, and ephemeral workloads. We are responsible for monitoring both the application layer (services, APIs, and business logic) and the underlying infrastructure (nodes, pods, networking, storage). Our customers include Microsoft, Fortune 500 companies, and startups. Our aim is to enable autonomous, integrated systems infused with artificial intelligence/machine learning (AI/ML) to proactively measure, monitor, detect, alert, diagnose, and mitigate health issues before they impact customers.
As a Principal Engineering Manager you will be overseeing a team of software engineers, ensuring project and development excellence, fostering career development and support, and cultivating a team culture characterized by customer passion, collaboration, diversity, and inclusion. This role necessitates the ability to collaborate across organizations and divisions and offers an opportunity to contribute to defining and executing the vision and technical strategy for enhancing our product. The individual in this role will be accountable for designing, implementing, and operating high-scale, highly optimized services capable of accommodating exponential growth in scale, volume, and usage over the years to come. As part of this role, you will have opportunities to participate and contribute to Cloud native technologies and influence the future of Open Telemetry, Prometheus & Kubernetes. We’re on a mission to build the observability services on the planet. If that mission excites you, come join us!
Microsoft’s mission is to empower every person and every organization on the planet to achieve more. As employees we come together with a growth mindset, innovate to empower others, and collaborate to realize our shared goals. Each day we build on our values of respect, integrity, and accountability to create a culture of inclusion where everyone can thrive at work and beyond.
In alignment with our Microsoft values, we are committed to cultivating an inclusive work environment for all employees to positively impact our culture every day.
Responsibilities
As the Principal Engineering Manager, you will be responsible for the following:
- Work with teams across Observability organization to design and build scalable, high-performance services that are highly reliable.
- You’ll build, foster, grow, and retain a team of high-performance engineers.
- Provide mentorship and coaching to engineers in, and beyond, your team.
- Own and deliver complete features across the development lifecycle, including design, architecture, implementation, testability, debugging, shipping, and servicing.
- Write and review clean, well-thought-out code with an emphasis on quality, performance, simplicity, durability, scalability, and maintainability.
- Be passionate about making customers successful.
- Stay informed about industry trends, emerging technologies, advancemen
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s