Jobs and Careers
MI

Senior Software Engineer

Microsoft
Washington, United Statesfull_timeVerifiedPosted 5 Aug 2026
💰 $234,700/yr($119,800/yr$234,700/yr)

About the role

Overview

Microsoft is a company where passionate innovators come to collaborate, envision what can be and take their careers further. This is a world of more possibilities, more innovation, more openness, and the sky is the limit thinking in a cloud-enabled world.

Microsoft’s Azure Data engineering team is leading the transformation of analytics in the world of data with products like databases, data integration, big data analytics, messaging & real-time analytics, and business intelligence. The products our portfolio include Microsoft Fabric, Azure SQL DB, Azure Cosmos DB, Azure PostgreSQL, Azure Data Factory, Azure Synapse Analytics, Azure Service Bus, Azure Event Grid, and Power BI. Our mission is to build the data platform for the age of AI, powering a new class of data-first applications and driving a data culture.

Within Microsoft Fabric, the Azure Monitor team builds services that enable customers to monitor, detect, troubleshoot, and mitigate issues with their services through an increasingly agentic experience. Azure Monitor includes Log Analytics, Application Insights, Container Insights, Hosted Prometheus, Azure Managed Grafana, and more. Additionally, Azure Monitor is the platform upon which Microsoft Sentinel is built. We have a multi-billion dollar business that is growing rapidly, and we run some of the world’s highest scale observability services both for Microsoft internally and for our external customers, processing over 1.5 Exabytes of logs daily and tracking over 100 billion active metrics.​

The Azure Monitor Site Reliability Team is hiring a Senior Software Engineer to drive architecture and delivery of agentic reliability platform capabilities for Microsoft Monitoring solutions. This role defines and builds platform systems for monitoring intelligence, telemetry, diagnostics, safe remediation, and operational automation used by external customers and internal Microsoft engineering teams. 

​​We do not just value differences or different perspectives. We seek them out and invite them in so we can tap into the collective power of everyone in the company. As a result, our customers are better served.



Responsibilities
  • Define architecture for agentic reliability systems spanning monitoring, telemetry, incident management, service topology, deployment signals, and operational knowledge.
  • Lead platform capabilities for automated detection, triage, root-cause assistance, mitigation recommendations, safe execution, and post-incident learning.
  • Establish engineering standards for safe agentic operations, including identity, access, compliance, rollback, auditability, change management, and human escalation.
  • Influence service teams to adopt consistent monitoring, SLOs, alert quality, incident automation, live-site readiness, and operational excellence practices.
  • Identify high-impact reliability gaps and convert them into platform investments, architectural improvements, and reusable automation.
  • Lead complex live-site investigations and drive systemic reliability improvements from incident patterns and customer-impact data.
  • Mentor engineers and shape long-term technical direction across monitoring, observability, incident response, and AI-assisted and agentic automation. 
  • Embody our culture and values 


Qualifications
Required Qualifications
Bachelor's Degree in Computer Science or related technical field AND 4+ years technical engineering experience with coding in languages including, but not limited to, C, C++, C#, Java, JavaScript, or Python OR equivalent experience.
 

Other Requirements: 

Ability to meet Microsoft, customer and/or government security screening requirements are required for this role. These requirements include, but are not limited to the following specialized security screenings: Microsoft Cloud Background Check:

  • This position will be required to pass the Microsoft Cloud background check upon hire/transfer and every two years thereafter.

 

Preferred Qualifications

  • Experience building production-scale platforms, cloud services, distributed systems, or reliability automation.
  • Experience architecting complex systems across service boundaries and driving execution across partner teams without direct authority.
  • Experience with observability architecture, monitoring systems, in

Apply for this role

Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.

Apply Now →Generate Application Kit

Free account required — sign up in 30s

Company

Microsoft

View company profile →