Jobs and Careers
OR

Senior Software Engineer – Cloud Infrastructure Reliability & Automation

Oracle
United States, United Statesfull_timeVerifiedPosted 26 Feb 2026
💰 $158,200/yr($79,100/yr$158,200/yr)

About the role

Join Oracle's Health Data Intelligence (HDI) team as a Software Engineer 3 and contribute to the reliability, scalability, and performance of our world-class analytics platform. In this role, you will develop, maintain, and optimize the infrastructure and data pipelines that power healthcare analytics globally. You will work within a collaborative team to implement robust solutions for business intelligence and reporting, ensuring our platform handles massive datasets with precision and speed.

U.S. citizenship is required for this position, as the successful candidate will be required to obtain (and maintain) a U.S. government security clearance after hire.

Required Skills

  • Infrastructure & Reliability: Experience implementing and maintaining high-availability systems with a focus on performance monitoring and fault tolerance.
  • Data Technologies: Proficiency in Data Warehousing platforms (e.g., Vertica, Snowflake) and ETL frameworks; understanding of columnar storage and large-scale data processing.
  • BI & Reporting: Practical experience integrating or supporting Business Intelligence tools (e.g., Tableau, Power BI, Oracle Analytics) to surface data-driven insights.
  • DevOps/SRE Practices: Competency in CI/CD pipelines (Jenkins, Kubernetes), Infrastructure as Code (Terraform), and observability tools (Prometheus, Grafana).
  • Cloud Ecosystems: Working knowledge of public cloud environments (OCI, AWS, or Azure) with an emphasis on deployment and resource management.
  • Problem-Solving: Strong ability to troubleshoot complex production issues, perform root-cause analysis, and document technical findings.
  • Programming & Tools: Solid foundation in Python, Java, or Go, along with containerization (Docker) and shell scripting.
     

Work with Site Reliability Engineering (SRE) team on the shared full stack ownership of a collection of services and/or technology areas. Understand the end-to-end configuration, technical dependencies, and overall behavioral characteristics of production services. Responsible for the design and delivery of the mission critical stack, with focus on security, resiliency, scale, and performance. Authority for end-to-end performance and operability. Partner with development teams in defining and implementing improvements in service architecture. Articulate technical characteristics of services and technology areas and guide Development Teams to engineer and add premier capabilities to the Oracle Cloud service portfolio. Understand and communicate the scale, capacity, security, performance attributes, and requirements of the service and technology stack. Demonstrate clear understanding of automation and orchestration principles. Act as ultimate escalation point for complex or critical issues that have not yet been documented as Standard Operating Procedures (SOPs). Utilize a deep understanding of service topology and their dependencies required to troubleshoot issues and define mitigations. Understand and explain the affect of product architecture decisions on distributed systems. Professional curiosity and a desire to a develop deep understanding of services and technologies.

Key Responsibilities

  • Develop & Maintain: Implement and tune infrastructure components for the Oracle HDI Analytics Platform to ensure system stability and uptime.
  • Data Pipeline Execution: Build and refine scalable data pipelines, leveraging Vertica and ETL processes to ensure efficient data ingestion and transformation.
  • BI Support: Assist in the integration and optimization of BI and reporting tools to ensure seamless data visualization for healthcare leaders.
  • Operational Excellence: Apply DevOps and SRE principles to automate routine tasks, manage deployments via CI/CD, and monitor system health using Prometheus/Grafana.
  • Cloud Integration: Support platform-agnostic initiatives across Oracle Cloud and AWS, ensuring cost-efficient and compliant resource usage.
  • Incident Response: Participate in on-call rotations or troubleshooting sessions to resolve production issues and implement preventative fixes.

Apply for this role

Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.

Apply Now →Generate Application Kit

Free account required — sign up in 30s

Company

Oracle

View company profile →