Site Reliability Engineer Jobs in San Francisco, USA

73 verified site reliability engineer openings in San Francisco

  • Site Reliability Engineer (Network)

    Loftorbital · San Francisco, CA

    On-site
    about 18 hours agoApply →

    Wanna join the adventure?   As a Senior Ground Segment Engineer, you play a pivotal role in building software to operate our fleet of in-orbit satellites, and to prepare for the challenges associated with operating different satellite buses in a variety of configurations. Your expertise will be instrumental in ensuring reliable, efficient, and scalable connectivity between our growing fleet of satellites via both ground stations and inter-satellite links. You will configure Cockpit, Loft’s mission control system, to expand our ground segment connectivity, working with our ground station providers and the product teams that build Cockpit. You will manage cross-fleet contact allocation and planning, optimizing resource usage to maximize communication opportunities and minimize latency. Work closely with our Space Infrastructure, Spacecraft Deployment, Mission Deployment and Execution, and regulatory teams to manage commissioned satellites and prepare for the commissioning of new ones. This position will also give you the opportunity to become a rotating Flight Director, responsible for the health and safety of our satellite fleet. There is opportunity for growth in this position. Your expertise in network infrastructure will be crucial in ensuring the reliability of our global satellite network, enabling us to deliver exceptional service to our customers. This role offers a broad scope to work on different aspects of satellite connectivity and play a significant part in shaping the future of satellite operations at Loft.

  • Senior Software Engineer, Site Reliability Engineer

    harvey · San Francisco, USA

    On-site
    2 days agoApply →
  • Staff Software Engineer, Site Reliability Engineer

    harvey · San Francisco, USA

    On-site
    2 days agoApply →
  • IT Spec (ENTARCH) "Platform/site Reliability Engineer", GS-2210-14 FPL GS-14 (DH)

    Federal Student Aid · San Francisco, California, United States

    On-siteUSD127,829–197,200/yr
    9 days agoApply →

    Minimum Qualification Requirements Specialized Experience for the IT Specialist (INFOSEC), GS-2210-14 One year of experience in either federal or non-federal service that is equivalent to at least a GS-13 performing two (2) out of three (3) of the following duties or work assignments: 1. Direct technical experience designing, deploying, and operating scalable cloud platforms using Infrastructure as Code (IaC), CI/CD, containers, and automated security controls to accelerate engineering delivery and ensure compliance. 2. Direct technical experience enhancing reliability and observability for distributed systems, including use of tracing/metrics/logs for observability, SLO/SLA development, incident response and analysis, and/or measurable performance/automation improvements. 3. Experience translating platform and reliability engineering concepts into clear documentation, technical standards, and architecture guidance for non-technical audiences, and influencing engineering practices across multiple teams. Basic Experience Requirements You must possess IT related experience (paid or unpaid experience and/or completion of specific, intensive training (e.g., IT certification), as appropriate) demonstrating each of the nine competencies listed below. 1. Attention to Detail - Is thorough when performing work and conscientious about attending to detail. 2. Customer Service - Works with clients and customers (that is, any individuals who use or receive the services or products that your work unit produces, including the general public, individuals who work in the agency, other agencies, or organizations outside the Government) to assess their needs, provide information or assistance, resolve their problems, or satisfy their expectations; knows about available products and services; is committed to providing quality products and services. 3. Decision Making - Makes sound, well-informed, and objective decisions; perceives the impact and implications of decisions; commits to action, even in uncertain situations, to accomplish organizational goals; causes change. 4. Information Management - Identifies a need for and knows where or how to gather information; organizes and maintains information or information management systems. 5. Interpersonal Skills - Shows understanding, friendliness, courtesy, tact, empathy, concern, and politeness to others; develops and maintains effective relationships with others; may include effectively dealing with individuals who are difficult, hostile, or distressed; relates well to people from varied backgrounds and different situations 6. Oral Communication - Expresses information (for example, ideas or facts) to individuals or groups effectively, taking into account the audience and nature of the information (for example, technical, sensitive, controversial); makes clear and convincing oral presentations; listens to others, attends to nonverbal cues, and responds appropriately. 7. Problem Solving - Identifies problems; determines accuracy and relevance of information; uses sound judgment to generate and evaluate alternatives, and to make recommendations. 8. Teamwork - Encourages and facilitates cooperation, pride, trust, and group identity; fosters commitment and team spirit; works with others to achieve goals. 9. Technical Competence – Uses knowledge that is acquired through formal training or on-the-job experience to perform one's job; works with, understands, and evaluates technical information related to the job; advises others on technical issues. Knowledge, Skills, and Abilities (KSAs) The quality of your experience will be measured by the extent to which you possess the following knowledge, skills and abilities (KSAs). You do not need to provide separate narrative responses to these KSAs, as they will be measured by your responses to the occupational questionnaire (you may preview the occupational questionnaire by clicking the link at the end of the Evaluations section of this vacancy announcement). 1. Skill in designing and implementing cloud and hybrid network solutions, including reusable platform services, container environments, identity integrations, and infrastructure components. 2. Skill in applying systems engineering and Site Reliability Engineering (SRE) concepts to ensure reliability, performance, scalability, security, and maintainability across complex, multi-cloud environments. 3. Knowledge of platform and reliability engineering principles and the ability to apply them through real-world implementation, debugging, optimization, and modernization of cloud environments. 4. Skill in computer engineering cloud automation, observability tooling, testing frameworks, and Continuous improvement/Continuous development (CI/CD) pipelines, including telemetry, logging, alerting, and distributed tracing. 5. Ability to leverage modern cloud, data, and security technologies to design, test, and deploy resilient platform and reliability systems that support mission-critical applications.

  • Staff TDI Site Reliability Engineer, Okta Federal

    okta · San Francisco, California

    On-site
    9 days agoApply →
  • Site Reliability Engineer

    cognition · San Francisco, USA

    On-site
    15 days agoApply →
  • Senior Site Reliability Engineer

    coderabbit · San Francisco, USA

    On-site
    16 days agoApply →
  • Site Reliability Engineer

    gamma · San Francisco, USA

    On-site
    16 days agoApply →
  • Site Reliability Engineer

    latent · San Francisco, USA

    On-site
    16 days agoApply →
  • Senior Site Reliability Engineer - SDN

    lambda · San Francisco Office (Fremont St), USA

    On-site
    16 days agoApply →
  • Senior Site Reliability Engineer

    drata · Hybrid - San Francisco, USA

    Hybrid
    16 days agoApply →
  • Lead Site Reliability Engineer

    stuut-ai · San Francisco, USA

    On-site
    16 days agoApply →
  • Director, Site Reliability Engineering

    stellar · San Francisco, USA

    On-site
    16 days agoApply →
  • Site Reliability Engineer, Compute

    fluidstack · San Francisco, CA, USA

    On-site
    16 days agoApply →
  • Senior Staff Site Reliability Engineer

    Ironcladhq · San Francisco, USA

    On-site
    17 days agoApply →

    Senior Staff Site Reliability Engineer at Ironcladhq. Apply via Ashby.

  • Senior Site Reliability Engineer - SDN

    Lambda · San Francisco Office (Fremont St), USA

    On-site
    17 days agoApply →

    Senior Site Reliability Engineer - SDN at Lambda. Apply via Ashby.

  • Site Reliability Engineer

    Gamma · San Francisco, USA

    On-site
    17 days agoApply →

    Site Reliability Engineer at Gamma. Apply via Ashby.

  • Senior Site Reliability Engineer, Platform Infrastructure (Foundations)

    anyscale · San Francisco, USA

    On-site
    17 days agoApply →
  • Site Reliability Engineer

    mercor · San Francisco, USA

    On-site
    17 days agoApply →
  • Senior Staff Site Reliability Engineer

    ironcladhq · San Francisco, USA

    On-site
    17 days agoApply →