Jobs and Careers
CO
Senior Site Reliability Engineer (Hybrid)
Colliers International EMEASpainfull_timeVerifiedPosted 21 Nov 2025
About the role
<h3>Company Description</h3><p>Colliers is a leading diversified professional services and investment management company. With operations in 68 countries, our 22,000 enterprising people work collaboratively to provide expert advice to maximize the potential of property and real assets to accelerate the success of our clients, our investors and our people.</p><p>We are at the forefront of the real estate industry, leading the way and backed by an exceptional record of success. We are building for our future – <em>and yours.</em></p><p>We strive to build our business at a competitive pace by augmenting internal growth with smart strategic acquisitions that increase market share, expand service offerings and extend our geographic reach for the benefit of our clients and shareholders.</p><p>For more than 29 years, Colliers has created value for shareholders that has resulted in superior returns and industry growth. Our people also own significant equity in our business, which brings pride of ownership to everything we do. We are passionate, take personal responsibility and always do what's right for our clients, people and communities.</p><h3>Job Description</h3><p>We are looking for an experienced and passionate <strong>Senior Side Reliability Engineer</strong> for our <strong>newly established Technology Hub</strong> in Madrid. This is a unique opportunity to become part of the founding team that helps shape the culture, practices, and technical direction. Working in a <strong>hybrid model</strong> (2 days on-site), you will enjoy significant freedom to <strong>innovate and influence</strong> the future of one of the world’s largest commercial real estate companies. With the leadership of the Global Technology Hub coming from a background of technology startups, you will benefit from a fast-paced learning environment, high visibility of your contributions and opportunities to shape processes, culture and technology strategies. At the same time, you take advantage of the stability, resources, and reach of a <strong>successful global company</strong>.</p><p>Guided by our global digital strategy, the Madrid Hub collaborates closely with international teams to deliver <strong>world-class technology solutions</strong> that power the future of commercial real estate.</p><p>As <strong>Senior Side Reliability Engineer</strong>, you are focused on ensuring the reliability, performance and availability of our applications and platforms across GCP and Azure, while enabling development teams to ship faster with confidence.</p><p>As a senior member of the DevOps team, you will help design and implement observability systems, reliability practices, and incident response processes in collaboration with Software Engineering and Infrastructure teams. Your mission is to bring an engineering-first mindset to operations, applying automation, data, and feedback loop continuously improve the resilience of our systems and platforms. You will contribute to global products and platforms that serve both internal and external customers in Commercial Real Estate across multiple regions. Working closely with international Product, Engineering, DevOps, Data, QA and Architecture teams, you will ensure delivery excellence, engineering quality and great consumer experience.</p><p>You are a hands-on problem-solver with strong design principles, who thrives in a collaborative and agile environment. As a senior member of the DevOps function, you will set technical direction, drive best practices. and mentor junior engineers while building solutions with real business impact.</p><p> </p><p><strong>Reliability Engineering, and Operational Excellence</strong></p><ul><li>Define and maintain Service Level Indicators, Service Level Objectives and Service Level Agreements across critical services in partnership with Product Owners; Engineering and Infrastructure Teams.</li><li>Identify resilience gaps and lead initiatives such as redundancy improvements and scaling strategies to address these.</li><li>Automate incident response, recovery and scaling where possible.</li><li>Build tooling for self-healing infrastructure and applications, reducing manual intervention.</li><li>Contribute to runbooks, playbooks and knowledge sharing for operations best practices.</li><li>Build mechanisms to ensure error budgets are respected and used to drive prioritization decisions.</li></ul><p><strong>Observability Monitoring and Incident Management </strong></p><ul><li>Design, implement and evolve monitoring, logging, and tracing systems (Azure Monitor, GCP Operations Suite, Prometheus, Grafana, Datadog).</li><li>Develop dashboards and alerting systems that provide actionable insights for engineers and stakeholders.</li><li>Design and implement a comprehensive ChatOps strategy and ensure close integration with Teams.</li><li>Collaborate with QA teams and Engineering teams to integrate performance and availability testing into CI/CD pipelines.</li><li>Lead
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s