Site Reliability Engineer (SRE) AI Infrastructure (Early Career)
Nebius · Amsterdam, Netherlands
On-siteabout 1 month agoApply →<div class="content-intro"><p><strong>About Nebius:</strong></p> <p>Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure.</p> <p>Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI.</p> <p>Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&amp;D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&amp;D.</p></div><p><strong>Summary:</strong></p> <p><strong>Location</strong>: Amsterdam<br><strong>Duration</strong>: 3 months<br><strong>Start date</strong>: June 2026&nbsp;<br><strong>Compensation</strong>: Paid<br><strong>Eligibility</strong>: Current University student (Computer Science or related field), Recent Graduate or Early Career specialist<br><strong>Work authorization</strong>: Permitted to work in the job’s location</p> <p><strong data-stringify-type="bold">Why work at Nebius<br></strong>Nebius is leading a new era in cloud computing to serve the global AI economy. We create the tools and resources our customers need to solve real-world challenges and transform industries, without massive infrastructure costs or the need to build large in-house AI/ML teams. Our employees work at the cutting edge of AI cloud infrastructure alongside some of the most experienced and innovative leaders and engineers in the field.</p> <p><strong>Where we work<br></strong>Headquartered in Amsterdam and listed on Nasdaq, Nebius has a global footprint with R&amp;D hubs across Europe, North America, and Israel. The team of over 1400+ employees includes more than 400 highly skilled engineers with deep expertise across hardware and software engineering, as well as an in-house AI R&amp;D team.</p> <p><strong>Your responsibilities will include</strong></p> <ul> <li>Operation: <ul> <li>assist, where possible in day-to-day SRE operations tasks in NetInfra</li> <li>deploy tested and approved changes following clear instructions</li> </ul> </li> <li>Project work: <ul> <li>execute small and well-defined tasks from the backlog</li> <li>together or instead: work on small and well-defined SRE project from backlog</li> <li>create tests for t
Site Reliability Engineer (SRE) - Security
IMC · Amsterdam, Netherlands
On-siteabout 1 month agoApply →<p data-local-id="827a50236684" data-prosemirror-content-type="node" data-prosemirror-node-name="paragraph" data-prosemirror-node-block="true" data-pm-slice="1 1 []">Platform Engineering at IMC Trading builds and runs the core platforms that power our business. This Site Reliability Engineer role within Platform Engineering is focused on security projects: embedding security into system design and infrastructure to protect and accelerate our rapidly evolving technology landscape.</p> <p data-local-id="2427b86258bb" data-prosemirror-content-type="node" data-prosemirror-node-name="paragraph" data-prosemirror-node-block="true">As a bridge between application development and production operations, Platform Engineering enables and applies security by design: embedding controls, guardrails, and secure defaults directly into our platforms and infrastructure. You will implement security into IMC systems in ways that reduce friction and raise our baseline of safety and reliability.</p> <p data-local-id="7d340fca6121" data-prosemirror-content-type="node" data-prosemirror-node-name="paragraph" data-prosemirror-node-block="true">We are hiring a Site Reliability Engineer within Platform Engineering to drive security-focused initiatives across IMC. You will collaborate closely with most technology teams, operate with high autonomy, and take ownership and responsibility for outcomes: delivering high-impact improvements to source control, CI/CD, observability, and the secure foundations our trading platforms rely on in a fast-changing, cutting-edge environment.</p> <p data-local-id="7d340fca6121" data-prosemirror-content-type="node" data-prosemirror-node-name="paragraph" data-prosemirror-node-block="true">&nbsp;</p> <p><strong>Your Core Responsibilities:&nbsp;</strong></p> <ul> <li data-local-id="90a22339-0a55-4ff3-bac6-6235fb8cc32e" data-prosemirror-content-type="node" data-prosemirror-node-name="listItem" data-prosemirror-node-block="true"> <p data-local-id="e9812b840d50" data-prosemirror-content-type="node" data-prosemirror-node-name="paragraph" data-prosemirror-node-block="true">Write and maintain code to extend, integrate, and improve the services we support - from custom plugins and API integrations to internal tooling and automation frameworks</p> </li> <li data-local-id="4a370f90ad89" data-prosemirror-content-type="node" data-prosemirror-node-name="listItem" data-prosemirror-node-block="true"> <p data-local-id="847adacaebb8" data-prosemirror-content-type="node" data-prosemirror-node-name="paragraph" data-prosemirror-node-bl
Senior Site Reliability Engineer
Manychat · Amsterdam, Netherlands
On-siteabout 2 months agoApply →<p><strong>WHO WE ARE 🌍</strong></p> <p>We help creators get more out of every conversation with Instagram-focused automations and support for other channels like Messenger, WhatsApp, and TikTok. The result? Better engagement, more sales, and real, sustainable growth.</p> <p>With a diverse team of 350+ people spread across three continents, we’re building the leading Chat Marketing platform that is used — and loved — by more than 1.5 million customers worldwide.</p> <p><br><span class="notion-enable-hover" data-token-index="0"><strong>WHO WE'RE LOOKING FOR </strong>🌟</span></p> <p>We’re looking for a Senior Site Reliability Engineer who thrives at the crossroads of classic Linux and AWS infrastructure and modern Site Reliability Engineering. This is a high-impact, hybrid role designed for someone who can manage cloud resources, harden Kubernetes clusters, and shape a more reliable and developer-friendly platform.</p> <p>We need you not just to maintain but to rethink and evolve our infrastructure, balancing hands-on operations with strategic improvements that future-proof our growing AI product landscape.<br><br>You’ll take over key responsibilities from our current Infra Lead who is transitioning to a software-focused role, giving you immediate ownership and space to shine.<br><br><strong>WHY THE ROLE IS SPECIAL</strong> <span class="notion-enable-hover" data-token-index="0">💡</span></p> <p>You won’t be a cog in a massive SRE org. You’ll be the bridge between Infrastructure and Engineering, shaping how we scale Kubernetes, how we approach platform reliability, and how developers ship fast without fear. You’ll get autonomy, ownership, and a smart, humble team excited to learn with you.<br><br><strong>WHAT YOU’LL DO 🤖</strong></p> <ul> <li>Maintain and harden AWS infrastructure (EC2, ALB/NLB, WAF, IAM, CloudWatch)</li> <li>Operate and evolve our EKS clusters powering Python-based AI services</li> <li>Migrate existing services to Kubernetes using Terraform and Helm</li> <li>Codify infrastructure with Terraform and manage host-level automation via Ansible</li> <li>Build and improve CI/CD pipelines with GitHub Actions</li> <li>Own observability efforts: Prometheus, Grafana, alerting, and on-call readiness</li> <li>Support OS-level patching, certs, WAF rules, and general infra hygiene</li> <li>Partner with engineers to guide best practices and drive platform reliability</li> <li>Create clean, maintainable infrastructure documentation and playbooks</li> <li>Occasionally support rare off-hours incidents (don’t worry, really rare)</li> </ul> <h4>TO SHINE IN THIS ROLE 💥</h4&
Site Reliability Engineer (SRE)
Airapps · Amsterdam, Netherlands
On-siteabout 2 months agoApply →Site Reliability Engineer (SRE) at Airapps. Apply via Ashby.
Site Reliability Engineer
optiverus · Amsterdam, Netherlands
On-siteabout 2 months agoApply →Senior Network Site Reliability Engineer (NetSRE)
nebius · Amsterdam, Netherlands; Remote - Europe
Remoteabout 2 months agoApply →Senior Site Reliability Engineer (SRE, Compute Node Team)
nebius · Amsterdam, Netherlands; Remote - Europe
Remoteabout 2 months agoApply →Senior Site Reliability Engineer — Token Factory (Inference Platform)
nebius · Amsterdam, Czech Republic; Remote - Europe
Remoteabout 2 months agoApply →Site Reliability Engineer (SRE) AI Infrastructure (Early Career)
nebius · Amsterdam, Netherlands
On-siteabout 2 months agoApply →Senior Site Reliability Engineer (Hardware Automation)
nebius · Amsterdam, Netherlands; Remote - Europe
Remoteabout 2 months agoApply →Senior Site Reliability Engineer
manychat · Amsterdam, Netherlands
On-siteabout 2 months agoApply →Site Reliability Engineer (SRE) - Security
imc · Amsterdam, Netherlands
On-siteabout 2 months agoApply →Site Reliability Engineer
Optiver · Amsterdam, Netherlands
On-siteabout 2 months agoApply →<p>Optiver's Production Engineering teams manage our live trading environment, which is active across 50+ global exchanges and hundreds of thousands of interconnected financial products. Our world-class infrastructure is a combination of vastly distributed systems, high-performance computing and low-latency trading algorithms, as well as high-throughput dataflows and data analysis. To keep pace with the ever-changing markets, Optiver must always react with speed, and therefore takes an in-house approach to building, analysing, managing and improving our custom infrastructure. Taking this autonomy a step further, Optiver's huge volumes of data are stored at our own data centre rather than in the cloud. Our technological independence enables us to evolve our systems on a daily basis, operating with tight feedback loops and quick development cycles. The challenge is balancing innovation with reliability and performance in such a complex, time-pressured environment – and that's exactly where you come in.&nbsp;</p> <p><strong>What you'll do<br></strong>As a Production Engineer, you'll be responsible for deploying, maintaining, monitoring and improving the reliability, scalability and performance of our in-house built trading systems. At Optiver, reliability means prevention and swift resolution across our software, to guarantee optimum functioning even in the most extreme market conditions. Having a decisive nature, engineering mindset and preference for simple solutions are essential to keeping our systems reliable. <br><br>The Production Engineer role is crucial for Optiver's trading activities. Your scope will cover market access, monitoring and compliance, strategy evaluation and management, performance tuning and trading automation. Sitting in the middle of our buzzing trading floor, you will have constant face-to-face interaction with end-users (Traders and Researchers as well as fellow Engineers). This is an engineering role, not a support role, so you'll set the standards for our production environment. As part of our core business, you will make a real, direct impact on our ability to trade and trading results. No two days are the same in our high-stakes, flat-hierarchy environment.<br><br>Learn more about a day in the life of a Production Engineer (we call this job Application Engineer internally) in<a href="https://www.optiver.com/insights/blog/a-day-in-the-life-of-an-application-engineer/">&nbsp;this blog post&nbsp;</a>and&nbsp;<a href="https://www.youtube.com/watch?v=c6jMMj1g02w">this video</a>.&nbsp;</p> <p><strong>Who you are</strong></p> <ul> <li>You are a pragmatic, logical thinker.</li> <li>With a facilitating and enabling attitude, you always strive to find clean and simple solutions.</li> <li>You are naturally curious