Your career compass awaits
Create a free account to unlock
- ✓ See how you match this role
- ✓ AI resume tailored to this specific job
- ✓ Inside Track Companion — find your insider contact
- ✓ Skills gap analysis and upskill plan
- ✓ Interview prep kit for this role
- ✓ Relocation concierge — salary, tax, cost of living
Free — no credit card required
Senior Site Reliability Engineer (DevTools)
ICT
💰 Salary Not shared
📊 Market 6,500 GBP–9,500 GBP/mo (78,000 GBP–114,000 GBP/yr)
About the Company
Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure. Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI. Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D.
About the Role
We're an SRE team within DevTools, looking for someone ready to help maintain and grow our systems. We run 25k builds a day in TeamCity, store 100 TB of artifacts in Artifactory, and work with a massive monorepo in GitLab - comparable in scale to what you'd find at FAANG companies. We modify GitLab and build our own TeamCity plugins to give users a product that meets their needs. We're also experimenting with AI - we have our own RAG setup and are figuring out how to operate in the age of agents. Our goal: understand users' problems and requests, define metrics that capture the problem, improve those metrics, and verify that the user's problem is actually gone.
Responsibilities
- ● You will improve services based on user feedback.
- ● You will build fault-tolerant, self-healing architecture.
- ● You will find ways to speed up our systems and reduce user friction.
- ● You will modify well-known closed- and open-source solutions to support our users.
Requirements
- ● You have a combination of SRE and SWE experience, with our code being in Java/Kotlin, Go, Python, and Ruby.
- ● You have an understanding of what's happening under the hood in Unix-like systems and the JVM.
- ● You have a passion for improving the user experience.
- ● You have the ability to adapt quickly on the fly in a fast-changing environment.
- ● It will be an added bonus if you have experience in Platform Engineering.
- ● It will be an added bonus if you have experience operating GitLab or another VCS and TeamCity or another CI system.
- ● It will be an added bonus if you have experience with Spring and operating Java monoliths.
Benefits
- ● You will receive competitive compensation.
- ● You will have career growth and learning opportunities.
- ● You will have flexibility and ownership.
- ● You will be part of a collaborative and innovative culture.
- ● You will have the opportunity to work on impactful AI projects.
- ● You will be part of an international environment and talented teams.
Confirm Application
Are you applying to ?
We'll track this in your dashboard as Applied.
Saving to your tracker…
Stand out: message someone at Nebius.
A short, personal note to a recruiter can get your application looked at sooner.
⚡ Full Resume Sandbox Canvas