Develocity brand banner

Senior Site Reliability Engineer

DevelocityUnited KingdomPosted 18d agoonsite
via Arbeitnow

Who We Are

AI is changing how software gets built. Code production is becoming a commodity. The focus is shifting from writing code to orchestrating, verifying, and governing change – and the toolchain is the new constraint.

We are at the center of this shift. We build Develocity, a toolchain observability and intelligence platform used by some of the world's leading software organizations – Netflix, Airbnb, Spotify, SAP, major global banks, and hundreds more. Develocity helps software teams achieve delivery excellence through deep observability, build and test acceleration, and AI-powered intelligence across the entire toolchain – with current support for Gradle Build Tool, Apache Maven™, sbt, npm, and Python.

We are an AI-native company. AI is not a feature we're bolting on – it's central to how we work, how we think about our product, and where we're heading. We're investing deeply in making Develocity's unique data and decades of domain expertise accessible to both humans and AI agents, with trust, evidence, and explainability at the core of everything we build.

We have partnered with the Apache Software Foundation, the Commonhaus Foundation, the Micronaut Foundation, and other OSS projects such as Spring, Quarkus, Kotlin, JUnit, AndroidX, and many more to bring the values of Develocity also to the OSS Community.

Our Values

Seek to Understand: Everything starts with listening and understanding, and we strive to understand different viewpoints, problems, and motivations. Before we take action, we ensure we truly grasp the challenges, perspectives, and goals. 

Know the Why: We approach our work with a clear sense of purpose, ensuring every step is deliberate and focused. We take meaningful action with urgency, but never at the expense of thoughtful consideration. 

Innovate & Iterate: We embrace challenges and are not afraid to try new things, even if they might fail. With deep understanding and a clear purpose, we can develop creative and bold solutions to tackle challenges.

Own the Outcome: We are empowered to take initiative and we maintain transparency in our work and its outcomes. When we execute, we take responsibility for our decisions, measure the success of our innovations, and learn from the results.

Who You Are

We're building a new SRE team and looking for founding members to help shape how we operate. You'll be responsible for the reliability, performance, and availability of Develocity instances serving paying customers, open-source projects, and public-facing services, plus supporting infrastructure like artifact registries.

You'll work on our internally-built Cloud Application Platform, Kubernetes on AWS, and develop deep expertise in it. When incidents happen, you'll troubleshoot issues across the stack, from application to infrastructure. You'll collaborate with the Cloud Platform team to improve the tooling you depend on, and with engineering teams to build reliability into how we ship software. If you like automating things and hate doing the same task twice, you'll fit in well.

You'll be part of a distributed, remote-first team that values asynchronous communication and written documentation. Strong self-direction and clear communication across time zones are essential.

Responsibilities

  • Operate and maintain all Develocity instances and supporting services.
  • Participate in a follow-the-sun on-call rotation, owning incident response and troubleshooting issues across the stack.
  • Drive automation across application deployment, upgrades, monitoring, self-healing, and recovery.
  • Build and maintain observability for all managed services (logging, metrics, tracing, and alerting).
  • Work with engineering teams to build reliability into features from the start.
  • Run incident response and retrospectives, and make sure we learn from them.
  • Own disaster recovery, backups, and business continuity.
  • Communicate with customers during incidents and maintenance windows.
  • Optimize performance, resource usage, and costs.
  • Help evolve our SaaS operations as we grow.

Minimum qualifications

  • 5+ years in SRE, DevOps, or equivalent role operating production services at scale.
  • Strong Kubernetes experience in production environments.
  • Cloud infrastructure expertise, preferably AWS (EKS, RDS, S3, EC2).
  • Proficiency with observability tools (Prometheus, Grafana) and Infrastructure as Code (Terraform).
  • Track record of incident management and response.
  • Knowledge of SRE best practices (SLAs, SLOs).
  • Scripting proficiency (Python, Bash) for automation.
  • Experience with 24/7 on-call rotations.
  • Strong written and verbal English communication.

Preferred qualifications

  • Experience operating SaaS platforms at scale.
  • Familiarity with Develocity.
  • JVM language experience (Java, Kotlin).
  • Disaster recovery planning and execution experience.
  • Customer-facing incident communication skills.
  • Experience establishing SRE practices in new or growing teams.

What We Offer

  • A ground-floor role in a new SRE team—you'll shape how we do things, not inherit someone else's decisions.
  • Real ownership of production systems used by engineers at companies you've heard of.
  • Direct interaction with customers when things go wrong (and when they go right).
  • A culture that values automation over heroics.
  • In-person meetings, such as our annual company offsite and team meetings.
  • Work from home in a remote-first environment.
  • Competitive salaries and equity grants.

Location

  • Remote from anywhere in Europe in the GMT timezone.
  • While our team works remotely and is spread across the globe, we deeply value daily interactions and collaboration.

Find more English Speaking Jobs in United Kingdom on Arbeitnow

Similar roles

  • Solirius Reply logo

    Junior Solution Architect

    Solirius Reply·London, United Kingdom

    About Us: Solirius Reply, part of the Reply Group, is a technology consultancy and digital transformation partner that helps organisations solve complex challenges through strategy, design, engineering, and delivery. We work closely with our clients to deliver secure, accessible, user-focused services that evolve with their needs. By combining deep technical expertise with people-centred design, we create solutions that deliver meaningful,…

    • Hybrid
    • Full-time
  • Solirius Reply logo

    Solution Architect

    Solirius Reply·London, United Kingdom

    About us Solirius Reply, part of the Reply Group, is a technology consultancy and digital transformation partner that helps organisations solve complex challenges through strategy, design, engineering, and delivery. We work closely with our clients to deliver secure, accessible, user-focused services that evolve with their needs. By combining deep technical expertise with people-centred design, we create solutions that deliver meaningful,…

    • Hybrid
    • Full-time
  • Millennium Hotel and Resorts UK logo

    Casual Room Attendant

    Millennium Hotel and Resorts UK·Cardiff, United Kingdom

    Here at Millennium Hotels UK where we value your skills, encourage growth by nurturing your personality, and your dedication rewarded. You'll learn not only from your fellow colleagues, but also through our academy where you’ll be able to excel your career with apprenticeship and develop your careers within our Brands. The Copthorne Cardiff is looking for a Casual Room Attendant…

    • Full-time
  • Hutch logo

    QA Embedded Tester (12 months)

    Hutch·London, United Kingdom

    QA Embedded Tester | QA | London | Permanent We’re Hutch, a mobile games developer & publisher with studios in central London, Dundee and Canada. Our mission is to build the most diverse and engaged automotive gaming community on mobile. Our games have been played by over 300 million people, with new titles in development. We believe in putting our…

    • Hybrid
    • Full-time
  • Focus Group logo

    Customer Infrastructure Specialist

    Focus Group·United Kingdom

    We’re Hiring – Customer Infrastructure Specialist Salary – £28,000 - £32,000 (DOE) REMOTE WORKING (Can be based out of one of our Focus group Hubs if preferred- Shoreham, Birmingham, Exeter) (Internal Job Level- Senior Associate) About Us: Established in 2003, Focus Group is one of the UK's fastest-growing tech providers, empowering over 30,000 businesses nationwide. With over 1,000 employees and…

    • Remote
    • Full-time
  • Janus Henderson logo

    Senior Investment Risk Analyst

    Janus Henderson·London, United Kingdom

    Why work for us? A career at Janus Henderson is more than a job, it’s about investing in a brighter future together. Our Mission at Janus Henderson is to help clients define and achieve superior financial outcomes through differentiated insights, disciplined investments, and world-class service. We will do this by protecting and growing our core business, amplifying our strengths and…

    • Full-time
  • ATA LTD logo

    Multi Drop Delivery Driver

    ATA LTD·Cumbernauld, United Kingdom

    Join ATA North’s Growing Delivery Team ATA North is looking for reliable, professional Delivery Drivers to join our growing Glasgow operation. This self-employed opportunity offers consistent work, competitive earnings and the support you need to succeed in a fast-paced delivery environment. Pay & Vehicle Package * £180 flat day rate * Weekly payments – paid two weeks in arrears *…

    • Full-time
  • Freedom Leisure logo

    Group Exercise Instructor - Casual - Ringwood Health and Leisure Centre

    Freedom Leisure·Ringwood, United Kingdom

    Join the Energy at Freedom Leisure – Do Good Feel Good! At Freedom Leisure, we’re all about positive vibes, great people and making a real impact. Yes, we run leisure and cultural facilities, gyms, and swimming pools - but at the heart of it all, it’s our people who make the difference. As one of the UK’s leading charitable leisure…

    • Temp