Senior DevOps Engineer

DevOps Engineer · Senior · Full Time

London - The River Building HQGBP 140k – 180k3w ago
Apply for this role

Opens Deliveroo's application page

Role

What you'll do.

This Senior DevOps Engineer role at Deliveroo involves designing and operating scalable Kubernetes and AWS infrastructure that powers engineering teams across Wolt, Deliveroo, and DoorDash. You'll split time between hands-on infrastructure ownership and building production-quality automation tooling, taking end-to-end responsibility for complex cloud platforms, CI/CD systems, and Infrastructure-as-Code frameworks while mentoring other engineers and driving platform reliability at global scale.

Responsibilities

  • End-to-End Infrastructure Initiative Ownership: Take complete ownership of complex infrastructure projects spanning Kubernetes cluster management, AWS cloud architecture, networking infrastructure, and Infrastructure-as-Code frameworks. Make autonomous technical decisions within scope and document architecture-level choices without requiring oversight.
  • Production Kubernetes Operations and Management: Design, deploy, and operate Kubernetes clusters at scale handling cluster scaling, security hardening, version upgrades, multi-tenancy configurations, and workload orchestration across environments. Ensure high availability and fault tolerance for mission-critical container platforms.
  • AWS Cloud Infrastructure Design and Optimization: Architect and maintain secure, scalable AWS infrastructure including compute resources, networking, identity and access management, account structuring, and cost optimization. Implement cloud security best practices and establish governance frameworks for infrastructure teams.
  • Infrastructure-as-Code Framework Development: Build and maintain self-service Infrastructure-as-Code solutions using Terraform and other provisioning frameworks that empower engineering teams to safely provision environments and infrastructure with consistency and reduced manual toil.
  • Automation and Tooling Development: Write production-quality code in Go, Python, or similar languages to build automation frameworks, operational tooling, and internal platforms that eliminate toil, improve developer experience, and increase engineering velocity across the organization.
  • Platform Reliability and SRE Leadership: Serve as on-call infrastructure responder, lead incident investigations and root cause analysis, drive mitigation strategies, and establish reliability standards including monitoring, alerting, dashboards, and operational runbooks across the platform.
  • Cross-Team Enablement and Knowledge Sharing: Unblock product engineering teams through code reviews, design reviews, and proactive technical guidance. Share expertise through technical documentation, playbooks, internal tech talks, and best practice evangelism to multiply platform team impact.
  • Developer Experience and Productivity Optimization: Drive quality objectives focused on improving developer experience and engineering productivity across the infrastructure platform. Implement self-service capabilities, reduce deployment friction, and measure success through adoption and engineering velocity metrics.
  • Multi-Organization Infrastructure Alignment: Partner closely with platform engineering teams across Wolt, Deliveroo, and DoorDash to establish shared infrastructure standards, align on tooling decisions, and coordinate the joint infrastructure roadmap across organizations.
  • Technical Leadership and Mentorship: Model engineering excellence through constructive feedback delivery, recognizing colleague contributions, and actively supporting hiring initiatives. Conduct technical interviews, mentor junior engineers, and contribute to team culture and capability development.

Qualifications

What we look for.

Technical

  • Kubernetes Production Operations

    Deep hands-on expertise operating Kubernetes in production environments with demonstrated experience in cluster scaling, security hardening, version upgrades, multi-tenancy architectures, and container orchestration at scale. Proven ability to diagnose and resolve Kubernetes networking, storage, and compute issues.

  • AWS Cloud Platform Expertise

    Strong production experience with AWS services including EC2, VPC networking, IAM authentication and authorization, account structuring, security groups, and load balancing. Competency in AWS cost management, resource optimization, and infrastructure security posture assessment.

  • Infrastructure-as-Code with Terraform

    Proficiency in Terraform for infrastructure provisioning and state management. Experience designing and maintaining IaC frameworks, modules, and patterns that scale across multiple teams and environments. Strong understanding of version control integration and IaC best practices.

  • Software Engineering and Automation

    Solid software engineering skills in Go, Python, or similar systems languages applied to building internal tooling, automation frameworks, and operational utilities. Ability to write clean, testable code that scales beyond simple scripting and includes proper error handling and observability.

  • SRE Principles and Incident Management

    Deep understanding of Site Reliability Engineering principles including chaos engineering, fault-tolerant architecture design, capacity planning methodologies, and incident response procedures. Experience implementing monitoring and alerting strategies that provide reliable incident detection.

  • CI/CD Pipeline Architecture

    Experience designing and maintaining continuous integration and continuous deployment pipelines using platforms like GitHub Actions, GitLab CI, or similar. Understanding of deployment automation, release orchestration, and build infrastructure optimization.

Education

  • Computer Science or Related Field

    Bachelor's degree in Computer Science, Engineering, Mathematics, or equivalent professional experience demonstrating strong technical foundation and systems thinking capability.

Experience

  • Complex Technical Initiative Ownership

    Demonstrated track record of owning and delivering large-scale infrastructure projects end-to-end with architecture-level decision-making authority. Evidence of understanding long-term consequences of technical decisions and ability to communicate rationale across engineering organizations.

  • Cross-Functional Team Collaboration

    Proven experience working effectively with multiple engineering teams, translating infrastructure requirements across diverse stakeholder groups, and establishing shared standards. Strong collaboration skills in matrix organizations and globally distributed teams.

  • Mentorship and Knowledge Leadership

    Experience unblocking other engineers through technical reviews, knowledge documentation, best practice guidance, and direct mentorship. Track record of elevating team capabilities and fostering knowledge sharing culture within engineering organizations.

  • Production System Operations

    Significant hands-on operational experience supporting production systems at scale. On-call incident response experience, root cause analysis facilitation, and driving follow-through on systemic improvements from incidents.

Skills

Required

  • Kubernetes

    Production-grade expertise in Kubernetes cluster operations, including node management, pod scheduling, networking models, storage orchestration, security policies, and cluster upgrades.

  • AWS

    Hands-on proficiency with AWS cloud platform services, networking, identity management, and infrastructure provisioning at production scale.

  • Terraform

    Strong capability in Infrastructure-as-Code using Terraform for managing cloud infrastructure, maintaining state files, and creating reusable module libraries.

  • Go or Python

    Proficiency in at least one systems programming language for building automation tools, internal platforms, and operational utilities with production quality standards.

  • Linux Systems Administration

    Deep understanding of Linux operating systems, shell scripting, system performance tuning, network configuration, and troubleshooting production systems.

  • Container Technologies

    Comprehensive knowledge of containerization concepts, Docker container runtime, image layer optimization, container security, and multi-container deployment patterns.

  • CI/CD Platforms

    Experience with continuous integration and deployment tooling such as GitHub Actions, GitLab CI, Jenkins, or equivalent for automating build and deployment workflows.

  • Monitoring and Observability

    Proficiency in implementing monitoring, alerting, logging, and tracing solutions. Experience with observability platforms and defining SLOs/SLIs for reliability measurement.

Preferred

  • GitOps Tooling

    Nice to have

    Hands-on experience with GitOps platforms such as Argo CD or Flux for declarative, version-controlled infrastructure and application deployments.

  • Multi-Cluster and Multi-Region Operations

    Nice to have

    Proven experience designing, deploying, and operating Kubernetes environments spanning multiple clusters or geographic regions at production scale.

  • Service Mesh Technologies

    Nice to have

    Familiarity with service mesh platforms like Istio or Linkerd for managing microservice communication, traffic policies, and observability across distributed systems.

  • GCP and Multi-Cloud Platforms

    Nice to have

    Experience working with Google Cloud Platform, Azure, or other cloud providers. Understanding of multi-cloud architecture patterns and cloud-agnostic infrastructure design.

  • Cloud Cost Optimization and FinOps

    Nice to have

    Experience implementing cloud cost management practices, reserved instance optimization, workload right-sizing, and FinOps methodologies for Kubernetes and cloud infrastructure.

  • CNCF Ecosystem Contributions

    Nice to have

    Contributions to open-source projects within the Cloud Native Computing Foundation ecosystem, demonstrating commitment to cloud-native technologies and community involvement.

  • Progressive Delivery Techniques

    Nice to have

    Expertise in advanced deployment strategies including blue-green deployments, canary releases, feature flagging, and traffic shadowing for reducing deployment risk.

  • Multi-Organization Infrastructure Unification

    Nice to have

    Experience collaborating across multiple organizations to establish shared infrastructure platforms, harmonize tooling, and coordinate technical standards at enterprise scale.

Tech stack

Languages

GoPythonBash/Shell

Frameworks

KubernetesGitHub ActionsArgo CDIstio

Databases

PostgreSQLMongoDBRedis

Tools

TerraformDockerAWS ServicesPrometheusGrafanaHelm

Other

gRPCKafkaSRE PracticesInfrastructure Security

Compensation

Pay and benefits.

Base·GBP 140,000 – 180,000

Equity·Stock options

Benefits

  • Comprehensive Healthcare Coverage

    Country-specific healthcare benefits including medical, dental, and vision coverage with options for family extension. Mental health and wellness support programs.

  • Generous Paid Time Off

    Competitive annual leave allowance with additional paid time to support charitable causes of your choice, flexible time off policies, and regional holiday allowances.

  • Pension and Retirement Planning

    Employer-funded pension contributions with matching options and retirement planning support to secure long-term financial wellbeing.

  • Parental and Family Support

    Comprehensive parental leave policies for birth, adoption, and surrogacy with flexible return-to-work options. Childcare support and family benefits.

  • Professional Development

    Learning and development budget for conferences, courses, certifications, and training programs. Internal tech talks and knowledge-sharing opportunities with leading-edge infrastructure teams.

  • Workplace Flexibility

    Flexible working arrangements supporting remote, hybrid, and office-based preferences. Global team environment with asynchronous-friendly collaboration tools.

  • Wellness Programs

    Gym membership allowances, wellness activities, stress management resources, and employee assistance programs supporting physical and mental wellbeing.

  • Diversity and Inclusion Initiatives

    Commitment to diverse hiring practices, inclusive workplace culture, and employee resource groups. Workplace adjustments and accessibility support for all candidates and employees.

Full posting

Original listing.

Senior DevOps Engineer

About the team

The Core Engineering teams build the systems and tooling that enable Wolt, Deliveroo, and DoorDash engineers to develop, ship, and operate software at scale. We treat infrastructure, CI / CD, and developer tooling as software products, engineered with the same care as customer-facing systems. Our mission is to increase engineering velocity, reliability, and security as we transform the way the world eats and shops.

About the role

Working in the Infrastructure team, you’ll be responsible for the foundational compute and cloud components that product engineering teams build on: our Kubernetes clusters, related cloud infrastructure and networking, and the Infrastructure-as-Code and provisioning frameworks that let our teams safely self-serve environments and capacity. Expect to split your time between hands-on operational ownership and writing production-quality code to build automation / tooling that removes toil and friction, while also designing and running scalable systems in AWS.

What you’ll do

  • Take ownership of complex infrastructure initiatives end-to-end, spanning Kubernetes, cloud infrastructure, networking, and infrastructure-as-code frameworks - making most technical decisions within scope without needing oversight.

  • Break initiatives down into executable work, identify dependencies, provide accurate estimates, and document project-level technical decisions and how infrastructure components interact.

  • Design, deploy, and operate secure, scalable systems in AWS, and know when and how to pull in other teams to support cross-cutting initiatives.

  • Independently plan and drive technical improvements to the compute and cloud platform, using data and clear success metrics to prioritize work.

  • Build and own self-service tooling, automation, and Infrastructure-as-Code frameworks that let other engineering teams provision environments and infrastructure safely and consistently.

  • Unblock other teams through code and design reviews, proactively answering questions, and sharing knowledge via guides, playbooks, and internal tech talks.

  • Serve as a first responder during on-call rotations for core infrastructure, lead root cause analysis, and drive incident mitigation and follow-through.

  • Drive team-wide quality objectives for the Infrastructure platform: reliability, operational readiness (alerting, dashboards, runbooks), and developer experience / productivity tooling.

  • Act as a role model for other engineers by sharing constructive feedback, recognizing colleagues' contributions, and supporting hiring through interviewing and recruiting activities.

  • Partner closely with platform teams across Wolt, Deliveroo, and DoorDash to align on shared standards, tooling, and the joint infrastructure roadmap.

Requirements

  • A track record of owning complex technical initiatives end-to-end, contributing to making architecture-level decisions and understanding their long-term consequences.

  • Deep, hands-on expertise operating Kubernetes and container orchestration in production, including cluster scaling, upgrades, and multi-tenancy.

  • Strong experience with cloud platforms (AWS preferred), including networking, IAM, account structuring, and cost management.

  • Proficiency in Infrastructure-as-Code (Terraform preferred) and building provisioning or automation frameworks that other engineers rely on.

  • Solid software engineering skills (Go, Python, or similar) applied to building internal tooling and automation, not just operational scripting.

  • Broader knowledge across supportive technologies (e.g. Kafka, gRPC, Postgres / MongoDB, Redis, CI / CD such as GitHub Actions) and the ability to weigh trade-offs between them for a given problem.

  • Solid understanding of SRE principles, including incident response, fault-tolerant architecture, and capacity planning.

  • Experience unblocking and mentoring other engineers (e.g. code / design reviews, knowledge sharing, or driving best practices), and willingness to contribute to hiring across interviewing, tech talks, etc.

  • Excellent collaboration and communication skills; this role works closely with many engineering teams, increasingly operating at global scale.

Nice-to-haves

  • Experience with GitOps tooling (e.g. Argo CD, Flux), progressive delivery, or service mesh technologies.

  • Experience operating multi-cluster or multi-region Kubernetes environments at scale.

  • GCP exposure, especially compute and networking aspects.

  • Familiarity with cloud cost optimization / FinOps practices for cloud and Kubernetes workloads.

  • Experience contributing to multi-org infrastructure unification efforts.

  • Contributions to open-source infrastructure or platform engineering projects, especially in the CNCF ecosystem.

Why Deliveroo

Our mission is to transform the way you shop and eat, bringing the neighbourhood to your door by connecting consumers, restaurants, shops and riders. We are transforming the way the world eats and shops by making access to food and products more convenient and enjoyable. We give people the opportunity to buy what they want, as they want it, when and where they want it.

We are a technology-driven company at the forefront of the most rapidly expanding industry in the world. We are still a small team, making a very large impact, looking to answer some of the most interesting questions out there. We move fast, value autonomy and ownership, and we are always looking for new ideas.

Workplace & Benefits

At Deliveroo we know that people are the heart of the business and we prioritise their welfare. Benefits differ by country, but we offer many benefits in areas including healthcare, well-being, parental leave, pensions, and generous annual leave allowances, including time off to support a charitable cause of your choice. Benefits are country-specific, please ask your recruiter for more information.

Diversity

At Deliveroo, we believe a great workplace is one that represents the world we live in and how beautifully diverse it can be. That means we have no judgement when it comes to any one of the things that make you who you are - your gender, race, sexuality, religion or a secret aversion to coriander. All you need is a passion for (most) food and a desire to be part of one of the fastest-growing businesses in a rapidly growing industry.

We are committed to diversity, equity and inclusion in all aspects of our hiring process. We recognise that some candidates may require adjustments to apply for a position or fairly participate in the interview process. If you require any adjustments, please don't hesitate to let us know. We will make every effort to provide the necessary adjustments to ensure you have an equitable opportunity to succeed.

Redirects to Deliveroo's application page.

Other roles

More at Deliveroo.

View all 20 roles