About the Role
Working in the Infrastructure team, you’ll be responsible for foundational compute and cloud components, including Kubernetes clusters, cloud infrastructure, networking, and Infrastructure-as-Code frameworks. You will split your time between operational ownership and writing code for automation and tooling, while also designing and running scalable systems in AWS.
Responsibilities
- Take ownership of complex infrastructure initiatives end-to-end, spanning Kubernetes, cloud infrastructure, networking, and infrastructure-as-code frameworks.
- Break initiatives down into executable work, identify dependencies, provide accurate estimates, and document project-level technical decisions.
- Design, deploy, and operate secure, scalable systems in AWS.
- Independently plan and drive technical improvements to the compute and cloud platform, using data and clear success metrics to prioritize work.
- Build and own self-service tooling, automation, and Infrastructure-as-Code frameworks.
- Unblock other teams through code and design reviews, proactively answering questions, and sharing knowledge.
- Serve as a first responder during on-call rotations for core infrastructure, lead root cause analysis, and drive incident mitigation.
- Drive team-wide quality objectives for the Infrastructure platform: reliability, operational readiness, and developer experience tooling.
- Act as a role model for other engineers by sharing constructive feedback and supporting hiring.
- Partner closely with platform teams to align on shared standards, tooling, and the joint infrastructure roadmap.
Requirements
- A track record of owning complex technical initiatives end-to-end, contributing to architecture-level decisions.
- Deep, hands-on expertise operating Kubernetes and container orchestration in production.
- Strong experience with cloud platforms (AWS preferred), including networking, IAM, account structuring, and cost management.
- Proficiency in Infrastructure-as-Code (Terraform preferred) and building provisioning or automation frameworks.
- Solid software engineering skills (Go, Python, or similar) applied to building internal tooling and automation.
- Broader knowledge across supportive technologies (e.g. Kafka, gRPC, Postgres / MongoDB, Redis, CI / CD such as GitHub Actions).
- Solid understanding of SRE principles, including incident response, fault-tolerant architecture, and capacity planning.
- Experience unblocking and mentoring other engineers.
- Excellent collaboration and communication skills.
Skills
- Kubernetes
- AWS
- Terraform
- Go
- Python
- Kafka
- gRPC
- Postgres
- MongoDB
- Redis
- GitHub Actions
- SRE principles
- Incident response
- Fault-tolerant architecture
- Capacity planning
- GitOps tooling
- Argo CD
- Flux
- Progressive delivery
- Service mesh technologies
- GCP
- Cloud cost optimization
- FinOps
Location
- Remote
Work Type
- Full-time
Experience Level
- Senior
Benefits
- Healthcare
- Well-being
- Parental leave
- Pensions
- Generous annual leave allowances
- Time off to support a charitable cause
About the Company
- Our mission is to transform the way you shop and eat, bringing the neighbourhood to your door by connecting consumers, restaurants, shops and riders.
- We are transforming the way the world eats and shops by making access to food and products more convenient and enjoyable.
- We give people the opportunity to buy what they want, as they want it, when and where they want it.
- We are a technology-driven company at the forefront of the most rapidly expanding industry in the world.
- We move fast, value autonomy and ownership, and we are always looking for new ideas.
Equal Opportunity
- We are committed to diversity, equity and inclusion in all aspects of our hiring process.
- We recognise that some candidates may require adjustments to apply for a position or fairly participate in the interview process.
- If you require any adjustments, please don't hesitate to let us know.
- We will make every effort to provide the necessary adjustments to ensure you have an equitable opportunity to succeed.
