Devops Engineer at Test Triangle | City of London, England, GB | Rezi

Devops Engineer at Test Triangle

Devops Engineer

Test Triangle · City of London, England, GB

1 months ago

Devops Engineer

Test Triangle · City of London, England, GB

2 months ago
Resume preview

Impress employers and recruiters.
Choose from hundreds of resume examples.

Target Resume Now

About the Role

We are seeking a skilled DevOps Engineer to manage and enhance our EMEIA infrastructure. This role involves maintaining existing VM systems while leading the development of a new managed environment, likely based on Kubernetes (Kube/EKS) or AWS@Apple. You will collaborate with various teams and vendors to ensure our infrastructure supports rapid, AI-assisted development cycles in a fast-paced, evolving environment.

Responsibilities

  • Lead the design and build-out of a new managed container environment.
  • Contribute to the selection of the new environment architecture.
  • Own the migration of existing VM-based workloads onto the new platform.
  • Establish and maintain the standard workflow for deploying solutions.
  • Configure and maintain networking between Kube and Apple’s internal systems.
  • Own namespace and compute provisioning on the shared Kube cluster.
  • Manage credentials, service accounts, and access controls.
  • Act as the go-to expert on internal network topology.
  • Own and manage cloud infrastructure across EMEIA using internal cloud tooling.
  • Manage certificates, firewalls, resource pools, networking, and access controls.
  • Ensure infrastructure is appropriately sized, resilient, and cost-efficient.
  • Maintain accurate documentation of infrastructure topology and configuration.
  • Maintain and operate existing virtual machines.
  • Build and maintain standardized, repeatable provisioning processes.
  • Manage package deployment, software repositories, databases, and web servers.
  • Own the patching and update lifecycle for managed systems.
  • Implement and maintain monitoring, alerting, and observability.
  • Proactively identify risks, bottlenecks, and failure patterns.
  • Define and track appropriate SLIs/SLOs for critical services.
  • Conduct post-incident reviews and drive lasting improvements.
  • Support AI-augmented development by providing the infrastructure scaffold for rapid iteration.
  • Act as a pragmatic partner to developers, unblocking deployment and catching risks.
  • Actively use AI tools to accelerate your own work.
  • Take ownership of vague or ambiguous production issues and drive them to resolution.
  • Deliver short-term fixes rapidly while tracking long-term root cause resolutions.
  • Maintain a pragmatic balance between speed-of-recovery and quality-of-fix.

Requirements

  • Proven experience in a DevOps, infrastructure, or platform engineering role.
  • Hands-on experience with Kubernetes — deploying, configuring, and operating workloads.
  • Experience containerizing applications: writing Dockerfiles, managing images, publishing to a registry, and debugging container-level issues.
  • Strong networking fundamentals: DNS, TLS/SSL certificates, firewall rules, load balancing, VPNs, and service-to-service connectivity.
  • Comfort operating in environments where the architecture is still being defined.
  • Hands-on experience with RHEL (or equivalent enterprise Linux) — provisioning, hardening, package management (yum/dnf), systemd services.
  • Experience managing cloud infrastructure, ideally in an enterprise private/hybrid cloud environment.
  • Experience with infrastructure-as-code or configuration management tooling (e.g. Terraform, Ansible, Puppet, or similar).
  • Solid scripting ability in Bash and at least one higher-level language (Python preferred).
  • Experience with monitoring and observability tooling (e.g. Prometheus, Grafana, Datadog, or similar).
  • Strong incident diagnosis skills — able to work from vague symptoms to root cause using logs, metrics, and reasoning.
  • Comfortable working with AI-generated or AI-assisted codebases.
  • Clear written and verbal communication — able to translate infrastructure complexity for non-technical stakeholders.
  • Experience with AWS or AWS@Apple, particularly EKS (Desirable).
  • Familiarity with Apple’s internal platform tooling: Kube, Shield, Appleconnect, Floodgate, or similar (Desirable).
  • Experience integrating with Snowflake (Desirable).
  • Experience with CI/CD pipelines (Desirable).
  • Exposure to security tooling, vulnerability scanning, or compliance frameworks (Desirable).
  • Familiarity with secrets management tooling (Desirable).
  • Experience working in a regulated or enterprise environment with change management processes (Desirable).

Skills

  • Kubernetes
  • Containerization (Docker)
  • Networking (DNS, TLS/SSL, Firewalls, Load Balancing, VPNs)
  • RHEL/Linux Administration
  • Cloud Infrastructure Management
  • Infrastructure as Code (Terraform, Ansible, Puppet)
  • Scripting (Bash, Python)
  • Monitoring and Observability (Prometheus, Grafana, Datadog)
  • Incident Diagnosis
  • AI-assisted Development Tools
  • Communication
  • AWS/EKS
  • CI/CD Pipelines
  • Security Tooling
  • Secrets Management

Location

  • EMEIA

Work Type

  • Full-time

Experience Level

  • Mid-level
  • Senior

About the Company

  • We are a company that produces solutions rapidly using AI-assisted, structured development, enabling ideas to move from concept to deployment faster than ever.