About the Role
This role offers a hybrid work offering to be present in our London office twice a week. Due to expansion, an opportunity has become available for a Site Reliability Engineer to join our team to help us transform our existing operational workloads to an SRE approach.
Responsibilities
- Integrate tightly with Product Engineering teams.
- Follow SRE practices and maintain high standards of compliance.
- Implement a new standard of observability utilizing SLI/SLO/Error Budgets.
- Continually evolve observability platforms for greater coverage.
- Use a code-first approach to build and changes to reduce TOIL.
- Advocate a strong focus on availability, reliability, and uptime.
- Liaise and embed with Engineering teams for the constant evolution of metrics.
- Work towards planned roadmap goals.
- Actively participate in daily stand-ups and keep sprints on track.
- Maintain up-to-date documentation in JIRA & Confluence tools.
- Participate in SRE Incident Management processes.
- Act as a key Incident Commander within the Incident Management process.
- Participate in SRE On Call.
- Ensure a focus on cost efficiency for platforms & services.
- Work with team members to foster collaboration and ongoing communication with stakeholders.
Requirements
- Good experience in DevOps or SRE, with a keen interest to learn and grow as a Site Reliability Engineer.
- Observability product experience (e.g., Datadog).
- Managing services using SLI/SLO & Error Budgets.
- Experience with AWS or other cloud providers.
- Experience in HA environments.
- Automation skills through Terraform, Python, Bash or similar.
- Good SRE skills with a good understanding of SRE practices.
- Some understanding of SQL, PHP, Kubernetes, CI/CD is advantageous.
- Ability to work both independently and as part of a team.
- Ability to work under pressure and be highly reliable.
- Adaptability and flexibility to change in a fast-moving environment.
- An ability to learn new tools and processes quickly and impart that knowledge.
Skills
- DevOps
- SRE
- Observability
- Datadog
- SLI/SLO
- Error Budgets
- AWS
- Cloud Providers
- HA environments
- Terraform
- Python
- Bash
- SQL
- PHP
- Kubernetes
- CI/CD
- JIRA
- Confluence
Location
- London
Work Type
- Full Time
- Hybrid
Salary/Compensations
- £60,000 - £65,000 / year
About the Company
- Reward Gateway|Edenred is a leading digital platform for services and payments for people at work, connecting 52 million users and 2 million partner merchants in 45 countries via close to 1 million corporate clients.
- Our shared mission of ‘Making the World a Better Place to Work' and ‘Enriching connections, For good’, guides our every action and charts a sustainable path to a better future.
Equal Opportunity
- At Reward Gateway, we want all of our employees to feel comfortable bringing their passion, creativity, and individuality to work. We value all cultures, backgrounds and experiences, as we truly believe that diversity drives innovation. Express yourself, join our community and help us Make the World a Better Place to Work.
- We hire BETTER. From perks to people, our BETTER approach to hiring earns us more trust, happier people, and more world-class talent that helps us to make the world a better place to work. Find out more about Reward Gateway's approach to benefits, equality, talent, technology, empathy, and what you’ll get in return for joining our Mission at rg.co/lifeatrg.
