Production Reliability Engineer at ASX | AU | Rezi

Production Reliability Engineer at ASX

Production Reliability Engineer

ASX · AU

3 weeks ago

Production Reliability Engineer

ASX · AU

a month ago
Resume preview

Impress employers and recruiters.
Choose from hundreds of resume examples.

Target Resume Now
Resume preview

Tailor your resume to this Production Reliability Engineer role.

Rezi rewrites your resume against ASX's job description. Free.

Resume score gauge reading 58 out of 100

Don't guess if your resume is good enough.

See how it scores against the Production Reliability Engineer posting at ASX — free, in seconds.

About the Role

The role is accountable for ensuring the reliability, availability, performance, observability, and operational resilience of critical business platforms through proactive monitoring, automation, incident management, capacity planning, and continuous service improvement, enabling secure and stable technology services that meet business and customer expectations.

Responsibilities

  • Ensure the availability, reliability, scalability, and resilience of critical technology platforms and services.
  • Design, implement, and maintain comprehensive monitoring, logging, alerting, and observability solutions.
  • Develop meaningful dashboards, operational metrics, and health indicators to provide real-time visibility of platform performance.
  • Drive continuous improvement initiatives to reduce service disruptions and improve platform stability.
  • Proactively identify, assess, and mitigate reliability risks across production environments.
  • Conduct root cause analyses (RCA) and post-incident reviews, driving permanent resolutions and preventative actions.
  • Maintain operational runbooks, playbooks, and recovery procedures to support rapid incident resolution.
  • Support 24x7 production environments and participate in on-call rotations where required.
  • Develop and maintain infrastructure, deployment, and operational automation using Infrastructure as Code (IaC) principles.
  • Build automated recovery mechanisms to minimise manual intervention.
  • Champion engineering best practices across operational and development teams.
  • Conduct capacity planning, performance analysis, and workload forecasting to ensure services can meet current and future demand.
  • Identify and resolve performance bottlenecks across applications, infrastructure, databases, and integrations.
  • Partner with development teams to improve deployment reliability, release automation, and operational supportability.
  • Work closely with Engineering, Architecture, Security, Infrastructure, Product, and Operations teams to improve service outcomes.

Requirements

  • 5+ years’ experience in Site Reliability Engineering, DevOps, Systems Engineering, Cloud Engineering, Platform Engineering, or Production Support environments.
  • Strong scripting experience in Python, Java or PowerShell.
  • Strong knowledge of Linux, Windows, networking, and distributed systems.
  • Experience with CI/CD tooling and automated deployment pipelines.
  • Strong understanding of containerisation and orchestration technologies, including Docker and Kubernetes.
  • Experience with observability platforms such as Grafana, Splunk, ITRS Geneos, OpenTelemetry, AWS CloudWatch or similar.
  • Knowledge of database technologies, including Oracle, SQL Server, PostgreSQL or NoSQL platforms.
  • Familiarity with integration technologies, including APIs, real-time messaging and streaming architectures such as Kafka.
  • Non-functional test planning and execution experience.
  • Experience working in the Capital Markets industry, with a solid understanding of product and transaction lifecycles.
  • Passionate about solving problems, troubleshooting software issues and triaging environment issues.
  • Excellent verbal and written communication skills, with the ability to work with internal and external stakeholders.
  • Lateral thinker who brings forward ideas that will automate repeatable tasks to reduce work effort and ensure quality deliverables.
  • Learns quickly and enjoys the challenge of learning new systems.
  • Compassionate, empathetic and self-motivated.
  • Able to take ownership and not be afraid to acknowledge failure.
  • Candidates must be legally authorised to work in Australia on a permanent basis without any restrictions.

Skills

  • Python
  • Java
  • PowerShell
  • Linux
  • Windows
  • Networking
  • Distributed systems
  • CI/CD
  • Docker
  • Kubernetes
  • Grafana
  • Splunk
  • ITRS Geneos
  • OpenTelemetry
  • AWS CloudWatch
  • Oracle
  • SQL Server
  • PostgreSQL
  • NoSQL
  • APIs
  • Kafka

Location

  • Australia

Work Type

  • Hybrid
  • Full-time
  • Part-time

Experience Level

  • 5+ years

About the Company

  • ASX powers Australia's financial markets, enabling a stronger economic future by providing a fair and dynamic marketplace.
  • ASX is a leading global securities exchange, known as a trusted market operator and an exciting data hub.
  • The ASX team brings together talented people from a diverse range of disciplines, running critical market infrastructure.
  • We are proud to foster a workplace where diversity is celebrated and inclusion is part of our everyday culture.
  • Our employee-led networks champion LGBTIQ+ inclusion, promote gender equality, accessibility and wellbeing, inspire giving and volunteering, and celebrate cultural and religious events, creating a sense of belonging for all.
  • We are an AWEI Bronze employer and a member of the Champions of Change Coalition for gender equality, committed to a fair and inclusive workplace where everyone can thrive.

Equal Opportunity

  • We make hiring decisions based on your skills, capabilities and experience, and of how you’ll help us to live our values.
  • We encourage you to apply even if you don’t meet all the criteria of this role.
  • If you need any adjustments during the application or interview process to help you present your best self, please let us know at careers@asx.com.au.
  • We support flexible working and offer hybrid working options.
  • Even if our roles are advertised as full-time, we encourage you to apply if you are interested in part-time or other flexible working arrangements.
  • We will arrange for successful candidates to have background checks, including reference and police checks, completed as part of the on-boarding process.