Principal Observability Platform Engineer at Optiver | AU | Rezi

Principal Observability Platform Engineer at Optiver

Principal Observability Platform Engineer

Optiver · AU

1 months ago

Principal Observability Platform Engineer

Optiver · AU

2 months ago
Resume preview

Impress employers and recruiters.
Choose from hundreds of resume examples.

Target Resume Now
Resume preview

Tailor your resume to this Principal Observability Platform Engineer role.

Rezi rewrites your resume against Optiver's job description. Free.

Resume score gauge reading 58 out of 100

Don't guess if your resume is good enough.

See how it scores against the Principal Observability Platform Engineer posting at Optiver — free, in seconds.

About the Role

Optiver is seeking a Principal Observability Platform Engineer to enhance observability as a critical platform capability. This role involves working on the shared platform for metrics, logs, traces, events, alerts, dashboards, diagnostics, and instrumentation. The engineer will build reliable systems for other engineers, transforming a heterogeneous observability foundation into a globally consistent, reliable, and easy-to-adopt platform embedded in production system operations.

Responsibilities

  • Design, build, and operate components of Optiver’s shared observability platform across telemetry collection, ingestion, storage, query, visualisation, alerting, diagnostics, and service health.
  • Build software, services, APIs, integrations, libraries, dashboards, automation, and reusable patterns that make observability easier to adopt and more reliable to operate.
  • Improve the scalability, reliability, performance, cost-effectiveness, and operational quality of high-volume telemetry systems.
  • Improve developer and operator experience through self-service workflows, golden paths, documentation, investigation tooling, and practical platform abstractions.
  • Work with engineering, infrastructure, trading systems, research, and regional operations teams to understand production debugging needs and improve observability adoption.
  • Own the reliability and operational quality of the components you build, including service health, failure modes, monitoring, incident learnings, and continuous improvement.
  • Raise the standard for telemetry quality, instrumentation, alerting, dashboards, diagnostic workflows, and service health across Optiver.

Requirements

  • Strong engineering experience in SRE, software engineering, platform engineering, infrastructure, observability, developer tooling, or distributed systems.
  • A production mindset, with the ability to reason about failure modes, debugging workflows, service reliability, operational impact, and how systems behave under pressure.
  • Technical understanding of modern observability practices across logs, metrics, traces, events, alerting, dashboards, telemetry pipelines, diagnostics, instrumentation quality, and service health.
  • Experience designing, building, or operating reliable services, platforms, pipelines, tools, or automation used by other engineering teams.
  • Good judgement in technical trade-offs across performance, scalability, reliability, complexity, cost, and maintainability.
  • A delivery mindset, with the ability to take ambiguous platform problems and turn them into practical, reliable solutions.
  • Strong preference for candidates with experience on observability, SRE, infrastructure, platform, production engineering, or developer tooling teams in large-scale distributed systems environments, including telemetry pipelines, streaming systems, time-series data, log platforms, query systems, alerting systems, or production diagnostics tooling.

Skills

  • Observability
  • Platform Engineering
  • SRE
  • Infrastructure
  • Distributed Systems
  • Software Engineering
  • Developer Tooling
  • Production Mindset
  • Telemetry Pipelines
  • Streaming Systems
  • Time-Series Data
  • Log Platforms
  • Query Systems
  • Alerting Systems
  • Production Diagnostics Tooling
  • Kafka
  • Grafana
  • ELK/OpenSearch
  • ClickHouse
  • VictoriaMetrics
  • InfluxDB
  • Telegraf
  • Vector
  • OpenTelemetry
  • Prometheus-style systems
  • Custom telemetry collectors

Location

  • Global

Work Type

  • Full-time

Experience Level

  • Principal
  • Senior

Benefits

  • Performance-based bonus structure
  • Global profit pool
  • Training, mentorship and personal development opportunities
  • Daily breakfast, lunch and an in-house barista
  • Gym membership plus weekly in-house chair massages
  • Regular social events, including a company trip every two years
  • Guided relocation, a competitive relocation package and visa sponsorship where necessary

About the Company

  • Optiver is a tech-driven trading firm and leading global market maker.
  • For over 35 years, Optiver has been improving financial markets worldwide, making them more transparent and efficient for all participants.
  • With more than 1,400 employees in offices around the world, we’re united in our commitment to improving the market through competitive pricing, execution and thorough risk management.
  • By providing liquidity on multiple exchanges across the world, we actively trade on 70+ exchanges, where we’re trusted to always provide accurate buy and sell pricing – no matter the market conditions.

Equal Opportunity

  • Optiver is committed to diversity and inclusion.
  • We encourage applications from candidates of all backgrounds, and welcome requests for reasonable adjustments during the process.