Performance & Observability Engineer at Herbert Smith Freehills Kramer | England | Rezi

Performance & Observability Engineer at Herbert Smith Freehills Kramer

Performance & Observability Engineer

Herbert Smith Freehills Kramer · England

1 weeks ago

Performance & Observability Engineer

Herbert Smith Freehills Kramer · England

12 days ago
Resume preview

Impress employers and recruiters.
Choose from hundreds of resume examples.

Target Resume Now

About the Role

The Performance & Observability Engineer ensures system reliability, scalability, and visibility across the technology stack. This role focuses on transitioning to full observability for deep performance insights, real-time issue detection, and proactive optimization, with expertise in performance engineering, distributed tracing, logging, metrics, and automation.

Responsibilities

  • Analyze application, database, and infrastructure performance to identify bottlenecks and inefficiencies.
  • Develop performance benchmarks and SLIs to measure service responsiveness and stability.
  • Collaborate with SRE and DevOps teams to optimize CI/CD pipelines for performance improvements.
  • Implement caching strategies, query optimization, and autoscaling to enhance system efficiency.
  • Design and implement end-to-end observability frameworks covering metrics, logs, traces, and events.
  • Instrument services using existing tools (e.g., Nexthink) to improve visibility.
  • Enable distributed tracing across microservices to enhance root cause analysis and performance debugging.
  • Standardize logging and telemetry collection across infrastructure, applications, and cloud services.
  • Define best practices and a consistent approach across development teams to improve monitoring consistency.
  • Transition from basic alerting to proactive insights, leveraging AI-driven anomaly detection.
  • Ensure comprehensive observability across frontend, backend, databases, cloud infrastructure, and networking.
  • Implement Service Level Indicators (SLIs), Service Level Objectives (SLOs), and Error Budgets to track system health.
  • Automate root cause analysis and incident detection through advanced monitoring techniques.
  • Reduce Mean Time to Detect (MTTD) and Mean Time to Resolve (MTTR) through improved observability.
  • Integrate monitoring and alerting tools with incident response platforms (e.g., ServiceNow).
  • Develop self-healing and auto-remediation mechanisms to reduce operational toil.
  • Improve alerting strategies by reducing false positives and improving signal-to-noise ratio.
  • Define observability best practices and governance models to ensure adoption across teams.
  • Ensure log retention, security, and compliance with standards (e.g., GDPR, SOC 2, PCI DSS).
  • Develop executive dashboards and reporting frameworks to showcase reliability and performance trends.

Requirements

  • Experience in using and maintaining Observability & APM Tools – Grafana experience is essential
  • Experience of using KQL
  • Performance Testing & Load Testing
  • Cloud & Infrastructure Monitoring – AWS CloudWatch, Azure Monitor, GCP Operations Suite, Kubernetes Observability.
  • Knowledge of Log Aggregation & Analysis – e.g: Grafana Loki, Splunk etc
  • Automation & Scripting – PowerAutomate, Terraform, PowerShell
  • Incident Response & ITSM – ServiceNow.
  • Sound understanding of firm’s applications, systems and tools across technology stack
  • DevOps experience is desirable
  • Nexthink experience is desirable
  • Strong problem-solving and root cause analysis skills.
  • Ability to translate observability insights into business impact for stakeholders.
  • A continuous improvement mindset, focused on reducing toil and improving efficiency.
  • Experience working in a DevOps, SRE, or Platform Engineering environment.

Skills

  • Observability
  • APM Tools
  • Grafana
  • KQL
  • Performance Testing
  • Load Testing
  • AWS CloudWatch
  • Azure Monitor
  • GCP Operations Suite
  • Kubernetes Observability
  • Log Aggregation
  • Log Analysis
  • Grafana Loki
  • Splunk
  • Automation
  • Scripting
  • PowerAutomate
  • Terraform
  • PowerShell
  • Incident Response
  • ITSM
  • ServiceNow
  • DevOps
  • SRE
  • Platform Engineering

Location

  • London

Work Type

  • Full time
  • Permanent Contract

About the Company

  • Herbert Smith Freehills Kramer is a world-leading global law firm, where our ambition is to help you achieve your goals.
  • Exceptional client service and the pursuit of excellence are at our core.
  • We invest in and care about our client relationships, which is why so many are longstanding.
  • We enjoy breaking new ground, as we have for over 170 years.
  • As a fully integrated transatlantic and transpacific firm, we are where you need us to be.
  • Our footprint is extensive and committed across the world’s largest markets, key financial centres and major growth hubs.
  • At our best tackling complexity and navigating change, we work alongside you on demanding litigation, exacting regulatory work and complex public and private market transactions.
  • We are recognised as leading in these areas.
  • We are immersed in the sectors and challenges that impact you.
  • We are recognised as standing apart in energy, infrastructure and resources.
  • And we’re focused on areas of growth that affect every business across the world.
  • All of this is achieved by supporting the growth of our people, who help us deliver on our ambition – which is to help you achieve yours.
  • Herbert Smith Freehills Kramer: Your goals. Our ambition

Equal Opportunity

  • We are committed to attracting people from all backgrounds and creating a respectful and inclusive culture where everyone thrives.
  • We see this as essential to our success, including our ability to innovate and achieve sustained high performance.
  • This is a key part of our Values—Human, Bold, and Outstanding.