About the Role
The Performance & Observability Engineer ensures system reliability, scalability, and visibility across the technology stack. This role focuses on transitioning to full observability for deep performance insights, real-time issue detection, and proactive optimization, with expertise in performance engineering, distributed tracing, logging, metrics, and automation.
Responsibilities
- Analyze application, database, and infrastructure performance to identify bottlenecks and inefficiencies.
- Develop performance benchmarks and SLIs to measure service responsiveness and stability.
- Collaborate with SRE and DevOps teams to optimize CI/CD pipelines for performance improvements.
- Implement caching strategies, query optimization, and autoscaling to enhance system efficiency.
- Design and implement end-to-end observability frameworks covering metrics, logs, traces, and events.
- Instrument services using existing tools (e.g., Nexthink) to improve visibility.
- Enable distributed tracing across microservices to enhance root cause analysis and performance debugging.
- Standardize logging and telemetry collection across infrastructure, applications, and cloud services.
- Define best practices and a consistent approach across development teams to improve monitoring consistency.
- Transition from basic alerting to proactive insights, leveraging AI-driven anomaly detection.
- Ensure comprehensive observability across frontend, backend, databases, cloud infrastructure, and networking.
- Implement Service Level Indicators (SLIs), Service Level Objectives (SLOs), and Error Budgets to track system health.
- Automate root cause analysis and incident detection through advanced monitoring techniques.
- Reduce Mean Time to Detect (MTTD) and Mean Time to Resolve (MTTR) through improved observability.
- Integrate monitoring and alerting tools with incident response platforms (e.g., ServiceNow).
- Develop self-healing and auto-remediation mechanisms to reduce operational toil.
- Improve alerting strategies by reducing false positives and improving signal-to-noise ratio.
- Define observability best practices and governance models to ensure adoption across teams.
- Ensure log retention, security, and compliance with standards (e.g., GDPR, SOC 2, PCI DSS).
- Develop executive dashboards and reporting frameworks to showcase reliability and performance trends.
Requirements
- Experience in using and maintaining Observability & APM Tools – Grafana experience is essential
- Experience of using KQL
- Performance Testing & Load Testing
- Cloud & Infrastructure Monitoring – AWS CloudWatch, Azure Monitor, GCP Operations Suite, Kubernetes Observability.
- Knowledge of Log Aggregation & Analysis – e.g: Grafana Loki, Splunk etc
- Automation & Scripting – PowerAutomate, Terraform, PowerShell
- Incident Response & ITSM – ServiceNow.
- Sound understanding of firm’s applications, systems and tools across technology stack
- DevOps experience is desirable
- Nexthink experience is desirable
- Strong problem-solving and root cause analysis skills.
- Ability to translate observability insights into business impact for stakeholders.
- A continuous improvement mindset, focused on reducing toil and improving efficiency.
- Experience working in a DevOps, SRE, or Platform Engineering environment.
Skills
- Observability
- APM Tools
- Grafana
- KQL
- Performance Testing
- Load Testing
- AWS CloudWatch
- Azure Monitor
- GCP Operations Suite
- Kubernetes Observability
- Log Aggregation
- Log Analysis
- Grafana Loki
- Splunk
- Automation
- Scripting
- PowerAutomate
- Terraform
- PowerShell
- Incident Response
- ITSM
- ServiceNow
- DevOps
- SRE
- Platform Engineering
Location
- London
Work Type
- Full time
- Permanent Contract
About the Company
- Herbert Smith Freehills Kramer is a world-leading global law firm, where our ambition is to help you achieve your goals.
- Exceptional client service and the pursuit of excellence are at our core.
- We invest in and care about our client relationships, which is why so many are longstanding.
- We enjoy breaking new ground, as we have for over 170 years.
- As a fully integrated transatlantic and transpacific firm, we are where you need us to be.
- Our footprint is extensive and committed across the world’s largest markets, key financial centres and major growth hubs.
- At our best tackling complexity and navigating change, we work alongside you on demanding litigation, exacting regulatory work and complex public and private market transactions.
- We are recognised as leading in these areas.
- We are immersed in the sectors and challenges that impact you.
- We are recognised as standing apart in energy, infrastructure and resources.
- And we’re focused on areas of growth that affect every business across the world.
- All of this is achieved by supporting the growth of our people, who help us deliver on our ambition – which is to help you achieve yours.
- Herbert Smith Freehills Kramer: Your goals. Our ambition
Equal Opportunity
- We are committed to attracting people from all backgrounds and creating a respectful and inclusive culture where everyone thrives.
- We see this as essential to our success, including our ability to innovate and achieve sustained high performance.
- This is a key part of our Values—Human, Bold, and Outstanding.
