About the Role
Our mission is to make all of our customers' security-relevant data continuously available for automated detection and response, threat hunting, and other Falcon platform use cases. The systems behind NG-SIEM are growing to accommodate over 100 PB of event and action data ingested every day, up to 10 years of retention, and millions of queries per hour. As a Senior Engineer II on the newly established NG-SIEM EPICS team, you will own the reliability and scalability of the security industry's largest SIEM platform, treating these as software engineering problems. You will build the observability, automation, and scaling systems that keep the entire platform performing.
Responsibilities
- Design, build, and maintain monitoring and synthetic test suites for deep visibility into the NG-SIEM pipeline.
- Engineer orchestrated scaling solutions that treat the NG-SIEM pipeline as a unified system.
- Serve as a subject matter expert during platform-wide incidents, diagnosing and resolving multi-component failures.
- Partake in on-call rotations, providing incident commander coordination for critical platform-wide events.
- Build and refine models for end-to-end capacity forecasting.
- Develop tooling to continuously track and surface cost drivers across the platform.
- Transform manual standard operating procedures into automated remediation workflows.
- Partner with teams to triage SLO breaches, drive problem management, and ensure consistent communication during incidents.
- Identify and drive systemic improvements across teams using broad NG-SIEM knowledge.
Requirements
- A passion for reliability engineering and curiosity about how large-scale running systems behave under pressure.
- 10+ years of experience in software engineering, site reliability engineering, or platform engineering, with significant time spent on large-scale distributed systems.
- Ability to make pragmatic tradeoffs between short-term delivery needs and long-term platform goals.
- Strong proficiency in at least one systems programming language (Go, Java, Rust, or C++) and one scripting language (Python, Bash).
- Deep experience with end-to-end observability — building monitoring pipelines, defining SLIs/SLOs, and creating dashboards that drive actionable insights across multi-service architectures.
- Demonstrated ability to diagnose and resolve complex incidents spanning multiple distributed components operating 24/7.
- Experience with coordinated capacity planning and scaling for systems with significant infrastructure footprints.
- Hands-on experience with streaming platforms (Kafka or similar) and understanding of backpressure, partition management, and consumer group dynamics at scale.
- Familiarity with infrastructure-as-code, CI/CD pipelines, and automated deployment practices.
- A can-do attitude — thrive collaborating in a team and are not afraid of taking on responsibilities.
- Strong written and verbal communication skills — will lead incident communications and produce post-incident analyses.
- Comfort working across time zones with globally distributed teams.
Skills
- Go
- Java
- Rust
- C++
- Python
- Bash
- Kafka
- Infrastructure-as-code
- CI/CD pipelines
- Automated deployment practices
- End-to-end observability
- SLIs/SLOs
- Dashboards
- Incident response
- Capacity planning
- Scaling
- Streaming platforms
- Backpressure
- Partition management
- Consumer group dynamics
- Log Management
- Cybersecurity products
- Security operations workflows
- Disaster recovery planning
- Cloud-native architectures
- Serverless computing
Location
- London (United Kingdom)
- Aarhus (Denmark)
- Dublin (Ireland)
Work Type
- Hybrid
Experience Level
- Senior Engineer II
- 10+ years of experience
Benefits
- Market leader in compensation and equity awards
- Comprehensive physical and mental wellness programs
- Competitive vacation and holidays
- Paid parental and adoption leaves
- Professional development opportunities
- Employee Networks, geographic neighborhood groups, and volunteer opportunities
- Vibrant office culture with world class amenities
- Great Place to Work Certified™
About the Company
- Global leader in cybersecurity, protecting people, processes, and technologies.
- Mission: stop breaches, redefined modern security with an AI-native platform.
- Processes almost 3 trillion events per day, with growing traffic.
- Customers span all industries.
- Mission-driven company with a culture of flexibility and autonomy.
- Looking for passionate, innovative, and customer-focused individuals.
- Founded in 2011 to address sophisticated attacks with a new approach combining advanced endpoint protection and expert intelligence.
Equal Opportunity
- CrowdStrike is proud to be an equal opportunity employer.
- Committed to fostering a culture of belonging where everyone is valued and empowered to succeed.
- Support veterans and individuals with disabilities through affirmative action.
- Provides equal employment opportunity for all employees and applicants.
- Does not discriminate on the basis of race, color, creed, ethnicity, religion, sex, sexual orientation, gender identity, marital or family status, veteran status, age, national origin, ancestry, physical disability, mental disability, medical condition, genetic information, membership or activity in a local human rights commission, status with regard to public assistance, or any other characteristic protected by law.
- Bases all employment decisions on valid job requirements.
- Provides assistance for accessing information, submitting applications, or requesting accommodations.
