About the Role
Join the Observability Platform team to build and scale a world-class observability platform for a reliable and profitable cloud. This platform provides a unified ecosystem for logs, metrics, traces, alerting, and troubleshooting across all layers of the stack, enabling engineers to understand and operate systems at scale. Responsibilities include working on high-volume telemetry ingestion, distributed storage, query engines, alerting pipelines, and developing AI-assisted troubleshooting solutions.
Responsibilities
- Build and scale core services behind the observability platform
- Work on high-volume telemetry ingestion
- Work on distributed storage
- Work on query engines
- Work on alerting pipelines
- Develop new ways to help engineers make sense of operational data, including AI-assisted troubleshooting
Requirements
- 5+ years of professional software engineering experience
- Strong knowledge of Golang or willingness to quickly switch to it
- Experience building distributed backend systems
- Solid understanding of software reliability, scalability, and performance
- Ability to troubleshoot complex production issues
- Teamwork-oriented approach
- Strong communication skills
Skills
- Experience building or contributing to observability platforms
- Experience building or contributing to telemetry systems
- Experience with Prometheus
- Experience with Grafana
- Experience with Loki
- Experience with Jaeger
- Experience with OpenTelemetry
- Experience with VictoriaMetrics
- Experience with Mimir
- Experience with Tempo
- Experience using ClickHouse in production
Location
- Amsterdam
- Remote
Work Type
- Onsite
- Remote
Experience Level
- 5+ years of professional software engineering experience
Salary/Compensations
- Competitive compensation
Benefits
- Competitive compensation
- Career growth opportunities
- Learning opportunities
- Flexibility
- Ownership
- Collaborative culture
- Innovative culture
- Opportunity to work on impactful AI projects
- International environment
- Talented teams
About the Company
- Nebius is leading a new era in cloud infrastructure for the global AI economy.
- Building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment.
- Platform eliminates the cost and complexity of building large in-house AI/ML infrastructure.
- Built by engineers, for engineers.
- Owns hard problems across compute, storage, networking, and applied AI, from large-scale GPU orchestration to inference optimization.
- Listed on Nasdaq (NBIS).
- Headquartered in Amsterdam.
- Global footprint with R&D hubs across Europe, UK, North America, and Israel.
- Team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software, and AI R&D.
Equal Opportunity
- Nebius is an equal opportunity employer.
- Committed to fostering an inclusive and diverse workplace.
- Provides equal employment opportunities in all aspects of employment.
- Does not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law.
- Applicants must be authorized to work in the country of application and provide proof of employment eligibility as a condition of hire.
- Accommodations available during the application process upon request.
