Impress employers and recruiters.
Choose from hundreds of resume examples.

Impress employers and recruiters.
Choose from hundreds of resume examples.
Tailor your resume to this AI Engineer - Data Platform role.
Rezi rewrites your resume against Clera's job description. Free.

Tailor your resume to this AI Engineer - Data Platform role.
Rezi rewrites your resume against Clera's job description. Free.
Don't guess if your resume is good enough.
See how it scores against the AI Engineer - Data Platform posting at Clera — free, in seconds.

Don't guess if your resume is good enough.
See how it scores against the AI Engineer - Data Platform posting at Clera — free, in seconds.
About the Role
Join a well-funded, Series A AI startup building the next generation of autonomous site reliability engineering for the enterprise. This role involves designing, building, and maintaining backend systems for an AI-driven observability platform, blending distributed systems engineering, low-level system design, performance optimization, observability, and AI integration across cloud and on-premises deployments.
Responsibilities
- Contribute to the design and implementation of scalable, resilient infrastructure systems powering AI-driven root cause analysis and observability workflows, including on-premises deployment environments.
- Work on the foundational building blocks of the infrastructure, ensuring efficient resource utilization and high performance at scale.
- Profile and tune backend systems to improve throughput, reduce latency, and eliminate bottlenecks across the stack.
- Build and maintain the internal observability stack — logs, metrics, and traces — used by AI agents to understand and act on production issues.
- Support cloud and on-premises architecture to serve both SaaS and enterprise customer deployment models.
- Work closely with engineers across the company to deliver resilient infrastructure that enables AI agents to diagnose and remediate production incidents in real time.
Requirements
- 2–5 years of hands-on backend or infrastructure engineering experience.
- Strong understanding of distributed systems design principles and trade-offs.
- Proven experience profiling and optimizing high-throughput, low-latency systems.
- Familiarity with observability tooling and concepts (logs, metrics, traces).
- Experience with hybrid or multi-environment infrastructure (cloud + on-premises).
- Interest in or experience building systems that support AI/ML workloads at scale.
- Prior experience at observability, incident management, or data infrastructure companies is highly valued.
- Visa sponsorship is not available for this role.
Skills
- Distributed systems
- Performance engineering
- Observability
- Cloud infrastructure
- On-premises infrastructure
- AI/ML integration
- Backend engineering
- Infrastructure engineering
- System design
- Performance optimization
- Observability tooling
- Datadog
- Grafana
- Splunk
Location
- New York, NY
Work Type
- On-site
Experience Level
- Mid-level
About the Company
- Well-funded, Series A AI startup.
- Building the next generation of autonomous site reliability engineering for the enterprise.
- Backed by top-tier investors.
- Trusted by some of the largest companies in the world.
- Tackling complex problems in AI: autonomously detecting, diagnosing, and remediating complex production incidents in real time.