AI Engineer - Data Platform at Clera | NY, US | Rezi

AI Engineer - Data Platform at Clera

AI Engineer - Data Platform

Clera · NY, US

1 months ago

AI Engineer - Data Platform

Clera · NY, US

2 months ago
Resume preview

Impress employers and recruiters.
Choose from hundreds of resume examples.

Target Resume Now
Resume preview

Tailor your resume to this AI Engineer - Data Platform role.

Rezi rewrites your resume against Clera's job description. Free.

Resume score gauge reading 58 out of 100

Don't guess if your resume is good enough.

See how it scores against the AI Engineer - Data Platform posting at Clera — free, in seconds.

About the Role

Join a well-funded, Series A AI startup building the next generation of autonomous site reliability engineering for the enterprise. This role involves designing, building, and maintaining backend systems for an AI-driven observability platform, blending distributed systems engineering, low-level system design, performance optimization, observability, and AI integration across cloud and on-premises deployments.

Responsibilities

  • Contribute to the design and implementation of scalable, resilient infrastructure systems powering AI-driven root cause analysis and observability workflows, including on-premises deployment environments.
  • Work on the foundational building blocks of the infrastructure, ensuring efficient resource utilization and high performance at scale.
  • Profile and tune backend systems to improve throughput, reduce latency, and eliminate bottlenecks across the stack.
  • Build and maintain the internal observability stack — logs, metrics, and traces — used by AI agents to understand and act on production issues.
  • Support cloud and on-premises architecture to serve both SaaS and enterprise customer deployment models.
  • Work closely with engineers across the company to deliver resilient infrastructure that enables AI agents to diagnose and remediate production incidents in real time.

Requirements

  • 2–5 years of hands-on backend or infrastructure engineering experience.
  • Strong understanding of distributed systems design principles and trade-offs.
  • Proven experience profiling and optimizing high-throughput, low-latency systems.
  • Familiarity with observability tooling and concepts (logs, metrics, traces).
  • Experience with hybrid or multi-environment infrastructure (cloud + on-premises).
  • Interest in or experience building systems that support AI/ML workloads at scale.
  • Prior experience at observability, incident management, or data infrastructure companies is highly valued.
  • Visa sponsorship is not available for this role.

Skills

  • Distributed systems
  • Performance engineering
  • Observability
  • Cloud infrastructure
  • On-premises infrastructure
  • AI/ML integration
  • Backend engineering
  • Infrastructure engineering
  • System design
  • Performance optimization
  • Observability tooling
  • Datadog
  • Grafana
  • Splunk

Location

  • New York, NY

Work Type

  • On-site

Experience Level

  • Mid-level

About the Company

  • Well-funded, Series A AI startup.
  • Building the next generation of autonomous site reliability engineering for the enterprise.
  • Backed by top-tier investors.
  • Trusted by some of the largest companies in the world.
  • Tackling complex problems in AI: autonomously detecting, diagnosing, and remediating complex production incidents in real time.