Product Manager, Managed Services at Fluidstack | USA | Rezi

Product Manager, Managed Services at Fluidstack

Product Manager, Managed Services

Fluidstack · USA

2 weeks ago

Product Manager, Managed Services

Fluidstack · USA

15 days ago
Resume preview

Impress employers and recruiters.
Choose from hundreds of resume examples.

Target Resume Now

About the Role

We are hiring a Product Manager to own our managed services portfolio, including SLURM and Kubernetes control planes. You will define the product vision and roadmap for how enterprises deploy, manage, and scale workloads on Fluidstack's infrastructure, covering everything from initial cluster provisioning through lifecycle management, observability, and optimization. This role is at the intersection of infrastructure, developer experience, and operational excellence, collaborating with engineering, datacenter operations, and customer-facing teams to build scalable control plane capabilities for 100k+ GPU megaclusters.

Responsibilities

  • Own the product roadmap for managed SLURM and Kubernetes offerings, including control plane architecture, autoscaling, multi-tenancy, and cluster lifecycle management.
  • Define requirements for control plane performance, reliability, and availability, including API rate limits, etcd scaling, provisioning tiers, and failure recovery mechanisms.
  • Work with engineering to design automated provisioning workflows, health monitoring systems, and node lifecycle controllers that minimize cluster downtime and maximize GPU utilization.
  • Partner with datacenter and networking teams to ensure control plane infrastructure scales seamlessly across geographic regions and supports hybrid deployment models.
  • Drive decisions on when to build vs. integrate with ecosystem tools (Rancher, OpenShift, Slurm accounting, workload orchestrators) based on customer requirements and competitive positioning.
  • Define metrics and SLAs for control plane uptime, API performance, scheduler throughput, and pod/job launch latency.
  • Conduct customer discovery to understand pain points around cluster management, job queueing, resource allocation, and multi-cluster orchestration.
  • Create product documentation, deployment guides, and reference architectures for enterprise customers running large-scale AI training and inference workloads.
  • Analyze competitive offerings from AWS EKS, Google GKE, DigitalOcean DOKS, and specialized HPC providers to inform feature prioritization and pricing strategy.

Requirements

  • 5+ years of product management experience with at least 3 years focused on infrastructure, platform, or cloud services.
  • Deep technical understanding of Kubernetes control plane architecture (kube-apiserver, etcd, scheduler, controller-manager) and SLURM job scheduling.
  • Experience building or managing infrastructure products that serve technical users (platform engineers, ML engineers, researchers).
  • Track record of shipping features that improved cluster reliability, reduced time-to-deployment, or increased resource efficiency at scale.
  • Strong grasp of distributed systems concepts: consensus protocols, failure modes, backpressure handling, and operational complexity tradeoffs.
  • Familiarity with GPU workload patterns (multi-node training, inference serving, batch processing) and how control plane design affects performance.
  • Ability to synthesize customer feedback, operational data, and competitive intelligence into clear product requirements and technical specifications.
  • Experience working with engineering teams to debug production incidents, analyze root causes, and translate findings into product improvements.
  • Comfortable navigating ambiguity and making pragmatic tradeoffs between feature completeness, time-to-market, and technical debt.

Skills

  • Kubernetes control plane architecture
  • SLURM job scheduling
  • Distributed systems concepts
  • GPU workload patterns
  • HPC schedulers (LSF, PBS, Grid Engine)
  • Cloud-native storage (Ceph, Lustre)
  • Datacenter automation

Experience Level

  • 5+ years product management experience
  • 3+ years focused on infrastructure, platform, or cloud services

Salary/Compensations

  • $180,000-$250,000

Benefits

  • Equity
  • Benefits

About the Company

  • Fluidstack exists to make humanity more free by delivering 10 to 100s of GWs of compute faster than anyone else, rethinking every layer of the stack. We acquire power, design and build data centers, and operate them with teams spanning hardware and software. Speed and scale are our key differentiators. We are building civilization-scale infrastructure for AI.
  • We hire people who care deeply about this problem space.

Equal Opportunity

  • Fluidstack is an Equal Employment Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability and protected veterans’ status, or any other characteristic protected by law.
  • Fluidstack will consider for employment qualified applicants with arrest and conviction records pursuant to applicable law.