About the Role
We’re looking for an Infrastructure Engineering Manager to help shape the future of AI-powered applications. In this role, you’ll bridge the gap between AI research and production, leading a team that is turning innovative prototypes into scalable, high-performance enterprise solutions. Your team will work on interactive AI applications, enterprise SaaS products, and platform capabilities that redefine how businesses leverage AI. We stay close to the customer experience, translating real-world deployment challenges into platform improvements, automation, and scalable playbooks that make every future rollout smoother than the last.
Responsibilities
- Define and execute the infrastructure roadmap aligned with business and engineering priorities.
- Lead the design and implementation of scalable, secure, and reliable infrastructure systems.
- Set and maintain SLAs/SLOs for platform uptime, performance, and developer experience.
- Manage the engineering team and drive technical delivery
- Design, build, and optimize backend services for advanced AI-driven applications, focusing on AI agents, evaluation tooling, and automation
- Influence the culture, values, and processes of a growing engineering team
- Inspire and mentor engineers.
- Work closely with product, security, and engineering leadership to align on goals and priorities.
Requirements
- At least 5 years of relevant experience and at least 2+ years of experience managing infrastructure or platform teams.
- Proven experience with cloud platforms such as AWS, GCP, Azure or OCI.
- Proven experience with kubernetes deployments on on-prem infrastructure.
- Deep understanding of CI/CD pipelines (e.g., Circle CI, Github Actions), infrastructure-as-code (e.g., Terraform), and container orchestration (e.g., Kubernetes).
- Experience managing production environments with high availability, reliability, and scalability requirements.
- Familiarity with monitoring, alerting, and incident response best practices.
- Experience working with modern developer platforms and internal tooling to improve engineering velocity.
- Solid foundation and real-world experience in network engineering.
Location
- Remote
Work Type
- Full-time
Experience Level
- Manager
About the Company
- At Scale, our mission is to develop reliable AI systems for the world's most important decisions. Our products provide the high-quality data and full-stack technologies that power the world's leading models, and help enterprises and governments build, deploy, and oversee AI applications that deliver real impact. We work closely with industry leaders like Meta, Ernst & Young, Mayo Clinic, Time Inc., the Government of Qatar, and U.S. government agencies including the Army and Air Force. We are expanding our team to accelerate the development of AI applications.
Equal Opportunity
- We believe that everyone should be able to bring their whole selves to work, which is why we are proud to be an inclusive and equal opportunity workplace. We are committed to equal employment opportunity regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability status, gender identity or Veteran status.
- We are committed to working with and providing reasonable accommodations to applicants with physical and mental disabilities. If you need assistance and/or a reasonable accommodation in the application or recruiting process due to a disability, please contact us at accommodations@scale.com. Please see the United States Department of Labor's Know Your Rights poster for additional information.
- We comply with the United States Department of Labor's Pay Transparency provision.
