Software Engineer - Tools & Infrastructure / DevOps at Cerebras Systems | Sunnyvale, California, US | Rezi

Software Engineer - Tools & Infrastructure / DevOps at Cerebras Systems

Software Engineer - Tools & Infrastructure / DevOps

Cerebras Systems · Sunnyvale, California, US

2 weeks ago

Software Engineer - Tools & Infrastructure / DevOps

Cerebras Systems · Sunnyvale, California, US

18 days ago
Resume preview

Impress employers and recruiters.
Choose from hundreds of resume examples.

Target Resume Now

About the Role

Cerebras Systems builds the world's largest AI chip, enabling industry-leading training and inference speeds that transform AI application user experiences and unlock real-time iteration. We partner with leading AI organizations, including OpenAI, to deploy massive-scale AI infrastructure.

Responsibilities

  • Contribute to the development and maintenance of CICD pipelines, ensuring reliable and efficient build, test, and release workflows.
  • Manage artifact lifecycle systems including versioning, storage, distribution, and dependency management.
  • Partner with development teams to design and improve code review workflows, branching strategies, and automated integration processes.
  • Provision, monitor, and optimize cloud infrastructure to support CI workloads, balancing cost efficiency with performance and reliability.
  • Troubleshoot build failures, pipeline bottlenecks, and infrastructure issues, driving root-cause analysis and implementing lasting fixes.
  • Contribute to internal build infrastructure, test infrastructure, tooling, and automation that improves developer velocity and engineering productivity.
  • Contribute to the company’s efforts on AI tooling to boost engineering productivity efficiently.

Requirements

  • 2-5 years of professional experience in a DevOps, infrastructure, or software engineering role.
  • Familiarity with CICD systems and hands-on experience in building or maintaining automated build and deployment pipelines.
  • Understanding of artifact repository management and software packaging concepts.
  • Experience with cloud computing platforms (AWS preferred) and programmatic resource provisioning.
  • Proficiency with distributed version control systems, code review processes, and repository management.
  • Foundational knowledge of operating system concepts (Linux/Unix), networking fundamentals, and scripting for automation.
  • Experience with containerization and container orchestration (Kubernetes preferred).
  • Strong troubleshooting skills and a methodical approach to debugging distributed systems.
  • Curiosity about how large-scale infrastructure is built, operated, and improved.
  • Willingness to participate in on-call and incident response as part of the Dev Productivity org.

Skills

  • CICD systems
  • Artifact repository management
  • Software packaging
  • Cloud computing platforms (AWS preferred)
  • Programmatic resource provisioning
  • Distributed version control systems
  • Code review processes
  • Repository management
  • Operating system concepts (Linux/Unix)
  • Networking fundamentals
  • Scripting for automation
  • Containerization
  • Container orchestration (Kubernetes preferred)
  • Troubleshooting
  • Infrastructure-as-code tools and practices (preferred)
  • Python scripting (preferred)
  • Shell scripting (preferred)
  • Build systems
  • Build graph optimization
  • Observability practices (monitoring, logging, alerting)

Location

  • Remote

Work Type

  • Full-time

Experience Level

  • 2-5 years of professional experience

Education Level

  • BS/MS in Computer Science or a related field, or equivalent practical experience

Benefits

  • Build a breakthrough AI platform beyond the constraints of the GPU.
  • Publish and open source cutting-edge AI research.
  • Work on one of the fastest AI supercomputers in the world.
  • Job stability with startup vitality.
  • Simple, non-corporate work culture that respects individual beliefs.

About the Company

  • Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs.
  • Our architecture delivers industry-leading training and inference speeds, over 10 times faster than GPU-based hyperscale cloud inference services.
  • We work with leading model labs, global enterprises, and cutting-edge AI-native startups.
  • OpenAI recently announced a multi-year partnership with Cerebras to deploy 750 megawatts of scale, transforming key workloads with ultra high-speed inference.

Equal Opportunity

  • Cerebras Systems is committed to creating an equal and diverse environment and is proud to be an equal opportunity employer.
  • We celebrate different backgrounds, perspectives, and skills.
  • We believe inclusive teams build better products and companies.
  • We strive to build a work environment that empowers people to do their best work through continuous learning, growth, and support of those around them.