About the Role
In this role, you’ll be a key contributor to our Platform Engineering team, honing your skills alongside experienced Platform Engineers with a focus on automation. Our team focuses on Service Delivery (Kubernetes), Infrastructure Management (Terraform), CI/CD (Argo, Atlantis), and SLOs/Operational Excellence. You will be empowered to fix root causes of issues and improve developer velocity and system reliability.
Responsibilities
- Become a key contributor to the team, taking responsibility for the success of subsystems.
- Participate in medium to large impact team initiatives and execute on such projects within a year.
- Help with interviewing potential teammates.
- Create technical designs that proactively address cost efficiency, security, and observability.
- Deliver technical plans, one-pagers, DRs, and other artifacts.
- Work with Kubernetes, GCP, Helm, Terraform, DataDog, ArgoCD, CircleCI, Atlantis, Docker to deliver work.
- Improve developer velocity across the company (leveraging frameworks like DORA) and harden reliability and observability.
- Participate in on-call rotations.
- Fix the root cause of issues and quell noisy monitors.
Requirements
- Systems expertise and experience with stateless/fault tolerant systems.
- Familiarity with eventing patterns and distributed paradigms.
- Leveraged agentic tooling to enhance daily work and detect/mitigate production issues.
- Ability to weigh technical and business trade-offs and anticipate future needs.
- Enjoy cross-team collaboration and believe in a 'rising tides lifts all boats' mentality.
- Data-oriented and have leveraged data to make concrete decisions.
- Familiarity with Kubernetes, GCP (or AWS/Azure), and Terraform.
Skills
- Distributed systems
- Evented systems
- Kubernetes
- GCP
- Helm
- Terraform
- DataDog
- ArgoCD
- CircleCI
- Atlantis
- Docker
- GraphQL
- Agentic tooling
- DORA frameworks
Location
- Remote
Work Type
- Full-time
Experience Level
- Proven track record
- Minimum requirements
About the Company
- Our team's current focus is on Service Delivery (Kubernetes), Infrastructure Management (Terraform), CI/CD (Argo, Atlantis), and in general all things SLOs and Operational Excellence.
- Teammates value code that is 80% of the value for 20% of the work, designs that are forward-thinking enough to be easily flexible for the next features, and leadership with a healthy dose of mindfulness and humility.
