Principal Site Reliability Engineer - Austin, Texas at ShipperHQ | TX, United States | Rezi

Principal Site Reliability Engineer - Austin, Texas at ShipperHQ

Principal Site Reliability Engineer - Austin, Texas

ShipperHQ · TX, United States

1 months ago

Principal Site Reliability Engineer - Austin, Texas

ShipperHQ · TX, United States

a month ago
Resume preview

Impress employers and recruiters.
Choose from hundreds of resume examples.

Target Resume Now

About the Role

ShipperHQ is seeking a Principal Site Reliability Engineer to lead the evolution of our cloud platform, reliability strategy, and infrastructure architecture. This hands-on leadership role focuses on designing scalable, resilient systems and establishing engineering best practices to enable rapid and confident team execution. The ideal candidate enjoys building platforms, thrives in a fast-paced, AI-first engineering culture, and is passionate about solving complex technical challenges and improving the developer experience.

Responsibilities

  • Own the technical vision and roadmap for ShipperHQ's cloud infrastructure, reliability, and platform engineering initiatives.
  • Design, build, and maintain highly available, scalable, and secure cloud infrastructure in AWS.
  • Architect and evolve Infrastructure as Code (Terraform) standards across all environments.
  • Design and optimize CI/CD pipelines for fast, reliable, and repeatable software delivery.
  • Define and implement reliability standards, SLOs, SLIs, error budgets, and incident management best practices.
  • Lead the design and implementation of observability, monitoring, logging, and alerting across the platform.
  • Build self-service platform capabilities and automation to empower engineering teams and reduce operational overhead.
  • Drive infrastructure modernization initiatives, including containerization, orchestration, and platform scalability.
  • Partner with Security to implement cloud security best practices, compliance controls, and governance.
  • Collaborate with Engineering teams to improve application reliability, performance, and operational excellence.
  • Lead technical decision-making for infrastructure architecture and serve as a trusted advisor across engineering.
  • Mentor engineers and promote best practices in cloud architecture, automation, reliability, and operational excellence.
  • Evaluate and introduce new technologies to improve scalability, reliability, developer productivity, and operational efficiency.
  • Participate in incident response, root cause analysis, and continuous improvement efforts for production systems.

Requirements

  • 10+ years of experience in Site Reliability Engineering, Platform Engineering, DevOps, Cloud Infrastructure, or Software Engineering.
  • Proven experience designing and operating large-scale, highly available cloud infrastructure in AWS.
  • Strong software engineering background with the ability to write production-quality code and automation.
  • Expert-level experience with Infrastructure as Code, preferably Terraform.
  • Deep experience designing and maintaining modern CI/CD pipelines using GitLab or similar platforms.
  • Strong knowledge of Kubernetes, containerized workloads, and cloud-native architectures.
  • Extensive experience with observability platforms, distributed tracing, logging, monitoring, and incident response.
  • Experience defining and implementing SLOs, SLIs, and reliability engineering best practices.
  • Strong understanding of networking, security, Linux systems administration, and cloud architecture.
  • Experience supporting high-traffic SaaS applications and mission-critical production environments.
  • Excellent problem-solving skills with the ability to simplify complex technical challenges.
  • Demonstrated ability to influence technical direction without direct authority while mentoring engineers across multiple teams.
  • Experience working in Agile development environments and partnering closely with cross-functional engineering teams.

Skills

  • AWS
  • Terraform
  • GitLab CI/CD
  • Kubernetes
  • Observability platforms
  • Distributed tracing
  • Logging
  • Monitoring
  • Incident response
  • SLOs
  • SLIs
  • Networking
  • Security
  • Linux systems administration
  • Cloud architecture
  • SaaS applications

Location

  • Austin, TX

Work Type

  • Hybrid
  • Full-time

Experience Level

  • Principal

Salary/Compensations

  • Compensation is based on experience

Benefits

  • 22 days of PTO plus public holidays
  • 401k Match
  • Medical, Dental, and Vision Insurance
  • Maternity and Paternity Leave

About the Company

  • ShipperHQ is a trusted leader in the e-commerce shipping space, with over 15 years of experience helping merchants deliver better checkout experiences.
  • Founded in 2009, we power shipping logic and checkout optimization for thousands of brands, from DTC disruptors to enterprise retailers, in 150+ countries.
  • Based in Austin with a global team, we’re a fast-moving, product-led company shaping the future of e-commerce logistics.
  • We are an agile, fast-moving team that likes to roll up our sleeves and solve some of the biggest issues in shipping.
  • We foster a collaborative learning culture that promotes continuous growth and innovation.
  • We are proud to be a team that’s as diverse as the merchants we serve.
  • With honesty, responsiveness, and innovation at the center of all we do, we remain committed to hiring the right people for the job, regardless of race, background, religion, or eccentricity.

Equal Opportunity

  • As a member of the e-commerce community, we take responsibility to empower shops large and small to grow and thrive through the power of technology to heart.
  • With honesty, responsiveness, and innovation at the center of all we do, we remain committed to hiring the right people for the job, regardless of race, background, religion, or eccentricity.