About the Role
As a Senior Site Reliability Engineer, you will be 100% hands-on across infrastructure and software development. You will support and help inform the evolution of major systems within our multi-cloud, multi-region, active-active content serving platform that serves upwards of 25 Billion requests daily. Through a combination of technical expertise and cross-team collaboration, you will help support the reliability initiatives and collaborate on the technical strategy that scales our platform to 50 Billion requests per day and beyond.
Responsibilities
- Improve the tooling and automation of our infrastructure to minimize manual work, increase performance, and decrease the frequency and severity of incidents
- Build, maintain, and support core applications
- Monitor our systems for capacity, performance, and troubleshoot issues
- Partner with the rest of the SRE team to ensure smooth, continued delivery of our service to clients
- Demonstrate a high level of autonomy in anticipating, identifying, and addressing systemic weaknesses and opportunities for platform improvement
Requirements
- Experience in Site Reliability or Software Engineering, building and maintaining scalable, resilient services.
- Building the tooling and automation to manage those services, as well as investigating system and application metrics to diagnose and resolve performance issues.
- 4+ years experience as an SRE or Software Engineer, with a focus on Cloud platforms (AWS/GCP)
- Experience architecting and leading large-scale observability platforms, including defining observability standards and SLO frameworks.
- Experience and willingness to operate in an on-call environment, evaluating and improving monitoring and alerting systems, and developing run books to investigate and debug issues
- Strong experience with infrastructure as code tools.
- Kubernetes experience, including cluster operations, multi-tenancy strategies, and supporting teams on container orchestration best practices.
- Experience with one or more high level programming languages; NodeJS, Go, Ruby, Python, in addition Shell Scripting
- Linux experience is a must
Skills
- Prometheus
- Thanos
- Grafana Alloy
- Loki
- Tempo
- Terraform
- EKS
- GKE
- NodeJS
- Go
- Ruby
- Python
- Shell Scripting
- Linux
- AWS
- GCP
Location
- New York City
Work Type
- Onsite
Experience Level
- Senior
- 4+ years
Salary/Compensations
- $140K - 182K/year CAD
Benefits
- additional bonus
- full range of medical, financial, and/or other benefits
About the Company
- Movable Ink scales content personalization for marketers through data-activated content generation and AI decisioning.
- The world’s most innovative brands rely on Movable Ink to maximize revenue, simplify workflow and boost marketing agility.
- Headquartered in New York City with close to 600 employees, Movable Ink serves its global client base with operations throughout North America, Central America, Europe, Australia, and Japan.
Equal Opportunity
- We are committed to building a diverse and inclusive culture where all Inkers can thrive.
- If you’re excited about the role but don’t meet all of the abovementioned qualifications, we encourage you to apply.
- Our differences bring a breadth of knowledge and perspectives that makes us collectively stronger.
- We welcome and employ people regardless of race, color, gender identity or expression, religion, genetic information, parental or pregnancy status, national origin, sexual orientation, age, citizenship, marital status, ethnicity, family or marital status, physical and mental ability, political affiliation, disability, Veteran status, or other protected characteristics.
- We are proud to be an equal opportunity employer.
