Impress employers and recruiters.
Choose from hundreds of resume examples.

Impress employers and recruiters.
Choose from hundreds of resume examples.
Tailor your resume to this Senior Service Reliability Engineer role.
Rezi rewrites your resume against Fitch Group's job description. Free.

Tailor your resume to this Senior Service Reliability Engineer role.
Rezi rewrites your resume against Fitch Group's job description. Free.
Don't guess if your resume is good enough.
See how it scores against the Senior Service Reliability Engineer posting at Fitch Group — free, in seconds.

Don't guess if your resume is good enough.
See how it scores against the Senior Service Reliability Engineer posting at Fitch Group — free, in seconds.
About the Role
The Senior Service Reliability Engineer will embed with Fitch Ratings development squads, partnering with developers across multiple locations to ensure excellence in Fitch Ratings services, with a focus on new AI development.
Responsibilities
- Lead the delivery of reliable, scalable, mission-critical services.
- Guide squads on Kubernetes and modern deployment patterns.
- Mentor associate engineers and set best practices.
- Partner closely with Fitch Ratings Development Squads and Operations to design and advance service builds, automation, AI tooling, and operational excellence.
- Partner with Core Engineering to architect and govern GitHub Actions CI/CD with quality gates, canary/blue‑green strategies, and AI‑assisted redeploy checks.
- Own observability in Datadog, defining SLIs/SLOs, dashboards, alerting, and MS Teams integrations.
- Reduce incidents via telemetry-driven automation and blameless postmortems.
- Champion AI‑enabled operations using AWS Bedrock/SageMaker and Model Context Protocol (MCP) for log analysis, anomaly detection, incident triage, and workflow orchestration.
- Establish adoption guardrails for AI-enabled operations.
- Define and enforce cloud guardrails and security controls in partnership with Security and Risk.
- Influence cross‑functional roadmaps, lead complex release planning, and drive strategic platform initiatives.
- Serve as an escalation point and participate in the L3 on‑call rotation.
Requirements
- Deep, hands-on experience in SRE, DevOps, or Platform Engineering across both AWS and Azure.
- Strong track record operating Docker and Kubernetes in production environments.
- Highly proficient in administering both Linux and Windows.
- Practical, enterprise-level experience supporting IIS/.NET applications as well as Java Spring Boot services.
- Experience building and maintaining CI/CD pipelines (primarily GitHub Actions; Bamboo experience a plus) with DevSecOps principles baked in.
- Experience integrating security scans, policy-as-code, and compliance gates.
- Confident scripting in Python, PowerShell, or Bash.
- Experience with cloud security best practices (IAM, secrets management, container/image scanning).
- Understanding of core infrastructure fundamentals (networking, storage, DNS).
- Experience with APM/telemetry tooling.
- Practical experience with agentic AI for operations—incident triage, runbooks, and change management—with clear guardrails, auditability, and human-in-the-loop controls.
- Experience supporting AI/ML workloads at scale: SageMaker endpoints, GPU node groups, autoscaling, and Kubernetes-based model serving.
- Experience with policy-as-code (OPA) and compliance implementation across CIS, NIST, ISO 27001, with automated remediation integrated via CSPM tools (e.g., Wiz).
- Experience applying AI in CI/CD, observability, and incident response using AWS DevOps Agent, Claude Code, or others with Skills and Model Context Protocol (MCP).
- Hands-on Agile delivery experience, actively participating in stand-ups and sprint ceremonies.
Skills
- SRE
- DevOps
- Platform Engineering
- AWS
- Azure
- Docker
- Kubernetes
- Linux administration
- Windows administration
- IIS/.NET applications
- Java Spring Boot services
- CI/CD pipelines
- GitHub Actions
- Bamboo
- DevSecOps
- Python scripting
- PowerShell scripting
- Bash scripting
- Cloud security best practices
- IAM
- Secrets management
- Container/image scanning
- Networking
- Storage
- DNS
- APM/telemetry tooling
- Agentic AI for operations
- AI/ML workloads
- SageMaker endpoints
- GPU node groups
- Autoscaling
- Kubernetes-based model serving
- Policy-as-code (OPA)
- CIS compliance
- NIST compliance
- ISO 27001 compliance
- CSPM tools
- AWS DevOps Agent
- Claude Code
- Model Context Protocol (MCP)
- Agile delivery
Location
- Toronto
Work Type
- Hybrid
Experience Level
- Senior
Salary/Compensations
- CAD 120,000 - CAD 150,000
Benefits
- Hybrid Work Environment (2-3 days in office)
- Dedicated trainings
- Leadership development programs
- Mentorship programs
- Retirement planning programs
- Tuition reimbursement programs
- Comprehensive healthcare offerings
- Family-friendly policies
- Generous global parental leave plan
- Collaborative workplace
- Employee Resource Groups
- Paid volunteer days
- Matched funding for donations
- Opportunities to volunteer in the community
About the Company
- Fitch Group is a leading, global financial information services provider delivering vital credit and risk insights, robust data, and dynamic tools to champion more efficient, transparent financial markets.
- With over 100 years of experience and colleagues in over 30 countries, Fitch Group’s culture of credibility, independence, and transparency is embedded throughout its structure.
- Fitch Group includes Fitch Ratings, one of the world’s top three credit ratings agencies, and Fitch Solutions, a leading provider of insights, data and analytics.
- Fitch Group has dual headquarters in London and New York and is owned by Hearst.
- Fitch's Technology & Data Team is a dynamic department where innovation meets impact.
- The Technology & Data Team includes the Chief Data Office, Chief Software Office, Chief Technology Office, Emerging Technology, Shared Technology Services, Technology, Risk and the Executive Program Management Office (EPMO).
- The team is driven by investment in cutting-edge technologies like AI and cloud solutions.
- Fitch's Technology & Data Team is recognized by Built In as a “Best Place to Work in Technology” 3 years in a row.
- Fitch Group SRE provides Service Reliability Engineering expertise to Fitch’s development organizations.
- This squad joins Core Engineering, Networking, and other SRE groups as part of Cloud Infrastructure & Platform Engineering (CI&PE).
- Fitch Group SRE serves as subject matter experts in cloud technologies, systems engineering, infrastructure automation, and DevOps tooling across Fitch Group.
- Fitch is committed to providing global securities markets with objective, timely, independent and forward-looking credit opinions.
- Fitch requires employees to take every precaution to avoid conflicts of interest or any appearance of a conflict of interest.
Equal Opportunity
- Fitch is proud to be an Equal Opportunity and Affirmative Action Employer.
- We evaluate qualified applicants without regard to race, color, national origin, religion, sex, sexual orientation, gender identity, disability, protected veteran status, and other statuses protected by law.