
Senior Site Reliability Engineer - Cloud Platform
jobgether · Canada
About The Role
This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Senior Site Reliability Engineer - Cloud Platform based in Canada.
Join a global cloud infrastructure team responsible for building and operating critical AWS platforms at <scale.You>’ll help engineer reliable, secure, and highly automated infrastructure used by teams across the organization.The role combines hands-on software engineering, cloud platform development, reliability, and operational <excellence.You>’ll work extensively with AWS, Python, Infrastructure as Code, observability, and automation-first practices.Your work will directly improve platform resilience, efficiency, security, and cloud cost <management.You>’ll also contribute to strategic initiatives across networking, identity, governance, and multi-account architecture.This is a remote opportunity suited to an experienced engineer who enjoys solving complex infrastructure challenges and mentoring others.
Accountabilities
- Operate and scale production AWS infrastructure, taking ownership of the reliability and health of services responsible for provisioning, securing, and managing cloud environments.
- Design, develop, and maintain cloud platform capabilities using Python, CloudFormation, AWS CDK, and automation-focused engineering practices.
- Drive cloud cost optimization initiatives, identifying opportunities to improve infrastructure efficiency and deliver measurable business value.
- Strengthen observability through monitoring, alerting, dashboards, operational tooling, and proactive identification of reliability issues.
- Participate in on-call rotations, lead incident response, troubleshoot production issues, and drive sustainable improvements through blameless post-incident reviews.
- Support strategic cloud initiatives spanning networking, identity and access management, governance, security guardrails, and multi-account architecture.
- Review technical designs and code, maintain documentation and operational runbooks, and mentor other engineers.
- Leverage AI-assisted engineering tools to accelerate development, automate repetitive tasks, and reduce operational toil.
Requirements
- 5+ years of experience in Site Reliability Engineering, Platform Engineering, Infrastructure Engineering, or a comparable role supporting production AWS environments.
- Strong Python development skills, with practical experience building automation, services, or infrastructure tooling for production use.
- Hands-on experience with Infrastructure as Code, particularly AWS CloudFormation and/or AWS CDK.
- Strong working knowledge of core AWS services, including IAM, VPC, EC2, Lambda, and managed storage or database services.
- Solid Linux systems administration, production troubleshooting, incident management, and operational experience.
- Experience working with multi-account AWS environments or AWS Organizations is highly valued.
- Knowledge of AWS networking technologies such as Transit Gateway, IPAM, or BYOIP, as well as cloud cost optimization or FinOps practices, is a plus.
- Experience with CI/CD pipelines, GitOps-style deployment practices, policy-as-code, compliance automation, or cloud governance tooling is advantageous.
- Familiarity with AI-assisted engineering workflows and a collaborative mindset focused on knowledge sharing, continuous improvement, and mentoring.
Benefits
- Competitive salary range of CAD $107,000–$161,000 per year, depending on skills, experience, and geographic location.
- Potential eligibility for a 10% discretionary annual cash bonus, based on individual and company performance.
- Potential participation in an equity plan.
- Remote work flexibility, with occasional opportunities to connect with colleagues in person.
- Health, dental, and vision insurance.
- Life insurance, critical illness coverage, and accidental death and dismemberment (AD&D) coverage.
- Healthcare spending account and employee assistance program.
- Paid sick time, personal time, holidays, wellness days, and parental leave.
- Retirement savings benefits and an employee stock purchase program.
- An inclusive environment that values diverse perspectives, professional growth, mentorship, and continuous learning.
- Opportunities to work with large-scale AWS infrastructure, modern automation practices, and AI-assisted engineering technologies.
Similar roles you might like
See all →This is an external listing. JobSpring does not represent or verify the employer. Report this listing
