Skip to content
← Back to job listings

Senior Site Reliability Engineer - Cloud Platform

jobgether · Canada

Software DevelopmentSenior LevelRemoteImported listingfull-time21 days ago

About The Role

This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Senior Site Reliability Engineer - Cloud Platform based in Canada.

Join a global cloud infrastructure team responsible for building and operating critical AWS platforms at <scale.You>’ll help engineer reliable, secure, and highly automated infrastructure used by teams across the organization.The role combines hands-on software engineering, cloud platform development, reliability, and operational <excellence.You>’ll work extensively with AWS, Python, Infrastructure as Code, observability, and automation-first practices.Your work will directly improve platform resilience, efficiency, security, and cloud cost <management.You>’ll also contribute to strategic initiatives across networking, identity, governance, and multi-account architecture.This is a remote opportunity suited to an experienced engineer who enjoys solving complex infrastructure challenges and mentoring others.

Accountabilities

  • Operate and scale production AWS infrastructure, taking ownership of the reliability and health of services responsible for provisioning, securing, and managing cloud environments.
  • Design, develop, and maintain cloud platform capabilities using Python, CloudFormation, AWS CDK, and automation-focused engineering practices.
  • Drive cloud cost optimization initiatives, identifying opportunities to improve infrastructure efficiency and deliver measurable business value.
  • Strengthen observability through monitoring, alerting, dashboards, operational tooling, and proactive identification of reliability issues.
  • Participate in on-call rotations, lead incident response, troubleshoot production issues, and drive sustainable improvements through blameless post-incident reviews.
  • Support strategic cloud initiatives spanning networking, identity and access management, governance, security guardrails, and multi-account architecture.
  • Review technical designs and code, maintain documentation and operational runbooks, and mentor other engineers.
  • Leverage AI-assisted engineering tools to accelerate development, automate repetitive tasks, and reduce operational toil.

Requirements

  • 5+ years of experience in Site Reliability Engineering, Platform Engineering, Infrastructure Engineering, or a comparable role supporting production AWS environments.
  • Strong Python development skills, with practical experience building automation, services, or infrastructure tooling for production use.
  • Hands-on experience with Infrastructure as Code, particularly AWS CloudFormation and/or AWS CDK.
  • Strong working knowledge of core AWS services, including IAM, VPC, EC2, Lambda, and managed storage or database services.
  • Solid Linux systems administration, production troubleshooting, incident management, and operational experience.
  • Experience working with multi-account AWS environments or AWS Organizations is highly valued.
  • Knowledge of AWS networking technologies such as Transit Gateway, IPAM, or BYOIP, as well as cloud cost optimization or FinOps practices, is a plus.
  • Experience with CI/CD pipelines, GitOps-style deployment practices, policy-as-code, compliance automation, or cloud governance tooling is advantageous.
  • Familiarity with AI-assisted engineering workflows and a collaborative mindset focused on knowledge sharing, continuous improvement, and mentoring.

Benefits

  • Competitive salary range of CAD $107,000–$161,000 per year, depending on skills, experience, and geographic location.
  • Potential eligibility for a 10% discretionary annual cash bonus, based on individual and company performance.
  • Potential participation in an equity plan.
  • Remote work flexibility, with occasional opportunities to connect with colleagues in person.
  • Health, dental, and vision insurance.
  • Life insurance, critical illness coverage, and accidental death and dismemberment (AD&D) coverage.
  • Healthcare spending account and employee assistance program.
  • Paid sick time, personal time, holidays, wellness days, and parental leave.
  • Retirement savings benefits and an employee stock purchase program.
  • An inclusive environment that values diverse perspectives, professional growth, mentorship, and continuous learning.
  • Opportunities to work with large-scale AWS infrastructure, modern automation practices, and AI-assisted engineering technologies.

This is an external listing. JobSpring does not represent or verify the employer. Report this listing