Open role
Site Reliability Engineer (SRE), Cloud Operations
Rbc
TORONTO, Ontario, CanadaPosted Jul 31, 2026 · 2d ago
Full-timeSeniorOn-siteFintech
About this role
Join the Platform Engineering & AI Operations team to shape how the bank operates, monitors, and self-heals its cloud platforms. This role focuses on reducing toil, building automation for enterprise-scale infrastructure, and establishing reliable operational practices for intelligent, autonomous operations.
What we are looking for
6- Support scalable, secure, and highly available architectures across private and public cloud platforms
- Automate infrastructure workflows and eliminate toil using Python, Ansible, and Terraform
- Extend self-healing automation capabilities for routine operational tasks
- Drive automation, CI/CD, and Infrastructure as Code practices
- Minimize risk of reliability failures through proactive alerting and anomaly detection
- Participate in on-call rotation for platform support and incident management
Skills mentioned
8PythonAWSAzureGCPKubernetesLinuxTerraformCI/CD
