Remote Site Reliability Engineer jobs open to Canada.
Canada? Yes. $$$? When they say. Stack? Before the essay.
12 openings. None older than 30 days.
Newest first
Senior Support Engineer - Toronto
2 days agoPython
8 years of experience, API platform expertise, Automation in support operations, Advanced monitoring and alerting, Incident response leadership, Scripting (Python), Cloud infrastructure knowledge, Cross-functional communication
Senior Site Reliability Engineer
3 days agoTypeScript · Python · Kubernetes
5 years of experience, Kubernetes (EKS, GKE), Terraform, Pulumi, GitOps (ArgoCD, Flux), Observability stack (Prometheus, Grafana, Datadog), SLI/SLO/error budget fluency, Production-quality code in Go, Python, or TypeScript, Incident response leadership, Service mesh (Istio, Cilium), Live-service game experience
Associate Infrastructure Engineer
4 days agoPython · AWS · GCP
2 years of experience, Kubernetes, AWS or GCP, GitOps tools, Observability stack, Infrastructure-as-code tools, Node, Python or Go, Debugging distributed systems, On-call rotation participation, Proactive embrace of AI
Staff Platform Engineer, Core Cloud Platform
4 days agoPython · Go · AWS
Kubernetes expertise, AWS cloud proficiency, Production SaaS systems experience, Networking and service mesh knowledge, Operational troubleshooting skills, Python and Golang proficiency, CI/CD practices advocacy, Technical roadmap design and leadership
Solutions Architect, Ethernet Networking - NVIS
8 days agoPython · Kubernetes · Docker
5 years of experience, Ethernet networking expertise, BGP, VxLAN, EVPN, Data center architecture, Network automation (Ansible, Salt, Python), Advanced network troubleshooting, Linux administration, Customer-facing experience, AI tools usage, Kubernetes, Docker, Networking simulation tools (NVIDIA Air, GNS3, EVE-NG), Network management tools (Grafana, Prometheus, Datadog)
Site Reliability Engineer
8 days agoAWS · Kubernetes
7 years of experience, Event-driven architecture, Deep AWS, Infrastructure as Code, Kubernetes, Observability and SLOs, Chaos engineering, Distributed systems debugging, Proven technical leadership, AI / MLOps infrastructure, Experience in payments industry
Operations Support Engineer
13 days ago5 years of experience, ITIL-aligned frameworks, Second-level support for mission-critical systems, Observability tools, Linux/Unix environments, Container orchestration, CI/CD pipelines, Middleware and integration technologies, Database performance monitoring, System hardening and patching, English fluency
Site Reliability Engineer III
20 days agoPython · Go · AWS
7 years of experience, AWS, Kubernetes, Terraform, Python, Golang, High-availability infrastructure, CI/CD pipelines, Automation technologies
System and Application Administrator
22 days agoAWS · Kubernetes · Docker
Linux, Perforce, GitLab, AWS, Kubernetes, Docker, Ansible, Terraform
Senior Engineer, SRE
28 days agoGCP · Kubernetes
Kubernetes, GCP, Cloud networking, Infrastructure-as-code, Distributed systems, CI/CD, Automated deployment, Operational reliability, Cost optimization, Debugging across systems
Senior Engineer, Infrastructure
28 days agoAWS · GCP · Kubernetes
GCP, AWS, Kubernetes, Infrastructure-as-code, Cloud networking, Cloud security, OpenSearch, Elasticsearch, PostgreSQL, CI/CD, Automated deployment
Cloud Operations Manager
about 1 month agoAWS
5 years of experience, AWS (EC2, RDS, S3, ECS, Fargate, IAM), CI/CD tools (GitHub Actions, Jenkins, GitLab CI, CircleCI), People management, Operational process implementation, Cloud security best practices