Remote Site Reliability Engineer jobs.
“Remote”? Depends where. $$$? When they say. Stack? Before the essay.
Where can you apply from?
79 openings. None older than 30 days.
Newest first
Service Mesh Engineer
2 days agoPython · Kubernetes
7 years of experience, Istio, Linkerd, Envoy, Kubernetes, mTLS, PKI, Go, Python, Distributed tracing, Cilium, SPIFFE/SPIRE, Zero-trust networking
Cloud Infrastructure Network Engineer
2 days agoKubernetes
6 years of experience, Cloud networking, VPC/VNet design, Hybrid connectivity, Infrastructure-as-code (Terraform), Cloud security and network controls, Kubernetes networking, Multi-cloud networking, SD-WAN familiarity, eBPF-based networking tools, Regulated industries exposure
Kafka Engineer
2 days agoKubernetes
7 years of experience, Kafka internals, Kafka security, Kafka Connect, Infrastructure-as-code, Observability tooling, DevOps practices, Kubernetes experience, Streaming data governance
Observability Engineer
2 days agoPython · Java
6 years of experience, Prometheus, Grafana, Datadog, New Relic, Splunk, OpenTelemetry, distributed tracing, structured logging, Go, Python, Java, high-cardinality metrics, SLOs, error budgets, SRE principles, CI/CD integration, Linux internals, networking, container platforms, Thanos, Mimir, Cortex, Loki, Tempo, eBPF, cost optimization, regulated environments
OpenShift Administrator
2 days agoPython · Kubernetes
6 years of experience, OpenShift, Kubernetes internals, Linux administration, Infrastructure-as-code (Ansible, Terraform, Helm), CI/CD pipelines (Tekton, Jenkins, Argo CD), Scripting (Bash, Python, Go), Container image security, Cluster monitoring tools, Service mesh (OpenShift Service Mesh, Istio, Linkerd)
Senior Site Reliability Engineer (SRE)
5 days agoPython · Java · AWS
Remote within the United States; must be legally authorized to work there without current or future employer-sponsored visa. Required: 6–10+ years in SRE, infrastructure or backend systems engineering; ownership of reliability for complex distributed production systems; cloud infrastructure (AWS, GCP or Azure); observability, incident management and performance; programming for automation (Go, Python or Java); incident leadership and cross-team influence. Preferred: building SLO/on-call practices, Kubernetes, infrastructure as code such as Terraform, performance engineering. Incident response and sustainable on-call workload are core responsibilities.
DevOps & SRE Engineer
5 days agoKubernetes
More than 5 years SRE/DevOps, Linux and Kubernetes at scale, observability, CI/CD automation
Telemetry Engineer
5 days ago6+ years overall, 5+ SRE/observability, Prometheus/Grafana/OpenTelemetry, Datadog/New Relic/Splunk
Platform Reliability Engineer
5 days agoPython · Java · Go
6+ years overall, 5+ SRE/DevOps, Linux/Kubernetes, Python or Golang or Java, distributed systems
Site Observability Engineer
5 days agoPython · Java · Go
6+ years SRE/observability, Prometheus/Grafana, OpenTelemetry, Datadog, Golang/Python/Java
Senior Platform Engineer I
6 days agoKubernetes
5+ years platform/SRE, production Kubernetes, Terraform/Crossplane/Pulumi, on-call incident response
Systems Observability Specialist
6 days ago5+ years SRE/observability, Prometheus + Grafana, 1 of Datadog/New Relic/Splunk, OpenTelemetry, no new H-1B
DataDog Observability Engineer - Part Time (R-00231)
7 days agoUS citizenship, 5+ years observability/SRE, Datadog, Terraform, 3+ listed certifications
Observability Engineer
7 days ago12 years of experience, Prometheus, Grafana, Datadog, OpenTelemetry, Distributed tracing, High-cardinality metrics, SLOs, CI/CD integration, Linux internals, eBPF-based observability
Site Reliability Engineer (SRE)
7 days agoPython · AWS · Azure
10 years of experience, Kubernetes, Python, Go, Prometheus, Grafana, CI/CD pipelines, Chaos engineering, Distributed systems, SLOs and error budgets, Cloud platforms (AWS, Azure, GCP)
MSP Engineer, Network and Security - Rotating shifts
7 days agoPython
3 years of experience, Cloudflare, WAF, DDoS mitigation, Terraform, Python, Bash, DNS, HTTP/HTTPS, French, English, Managed services experience
Cloud Infrastructure Engineer – AWS
8 days agoPython · AWS · Kubernetes
Bachelor’s or Master’s degree in CS, IT, Engineering or related field; 10+ years in IT/cloud engineering including 5+ years designing and operating enterprise AWS; EC2/VPC/IAM/S3/RDS/Lambda and other AWS services; Terraform, AWS CDK or CloudFormation; production EKS, ECS or Kubernetes; CI/CD/DevOps/GitOps; Python and Bash (Go/PowerShell preferred); IAM, encryption, compliance and observability. Preferred: advanced AWS certifications and other qualifications in posting.
Cloud Infrastructure Network Engineer
8 days agoKubernetes
Bachelor’s degree in CS, Networking or related field; 5+ years networking with substantial cloud networking; at least one major cloud provider; routing, switching and BGP; Direct Connect, ExpressRoute or equivalent hybrid connectivity; Terraform for cloud networking; security controls; Kubernetes networking/service mesh fundamentals; packet-level troubleshooting. Preferred: cloud networking certification, multi-cloud, SD-WAN/SASE, eBPF, regulated environments.
Cloud Networking Engineer
8 days agoPython · Kubernetes
Bachelor’s degree in Computer Science or related field; 5+ years in platform engineering, SRE or networking; production Istio or Linkerd; Envoy, Kubernetes networking/CNI/ingress, mTLS/PKI and certificate lifecycle, distributed tracing; Go or Python; networking troubleshooting. Preferred: multi-cluster mesh, Cilium/eBPF, SPIFFE/SPIRE, open-source contributions, enterprise zero-trust.
Container Platform Engineer
8 days agoPython · Kubernetes
Bachelor’s degree in CS, Engineering or related field; 5+ years operating production container platforms including 3+ years on Red Hat OpenShift; Kubernetes/OpenShift internals, Linux, Ansible/Terraform/Helm, Tekton/Jenkins/Argo CD, Bash/Python/Go, cluster observability and image security. Preferred: Red Hat certification, public cloud, service mesh, regulated environments and GitOps.
Monitoring Engineer
8 days agoPython · Java
Bachelor’s degree in Computer Science or related field; 5+ years in SRE, platform engineering or observability; Prometheus, Grafana and one commercial platform such as Datadog/New Relic/Splunk; OpenTelemetry, tracing, structured logging; Go, Python or Java; high-throughput metrics/log pipelines; SLOs/error budgets; CI/CD and incident management; Linux, networking and containers. Preferred: Thanos/Mimir/Cortex/Loki/Tempo, eBPF, observability cost optimization.
Service Mesh Architect
8 days agoPython · Kubernetes
Bachelor’s degree in Computer Science or related field; 5+ years in platform engineering, SRE or networking; production Istio or Linkerd; Envoy, Kubernetes networking/CNI/ingress, mTLS/PKI/certificates, distributed tracing; Go or Python; networking and control-plane troubleshooting. Preferred: multi-cluster mesh, Cilium/eBPF, open-source contributions, SPIFFE/SPIRE, enterprise zero-trust.
Reliability Engineer
8 days agoPython · Java · AWS
Bachelor’s degree in Computer Science, Engineering or related field; 5+ years of SRE, DevOps or production engineering for large distributed systems; Python, Go or Java; Linux at scale; production Kubernetes; observability; CI/CD; distributed system design; incident response and post-incident reviews. Preferred: SLOs/error budgets, chaos engineering, AWS/Azure/GCP, capacity planning and service mesh.
Cloud Infrastructure Engineer (Open LMS) Colombia, Remote
9 days agoPython · AWS
AWS, Terraform, Puppet or similar, Python daemons, Linux, distributed systems, on-call
Service Mesh Engineer
9 days agoPython · Kubernetes
Headline says 7+ years; detailed required qualifications specify 5+ years of platform engineering, SRE or networking, production Istio or Linkerd, Envoy, Kubernetes CNI, mTLS/PKI, distributed tracing and Go or Python. Cilium, SPIFFE/SPIRE and zero-trust networking are preferred.
OpenShift Platform Engineer
9 days agoPython · Kubernetes
Headline says 8+ years; detailed required qualifications specify 5+ years operating production container platforms including at least 3 years on Red Hat OpenShift, bachelor's degree, Kubernetes/OpenShift internals, Linux, Ansible or Terraform or Helm, Tekton or Jenkins or Argo CD, Bash or Python or Go, observability and image security. Service mesh and GitOps are preferred.
DevOps & SRE Engineer
12 days agoPython · Java · AWS
6 years of experience, Python, Go, Java, Kubernetes, Linux at scale, Prometheus, Grafana, CI/CD pipelines, SLOs, Chaos engineering, AWS, Azure, GCP
Telemetry Engineer
12 days agoPython · Java
6 years of experience, Prometheus, Grafana, OpenTelemetry, Datadog, New Relic, Splunk, Go, Python, Java, SRE principles, distributed tracing, structured logging, Linux, containers, CI/CD, observability cost optimization, eBPF-based observability
Senior Infrastructure Engineer | Permanent WFH | Night Shift
12 days ago5 years of experience, Windows Server Administration, Active Directory, Microsoft 365 Administration, VMware vSphere, Backup and recovery solutions, Networking technologies, PowerShell, Cloud platforms
Kubernetes Service Engineer
12 days agoPython · Kubernetes
6 years of experience, Istio, Linkerd, Envoy, mTLS, Kubernetes, Go, Python, Distributed tracing, Service mesh architecture, Traffic management policies, Zero-trust networking