Remote Site Reliability Engineer jobs.

“Remote”? Depends where. $$$? When they say. Stack? Before the essay.

Where can you apply from?

90 openings. None older than 30 days.

Newest first

Cloud Operations Engineer

2 days ago
$110,000 to $127,000 per yearopen to United StatesFull-time

AWS · Azure

7 years of experience, Microsoft Azure, AWS, Linux administration, Office 365 migrations, Intune configuration, AWS GovCloud, Cloud security practices, System troubleshooting, Automation tools, Security tools knowledge, Analytical skills, Interpersonal communication, Team collaboration

CyberSheathCyberSheath is a managed security services provider specializing in cybersecurity compliance for the U.S. Defense Industrial Base (DIB), operating as a B2B entity focused on DoD contractors.

Senior Staff Site Reliability Engineer

2 days ago
$150,000 to $200,000 per yearopen to IndiaStaffFull-time

Python · AWS · Kubernetes

10 years of experience, Kubernetes, AI/ML infrastructure, OpenTelemetry, AWS CDK, Terraform, Python, Go, Distributed systems debugging, Agentic AI platforms, Automation systems, Technical strategy steering

NVIDIANVIDIA Corporation is a Santa Clara-based technology company specializing in designing GPUs and AI solutions for gaming, professional visualization, and cloud services, operating in both B2B and B2C markets globally.

Systems Operations and Administrator

2 days ago
$112,000 to $178,250 per yearopen to United StatesFull-time

5 years of experience, Linux administration, Enterprise networking standards, Data-center infrastructure, Networking knowledge (TCP/IP, DNS, DHCP), Scripting and automation, Infrastructure monitoring tools, Capacity planning, Operational process establishment

NVIDIANVIDIA Corporation is a Santa Clara-based technology company specializing in designing GPUs and AI solutions for gaming, professional visualization, and cloud services, operating in both B2B and B2C markets globally.

Sr. Staff Platform/Data Reliability Engineer, Databricks (R5537)

2 days ago
$180,000 to $270,000 per yearopen to United States + United Arab Emirates + UkraineStaffFull-time

12 years of experience, Databricks, CI/CD for data platforms, Observability, Incident management, Compute policy design, Regulated environment experience, Delta Lake, Infrastructure-as-code, Defense or aerospace industry experience

Shield AIShield AI is a San Diego-based defense technology company specializing in AI-powered autonomous drones and aircraft systems for military and commercial applications, operating primarily in the B2B sector.

Infrastructure Security Engineer

3 days ago
$140,000 to $160,000 per yearopen to United StatesFull-time

AWS · Kubernetes

5 years of experience, AWS, Kubernetes, CI/CD security integrations, Terraform, Helm, Flux, ArgoCD, Okta, Web Application Firewalls, monitoring and logging platforms, security policies as code, incident response, active Secret clearance

Sphinx DefenseSphinx Defense is a Washington, DC-based B2G SaaS provider specializing in advanced software solutions for satellite operations and national security, serving the U.S. Space Force and allied forces.

Apache Kafka Developer

3 days ago
$100,000 to $150,000 per yearopen to United StatesFull-time

Python · Kubernetes

6 years of experience, Kafka internals, Kafka security, Kafka Connect, Schema Registry, Kafka Streams, HA/DR strategies, Python scripting, Terraform, Observability tooling, Confluent Certified Administrator, Kafka on Kubernetes

Bright Vision TechnologiesBright Vision Technologies is a New Jersey-based IT staffing firm specializing in placing technical professionals in software development roles across the U.S. government and enterprise sectors.

Monitoring Engineer

3 days ago
$75,000 to $85,000 per yearopen to United StatesFull-time

Python · Java

6 years of experience, Prometheus, Grafana, Datadog, New Relic, Splunk, OpenTelemetry, distributed tracing, structured logging, Go, Python, Java, high-cardinality metrics, SLOs, error budgets, SRE principles, CI/CD integration, Linux internals, networking, container platforms, Thanos, Mimir, Cortex, Loki, Tempo, eBPF-based tooling, observability cost optimization, regulated environments

Bright Vision TechnologiesBright Vision Technologies is a New Jersey-based IT staffing firm specializing in placing technical professionals in software development roles across the U.S. government and enterprise sectors.

Observability Engineer

3 days ago
$89,000 to $110,000 per yearopen to United StatesFull-time

Python · Java

6 years of experience, Prometheus, Grafana, Datadog, New Relic, Splunk, OpenTelemetry, Distributed tracing, Structured logging, Go, Python, Java, High-cardinality metrics, SLOs, Error budgets, SRE principles, CI/CD integration, Linux internals, Networking, Container platforms, Thanos, Mimir, Cortex, Loki, Tempo, eBPF, Cost optimization, Regulated environments

Bright Vision TechnologiesBright Vision Technologies is a New Jersey-based IT staffing firm specializing in placing technical professionals in software development roles across the U.S. government and enterprise sectors.

Reliability Engineer

3 days ago
$75,000 to $95,000 per yearopen to United StatesFull-time

Python · Java · AWS

6 years of experience, Python, Go, Java, Kubernetes, Linux at scale, Prometheus, Grafana, CI/CD pipelines, Distributed system design, Incident response, SLOs and error budgets, Chaos engineering, AWS, Azure, GCP

Bright Vision TechnologiesBright Vision Technologies is a New Jersey-based IT staffing firm specializing in placing technical professionals in software development roles across the U.S. government and enterprise sectors.

IT Systems Engineer

4 days ago
12,900 to 17,700 PLN per monthopen to PolandFull-time

Python · AWS · GCP

3 years of experience, Python, PowerShell, Infrastructure as Code (IaC), CI/CD tools, AWS, GCP, Ansible, Terraform, Zabbix, Prometheus, Grafana, Analytical thinking, Bilingual (Polish and English)

CD PROJEKT REDCD PROJEKT RED is a Warsaw-based video game developer specializing in story-driven RPGs, including The Witcher series and Cyberpunk 2077, operating primarily in the B2C gaming industry with a global target market.

Service Reliability Engineer

4 days ago
$168,000 to $333,500 per yearopen to United StatesFull-time

Python · AWS · Azure

8 years of experience, Kubernetes, SLURM, large-scale cluster management, GPU hardware, high-performance computing, observability tools, incident management tools, AWS, Azure, GCP, OCI, Linux system administration, Ansible, Python, shell scripting, DNS, DHCP, storage systems, core networking, problem-solving, bare-metal infrastructure

NVIDIANVIDIA Corporation is a Santa Clara-based technology company specializing in designing GPUs and AI solutions for gaming, professional visualization, and cloud services, operating in both B2B and B2C markets globally.

Solutions Architect, Ethernet Networking - NVIS

4 days ago
$124,000 to $195,500 per yearopen to North AmericaFull-time

Python · Kubernetes · Docker

5 years of experience, Ethernet networking expertise, BGP, VxLAN, EVPN, Data center architecture, Network automation (Ansible, Salt, Python), Advanced network troubleshooting, Linux administration, Customer-facing experience, AI tools usage, Kubernetes, Docker, Networking simulation tools (NVIDIA Air, GNS3, EVE-NG), Network management tools (Grafana, Prometheus, Datadog)

NVIDIANVIDIA Corporation is a Santa Clara-based technology company specializing in designing GPUs and AI solutions for gaming, professional visualization, and cloud services, operating in both B2B and B2C markets globally.

Senior Site Reliability Engineer - Storage

4 days ago
$168,000 to $322,000 per yearopen to United StatesSeniorFull-time

Python · Go · AWS

8 years of experience, HPC storage solutions, Enterprise NAS solutions, Distributed filesystems (Lustre, GPFS), Python/Bash/Golang, Cloud services (AWS, Azure, GCP), Monitoring stacks (Prometheus, Grafana, etc.), RDMA fabrics (InfiniBand, RoCE), HPC cluster management tools (Slurm, PBS, LSF), Containerization (Docker, Kubernetes)

NVIDIANVIDIA Corporation is a Santa Clara-based technology company specializing in designing GPUs and AI solutions for gaming, professional visualization, and cloud services, operating in both B2B and B2C markets globally.

Site Reliability Engineer

4 days ago
salary not disclosedopen to WorldwideFull-time

AWS · Kubernetes

7 years of experience, Event-driven architecture, Deep AWS, Infrastructure as Code, Kubernetes, Observability and SLOs, Chaos engineering, Distributed systems debugging, Proven technical leadership, AI / MLOps infrastructure, Experience in payments industry

YunoYuno is a global payment orchestration platform headquartered in [location], specializing in B2B payment infrastructure that enables businesses to integrate over 1,000 payment methods through a single API.

Site Reliability Engineer

4 days ago
salary not disclosedopen to United StatesFull-time

Node.js · Python · Java

Ansible, Terraform, Kubernetes, Linux, Windows, Python, Java, Golang, Node.js, Nginx, HAProxy, Docker, cloud-first mindset, security-first mindset

Ontrac SolutionsOntrac Solutions is a technology consulting firm specializing in AI-driven systems and digital transformation, headquartered remotely with a focus on B2B services for healthcare, retail, and wellness industries.

Cloud Engineer – Windows & Linux Platform Automation

4 days ago
salary not disclosedopen to United StatesFull-time

Python · SQL · AWS

4 years of experience, Microsoft platform management, Linux platform management, Public cloud services (Azure, AWS, GCP), Production networking, Production automation, SQL Server, PostgreSQL, MongoDB, Elasticsearch, Python, Terraform, PowerShell, Container scheduling engines (Mesos, Docker, Kubernetes), CI/CD solutions (GitHub Actions, Jenkins, CircleCI, ArgoCD), Level-3 support in ticketing systems

Ontrac SolutionsOntrac Solutions is a technology consulting firm specializing in AI-driven systems and digital transformation, headquartered remotely with a focus on B2B services for healthcare, retail, and wellness industries.

Cloud Infrastructure Engineer – AWS

4 days ago
$130,000 to $180,000 per yearopen to United StatesFull-time

Python · AWS · Kubernetes

10 years of experience, AWS, Terraform, AWS CloudFormation, Kubernetes, DevOps, Site Reliability Engineering, Cloud security, CI/CD pipelines, Python, Bash, FinOps, AWS Organizations, Zero Trust architecture, Observability solutions

Bright Vision TechnologiesBright Vision Technologies is a New Jersey-based IT staffing firm specializing in placing technical professionals in software development roles across the U.S. government and enterprise sectors.

Cloud Infrastructure Network Engineer

4 days ago
$100,000 to $115,000 per yearopen to United StatesFull-time

Kubernetes

6 years of experience, Cloud networking, VPC/VNet design, Hybrid connectivity, Infrastructure-as-code (Terraform), Cloud security, Kubernetes networking, Multi-cloud networking, SD-WAN familiarity, eBPF-based networking tools, Regulated industries exposure

Bright Vision TechnologiesBright Vision Technologies is a New Jersey-based IT staffing firm specializing in placing technical professionals in software development roles across the U.S. government and enterprise sectors.

DevOps & SRE Engineer

4 days ago
$100,000 to $150,000 per yearopen to United StatesFull-time

Python · Java · Kubernetes

6 years of experience, Python, Go, Java, Kubernetes, Linux at scale, Observability tooling, CI/CD pipelines, SLOs, Chaos engineering, Cloud platforms

Bright Vision TechnologiesBright Vision Technologies is a New Jersey-based IT staffing firm specializing in placing technical professionals in software development roles across the U.S. government and enterprise sectors.

OpenShift Administrator

4 days ago
$70,000 to $100,000 per yearopen to United StatesFull-time

Python · Kubernetes

6 years of experience, OpenShift, Kubernetes internals, Linux administration, Infrastructure-as-code (Ansible, Terraform, Helm), CI/CD pipelines (Tekton, Jenkins, Argo CD), Scripting (Bash, Python, Go), Cluster monitoring and logging, Container image security, Service mesh (OpenShift Service Mesh, Istio, Linkerd), GitOps workflows (Argo CD, Flux)

Bright Vision TechnologiesBright Vision Technologies is a New Jersey-based IT staffing firm specializing in placing technical professionals in software development roles across the U.S. government and enterprise sectors.

Reliability Monitoring Engineer

4 days ago
$100,000 to $150,000 per yearopen to United StatesFull-time

6 years of experience, Prometheus, Grafana, Datadog, OpenTelemetry, SLOs, High-cardinality metrics, CI/CD integration, Linux internals, Container platforms, Observability cost optimization

Bright Vision TechnologiesBright Vision Technologies is a New Jersey-based IT staffing firm specializing in placing technical professionals in software development roles across the U.S. government and enterprise sectors.

Service Mesh Architect

4 days ago
$150,000 to $190,000 per yearopen to United StatesFull-time

Python · Kubernetes

6 years of experience, Istio, Linkerd, Envoy, Kubernetes, mTLS, PKI, Go, Python, Distributed tracing, Cilium, SPIFFE/SPIRE, Zero-trust networking

Bright Vision TechnologiesBright Vision Technologies is a New Jersey-based IT staffing firm specializing in placing technical professionals in software development roles across the U.S. government and enterprise sectors.

Systems Reliability Engineer

4 days ago
$100,000 to $150,000 per yearopen to United StatesFull-time

Python · Java · AWS

6 years of experience, Python, Go, Java, Kubernetes, Linux at scale, Prometheus, Grafana, CI/CD pipelines, Distributed system design, Incident response, SLOs and error budgets, Chaos engineering, AWS, Azure, GCP

Bright Vision TechnologiesBright Vision Technologies is a New Jersey-based IT staffing firm specializing in placing technical professionals in software development roles across the U.S. government and enterprise sectors.

Telemetry Engineer

4 days ago
$100,000 to $150,000 per yearopen to United StatesFull-time

Python · Java

6 years of experience, Prometheus, Grafana, OpenTelemetry, Datadog, New Relic, Splunk, Go, Python, Java, SRE principles, incident management, distributed tracing, structured logging, Linux, containers, CI/CD, observability cost optimization, eBPF-based observability

Bright Vision TechnologiesBright Vision Technologies is a New Jersey-based IT staffing firm specializing in placing technical professionals in software development roles across the U.S. government and enterprise sectors.

Staff Platform Engineer

5 days ago
$250,000 to $285,000 per yearopen to United StatesStaffFull-time

Kubernetes

10 years of experience, Kubernetes, Infrastructure-as-code (Terraform), Database proficiency (Postgres), GitOps model (Argo), Security mindset, AI in engineering, Architectural judgment, Multi-tenant isolation design, Compliance-heavy deployments (FedRAMP), Excellent communication and collaboration

CortexCortex is a fully remote B2B software company specializing in an AI-powered Internal Developer Portal that enhances engineering productivity for teams across the US.

Senior Site Reliability Engineer (SRE & Platform Reliability)

5 days ago
308,000 zł to 428,000 zł per yearopen to PolandSeniorContract

Python · Kotlin · AWS

4 years of experience, Bash, Python, Kotlin, AWS, MySQL, Kubernetes, Incident Lifecycle, Distributed systems, Capacity management, Automation, Observability, Configuration management

AffirmAffirm is a U.S.-based fintech company offering a buy now, pay later (BNPL) service that allows consumers to make purchases in installments, primarily targeting the retail sector with both B2C and B2B business models.

Senior Site Reliability Engineer (SRE & Platform Reliability)

5 days ago
€86,000 to €122,000 per yearopen to SpainSeniorFull-time

Python · Kotlin · AWS

4 years of experience, Bash, Python, Kotlin, AWS, MySQL, Kubernetes, Incident Lifecycle, Distributed systems, Capacity management, Automation, Observability, Configuration management

AffirmAffirm is a U.S.-based fintech company offering a buy now, pay later (BNPL) service that allows consumers to make purchases in installments, primarily targeting the retail sector with both B2C and B2B business models.

Cloud Engineer – Windows & Linux Platform Automation

5 days ago
salary not disclosedopen to United StatesFull-time

Python · SQL · AWS

4 years of experience, Microsoft platform management, Linux platform management, Public cloud services (Azure, AWS, GCP), Production networking, Configuration-management tools, SQL Server, PostgreSQL, MongoDB, Elasticsearch, Python, Terraform, PowerShell, Container scheduling engines (Mesos, Docker, Kubernetes), CI/CD solutions (GitHub Actions, Jenkins, CircleCI, ArgoCD), Level-3 support in ticketing systems

Ontrac SolutionsOntrac Solutions is a technology consulting firm specializing in AI-driven systems and digital transformation, headquartered remotely with a focus on B2B services for healthcare, retail, and wellness industries.

Principal Site Reliability Engineer

7 days ago
salary not disclosedopen to IndiaPrincipalFull-time

Python

10 years of experience, Openshift, Nutanix AHV, VMware vSphere, RedHat OpenShift, Python, Go, Ansible, GitHub, 24x7 operations

DellDell Technologies designs and develops hardware and software for infrastructure, enterprise solutions, and data management.

Senior Platform Engineer - Platform Metal | Ireland | Remote

7 days ago
€104,000 to €124,800 per yearopen to United Kingdom + Ireland + SpainSeniorFull-time

Python · Kubernetes

Kubernetes, Terraform, Datacenter experience, Distributed systems, Go, Python, Shell, Crossplane, Cluster-API, Tinkerbell, Talos, Ceph, CSP experience

Grafana LabsGrafana Labs is a San Francisco-based B2B observability platform specializing in open-source IT systems monitoring and data visualization for enterprises globally.