Remote Site Reliability Engineer jobs.

“Remote”? Depends where. $$$? When they say. Stack? Before the essay.

Where can you apply from?

58 openings. None older than 30 days.

Newest first

Senior Site Reliability Engineer

1 day ago
salary not disclosedopen to United StatesSeniorFull-time

AWS · Kubernetes

8 years of experience, AWS architecture expertise, Event-driven system design, Kubernetes management, Distributed ML training, Terraform for infrastructure as code, Observability standards, Architectural judgment, Experience in ML-heavy environments, Containerized product delivery, Governance in adtech integrations

Chalice AI

Senior Site Reliability Engineer, DGX Cloud

2 days ago
$168,000 to $270,250 per yearopen to United StatesSeniorFull-time

Python · Kubernetes

8 years of experience, Kubernetes administration, GPU workloads optimization, Infrastructure automation (Terraform, Ansible), High-level programming (Python, Go), Linux operating systems, SRE principles (SLOs, SLIs), Observability stacks (OpenTelemetry, Prometheus), GPU-accelerated clusters (KubeVirt), AI inference workloads (vLLM, PyTorch)

NVIDIANVIDIA Corporation is a Santa Clara-based technology company specializing in designing GPUs and AI solutions for gaming, professional visualization, and cloud services, operating in both B2B and B2C markets globally.

Site Reliability Engineer Technical Lead

2 days ago
$100,000 to $150,000 per yearopen to United StatesLeadFull-time

8 years of experience, Deep SRE and Systems Expertise, Automation and Tooling, Observability and Analysis, Exceptional leadership and communication skills

Bright Vision TechnologiesBright Vision Technologies is a New Jersey-based IT staffing firm specializing in placing technical professionals in software development roles across the U.S. government and enterprise sectors.

Systems Observability Specialist

2 days ago
$135,000 to $155,000 per yearopen to United StatesFull-time

6 years of experience, Prometheus, Grafana, Commercial observability platforms, OpenTelemetry, Distributed tracing, Structured logging, High-cardinality metrics, SLOs, SRE principles, CI/CD integration, Linux internals, Networking, Container platforms

Bright Vision TechnologiesBright Vision Technologies is a New Jersey-based IT staffing firm specializing in placing technical professionals in software development roles across the U.S. government and enterprise sectors.

Site Reliability Engineer, Tech Lead

3 days ago
salary not disclosedopen to BrazilLeadFull-time

Python · AWS · Kubernetes

5 years of experience, AWS, Kubernetes, Cloud Computing, SRE/DevOps, Reliability Engineering leadership, CI/CD pipelines, UNIX/Linux, Networking, Automation tools, Python scripting, Monitoring and incident management

LoadsmartLoadsmart is a Chicago-based logistics technology company specializing in innovative freight management solutions for the B2B market.

Incident Operations Lead (EMEA/AMER)

4 days ago
salary not disclosedopen to United States + Canada + United Kingdom + Brazil + Japan + NigeriaLeadFull-time

5 years of experience, Incident command experience, FinTech understanding, 24x7 team leadership, Reliability metrics development, AI automation for incident management, Distributed team management, Severity model ownership, Effective communication under pressure

AlpacaAlpaca is a US-based fintech company providing self-clearing brokerage infrastructure and APIs for stocks, ETFs, options, and crypto, serving financial institutions globally.

Principal Infrastructure Engineer

4 days ago
$150,000 to $249,600 per yearopen to United StatesPrincipalFull-time

Python · Go · AWS

12 years of experience, AWS, Kubernetes, Aurora RDS (MySQL/Postgres), Infrastructure-as-code, Golang, Python, AI-assisted tooling, Disaster recovery, High-availability platform, Observability systems, Fintech experience

SezzleSezzle is a Minneapolis-based fintech B2C company specializing in Buy Now, Pay Later (BNPL) solutions, aiming to enhance financial inclusion for consumers across the U.S. market.

Principal Infrastructure Engineer

4 days ago
$150,000 to $249,600 per yearopen to United StatesPrincipalFull-time

Python · Go · AWS

12 years of experience, AWS, Kubernetes, Aurora RDS (MySQL/Postgres), Infrastructure-as-code, Golang, Python, AI-assisted tooling, Disaster recovery, High-availability platform, Observability systems, Fintech experience

SezzleSezzle is a Minneapolis-based fintech B2C company specializing in Buy Now, Pay Later (BNPL) solutions, aiming to enhance financial inclusion for consumers across the U.S. market.

Senior Site Reliability Engineer - FedRAMP

5 days ago
salary not disclosedopen to United StatesSeniorFull-time

Python · Azure · Kubernetes

8 years of experience, Azure, Datadog, SLIs, SLOs, Incident response, Kubernetes, Terraform, PowerShell, Python, Networking fundamentals, Observability platform management, FedRAMP compliance, Infrastructure as code, CI/CD pipelines, Web application firewall administration, Disaster recovery strategies, Clear written communication

DelineaDelinea is a Redwood City-based B2B AI-driven identity security platform specializing in Privileged Access Management (PAM) for enterprises across cloud and hybrid infrastructures.

Incident Operations Commander

5 days ago
salary not disclosedopen to United States + Canada + United Kingdom + Brazil + Japan + NigeriaFull-time

4 years of experience, Incident command, High-severity incident management, Technical operations, FinTech understanding, AI tools usage, Cross-region handoffs, Clear communication, Accountability under pressure

AlpacaAlpaca is a US-based fintech company providing self-clearing brokerage infrastructure and APIs for stocks, ETFs, options, and crypto, serving financial institutions globally.

Site Reliability Engineer (Principal Systems)

6 days ago
salary not disclosedopen to SingaporePrincipalFull-time

What You Bring to the Role; Experience with ServiceNow and ITIL principles.; This position will be 50% on site and 35% travel as required

Harris Computer

Support Engineer

6 days ago
salary not disclosedopen to UkraineFull-time

Python · AWS · Kubernetes

0.5 years of experience, Linux, Git, sh/bash, Diagnostic tools, Monitoring systems (Prometheus, Grafana), Configuration management (Ansible, Terraform), Scripting (Python), Docker, Kubernetes, CI/CD (Gitlab CI), Cloud (AWS), Relational databases (PostgreSQL)

KyivstarKyivstar is a Kyiv-based B2C telecommunications provider, offering mobile, broadband, and digital services, with a strong presence in Ukraine's telecom and digital media industries.

VMware Infrastructure Engineer

6 days ago
$100,000 to $150,000 per yearopen to United StatesFull-time

AWS · Kubernetes

6 years of experience, vSphere, vSAN, NSX-T, PowerCLI, Tanzu Kubernetes Grid, disaster recovery patterns, VMware Cloud on AWS, VMware Certified Professional (VCP), Aria Operations

Bright Vision TechnologiesBright Vision Technologies is a New Jersey-based IT staffing firm specializing in placing technical professionals in software development roles across the U.S. government and enterprise sectors.

Site Reliability Engineer, Infrastructure Platforms — UK (Intermediate to Senior Staff)

7 days ago
salary not disclosedopen to United KingdomStaffFull-time

AWS · GCP · Kubernetes

Kubernetes, Terraform, Go, AWS, GCP, Infrastructure as Code, Observability practices, Automation, Incident response, Strong written communication

GitLabGitLab is a San Francisco-based DevOps platform offering B2B and B2C solutions for software development, security, and collaboration, with a global presence.

Senior Software Engineer - SRE

7 days ago
$200,700 to $250,900 per yearopen to United States + CanadaSeniorFull-time

Site Reliability Engineering experience, PostgreSQL, Temporal workflows, Grafana or Honeycomb, OpenTelemetry

MercuryMercury Bank is a B2B financial institution offering simplified business banking services, headquartered in an unspecified location, targeting businesses primarily in the North American market.

Staff Security Engineer - PAM and Agentic Identity

7 days ago
$168,000 to $270,250 per yearopen to United StatesStaffFull-time

Kubernetes

8 years of experience, PAM expertise, NHI solutions, Linux/Windows environments, Kubernetes, CI/CD systems, Idira/CyberArk, HashiCorp Vault, Operational tooling, Short-lived credentials, Identity-aware proxies, Session recording

NVIDIANVIDIA Corporation is a Santa Clara-based technology company specializing in designing GPUs and AI solutions for gaming, professional visualization, and cloud services, operating in both B2B and B2C markets globally.

Senior Storage Production Engineer - DGX Cloud

7 days ago
salary not disclosedopen to OceaniaSeniorFull-time

Kubernetes

8 years of experience, Distributed storage solutions, High-performance storage systems, Storage networking protocols, Linux-based storage automation, Infrastructure configuration management, Observability tools, Capacity planning, Disaster recovery strategies, Kubernetes storage solutions, Replication strategies

NVIDIANVIDIA Corporation is a Santa Clara-based technology company specializing in designing GPUs and AI solutions for gaming, professional visualization, and cloud services, operating in both B2B and B2C markets globally.

Senior Infrastructure Engineer - Infrastructure Security and Core Services

8 days ago
$208,000 to $333,500 per yearopen to United StatesSeniorFull-time

12 years of experience, Infrastructure security, Large scale hybrid networks, Automation pipelines, Mellanox networking, Compute technologies, Storage architectures, Data lakes, Test Automation infrastructures, CPU and GPU workloads

NVIDIANVIDIA Corporation is a Santa Clara-based technology company specializing in designing GPUs and AI solutions for gaming, professional visualization, and cloud services, operating in both B2B and B2C markets globally.

Senior GCP Cloud Infrastructure Engineer

8 days ago
salary not disclosedopen to United StatesSeniorContract

GCP

8 years of experience, GCP landing zones, FedRAMP compliance, IAM models with external IdP, Assured Workloads provisioning, Infrastructure-as-Code, VPC/VPN configuration, Cloud Operations Suite, API Gateway implementation

Ontrac SolutionsOntrac Solutions is a technology consulting firm specializing in AI-driven systems and digital transformation, headquartered remotely with a focus on B2B services for healthcare, retail, and wellness industries.

Kubernetes & OpenShift Engineer

8 days ago
$135,000 to $155,000 per yearopen to United StatesFull-time

Python · Kubernetes

6 years of experience, Kubernetes internals, OpenShift internals, Linux administration, Infrastructure-as-code (Ansible, Terraform, Helm), CI/CD pipelines (Tekton, Jenkins, Argo CD), Scripting (Bash, Python, Go), Cluster monitoring and logging, Container image security

Bright Vision TechnologiesBright Vision Technologies is a New Jersey-based IT staffing firm specializing in placing technical professionals in software development roles across the U.S. government and enterprise sectors.

Platform Networking Engineer

8 days ago
$140,000 to $155,000 per yearopen to United StatesFull-time

Python · Kubernetes

6 years of experience, Istio, Linkerd, Envoy, Kubernetes, mTLS, PKI, Go, Python, Distributed tracing, Service mesh production experience

Bright Vision TechnologiesBright Vision Technologies is a New Jersey-based IT staffing firm specializing in placing technical professionals in software development roles across the U.S. government and enterprise sectors.

Platform Engineer (remote work)

8 days ago
salary not disclosedopen to WorldwideFull-time

Kubernetes

Linux systems administration, Kubernetes in production, Infrastructure as code (Ansible, Terraform), GitLab administration, Prometheus and Grafana, Technical writing for engineers, Strong communication skills, Advanced use of AI engineering assistants

CloudlinuxCloudLinux is a B2B commercial Linux distribution provider headquartered in an unspecified location, specializing in server stability and security solutions for the web hosting industry, serving shared hosting and VPS providers globally.

Senior Site Reliability Engineer

9 days ago
$191,000 to $226,000 per yearopen to United StatesSeniorFull-time

Python · AWS · Kubernetes

4 years of experience, Kubernetes, Terraform, AWS, Production observability, Python, Go, AI tools, HIPAA compliance, Cloud cost-efficiency

Garner HealthGarner Health is a U.S.-based healthcare technology B2B company offering a doctor quality analytics platform through a mobile app and concierge service, primarily targeting employers to enhance employee healthcare benefits.

Senior Staff Site Reliability Engineer

13 days ago
salary not disclosedopen to IndiaStaffFull-time

Python · Java · AWS

10 years of experience, Incident Commander experience, Distributed systems expertise, Kubernetes and cloud-native infrastructure, Automation for incident management, AI/ML applied to operations, Infrastructure-as-code tooling, Observability tooling expertise, Strong programming skills (Python, Go, Java), Experience with public cloud platforms (AWS, Azure, GCP), Scaling reliability across distributed teams

NVIDIANVIDIA Corporation is a Santa Clara-based technology company specializing in designing GPUs and AI solutions for gaming, professional visualization, and cloud services, operating in both B2B and B2C markets globally.

Senior Site Reliability Engineer - HPC

13 days ago
$152,000 to $287,500 per yearopen to United StatesSeniorFull-time

Python · Ruby · Kubernetes

5 years of experience, HPC cluster support, Slurm or LSF or Kubernetes, Infrastructure as Code (IaC), CI/CD techniques, Automated host lifecycle management, E2E observability, Python or Go or Perl or Ruby, Technical mentoring, Published technical write-ups, Open source component maintenance

NVIDIANVIDIA Corporation is a Santa Clara-based technology company specializing in designing GPUs and AI solutions for gaming, professional visualization, and cloud services, operating in both B2B and B2C markets globally.

Infiniband Network Engineer

13 days ago
€50,250 to €114,400 per yearopen to ItalyFull-time

5 years of experience, Infiniband Network troubleshooting, Infiniband Network certifications, Linux background, Ethernet networking

NVIDIANVIDIA Corporation is a Santa Clara-based technology company specializing in designing GPUs and AI solutions for gaming, professional visualization, and cloud services, operating in both B2B and B2C markets globally.

DevOps & SRE Engineer

13 days ago
$100,000 to $150,000 per yearopen to United StatesFull-time

Python · Java · AWS

6 years of experience, Python, Go, Java, Kubernetes, Linux at scale, Prometheus, Grafana, CI/CD pipelines, SLOs, Chaos engineering, AWS, Azure, GCP

Bright Vision TechnologiesBright Vision Technologies is a New Jersey-based IT staffing firm specializing in placing technical professionals in software development roles across the U.S. government and enterprise sectors.

Telemetry Engineer

13 days ago
$100,000 to $150,000 per yearopen to United StatesFull-time

Python · Java

6 years of experience, Prometheus, Grafana, OpenTelemetry, Datadog, New Relic, Splunk, Go, Python, Java, SRE principles, distributed tracing, structured logging, Linux, containers, CI/CD

Bright Vision TechnologiesBright Vision Technologies is a New Jersey-based IT staffing firm specializing in placing technical professionals in software development roles across the U.S. government and enterprise sectors.

Platform Engineer

14 days ago
$123,250 to $166,750 per yearopen to United StatesFull-time

AWS · Kubernetes

Kubernetes, Bare-metal infrastructure, Infrastructure as code (Terraform), CI/CD (GitLab CI), Linux/Unix administration, AWS (EC2, EKS, IAM), GPU-enabled infrastructure, YAML for Kubernetes manifests, Security clearance eligibility

Defense UnicornsDefense Unicorns is a USA-based B2B company providing secure, open source software solutions for the defense sector, specializing in DevSecOps and serving U.S. government agencies.

Kubernetes Service Engineer

14 days ago
$100,000 to $150,000 per yearopen to United StatesFull-time

Python · Kubernetes

6 years of experience, Istio, Linkerd, Envoy, mTLS, Kubernetes, Go, Python, Distributed tracing, Traffic management policies, Service mesh architecture

Bright Vision TechnologiesBright Vision Technologies is a New Jersey-based IT staffing firm specializing in placing technical professionals in software development roles across the U.S. government and enterprise sectors.