Remote Site Reliability Engineer Jobs

Explore 50 fresh remote Site Reliability Engineer jobs. Whether you're working from home or from anywhere in the world, our curated listings deliver clear insights for your next move.

Filter by Location

Subscribe to our Telegram bot to receive instant notifications about new remote jobs

TelegramSubscribe Now

Latest Site Reliability Engineer Jobs (50)

AI Infrastructure & Platform Operations Engineer (remote in the EU)

1 day ago
Full-time
Germany, France, Netherlands, Ireland, Spain, Italy, Poland, Portugal, Belgium, Malta, Romania, Sweden, Finland, Denmark, Austria, Greece, Bulgaria, Lithuania, Czechia, Croatia
Key requirements: 3 years of experience, NVIDIA GPU infrastructure, Kubernetes, AI infrastructure, High-performance networking, Observability platforms, Infrastructure automation, Site Reliability Engineering (SRE)
Mirantis

Mirantis is a Campbell, California-based B2B company specializing in open source cloud computing and Kubernetes-native AI infrastructure, serving diverse sectors including automotive, healthcare, and financial services.

Remote policy: Mirantis supports flexible remote work, primarily within the EU, and offers the option to work from their Helsinki hub, fostering a remote-first culture.

NCX Senior Engineer

1 day ago
Full-time
United States
$184,000 to $356,500 per year
Key requirements: 8 years of experience, Kubernetes, AI/ML experience, Python, Go, NVIDIA ecosystem, MLOps, Infrastructure as code, Distributed computing, GPU scheduling, Observability stacks
NVIDIA

NVIDIA Corporation is a Santa Clara-based technology company specializing in designing GPUs and AI solutions for gaming, professional visualization, and cloud services, operating in both B2B and B2C markets globally.

Remote policy: NVIDIA supports flexible remote work arrangements and hires from various regions globally, including the Americas, Europe, Asia, and the Middle East, with roles that may require collaboration across time zones.

Staff Platform Engineer (IC-4)

2 days ago
Full-time
Brazil, Chile, Colombia
Key requirements: 5 years of experience, AWS, Datadog, CI/CD (GitHub Actions), Cloud networking, Security fundamentals, Incident response, Observability tooling, Production systems management, Local development tooling
Moxie

Moxie is a remote-first company empowering aesthetic entrepreneurs to build profitable practices, operating in the aesthetics industry with a global reach.

Remote policy: Moxie operates as a remote-first company, hiring from various regions globally, with team members supporting practices across multiple locations.

Staff Platform Engineer (IC-4)

2 days ago
Full-time
Brazil, Chile, Colombia
Key requirements: 5 years of experience, AWS, Datadog, CI/CD (GitHub Actions), Cloud networking, Production systems, Incident response, Observability tooling, Vercel, GCP, Cloudflare
Moxie

Moxie is a remote-first company empowering aesthetic entrepreneurs to build profitable practices, operating in the aesthetics industry with a global reach.

Remote policy: Moxie operates as a remote-first company, hiring from various regions globally, with team members supporting practices across multiple locations.

Senior HPC Cluster Administrator - Deep Learning Frameworks Infrastructure

2 days ago
Full-time
Poland
₱292,500 to ₱507,000 per year
Key requirements: 5 years of experience, HPC cluster administration, Deep learning frameworks, Linux systems administration, Slurm, Ansible, Python, High-speed networking, Distributed filesystems, NVIDIA GPU tools, MLOps tooling
NVIDIA

NVIDIA Corporation is a Santa Clara-based technology company specializing in designing GPUs and AI solutions for gaming, professional visualization, and cloud services, operating in both B2B and B2C markets globally.

Remote policy: NVIDIA supports flexible remote work arrangements and hires from various regions globally, including the Americas, Europe, Asia, and the Middle East, with roles that may require collaboration across time zones.

Senior Reliability Engineer

2 days ago
Full-time
United States
$168,000 to $264,500 per year
Key requirements: 8 years of experience, Wafer fab expertise, Flip Chip BGA packaging, COWOS assembly process, CMOS/Fin FET device physics, Reliability statistics, Failure mechanisms, Industry standards (JEDEC, IPC, AEC-Q100), ESD/LU testing procedures, Reliability data analysis (JMP, Weibull++, Minitab)
NVIDIA

NVIDIA Corporation is a Santa Clara-based technology company specializing in designing GPUs and AI solutions for gaming, professional visualization, and cloud services, operating in both B2B and B2C markets globally.

Remote policy: NVIDIA supports flexible remote work arrangements and hires from various regions globally, including the Americas, Europe, Asia, and the Middle East, with roles that may require collaboration across time zones.

Senior Platform Engineer

3 days ago
Contract
United States
Key requirements: Cloud infrastructure, Platform engineering, Distributed systems, IAM, Observability, Infrastructure automation, Terraform or OpenTofu, AWS or Azure or GCP or Kubernetes, Chaos engineering
CONFISA INTERNATIONAL GROUP

Confisa International Group is a Pembroke Pines, Florida-based professional services firm specializing in retained executive search and human resources solutions for B2B clients across multiple industries globally.

Remote policy: Confisa International Group supports remote work for certain roles, with current hiring focused on candidates located in the USA. The company has a global reach, operating across multiple continents including Europe, Africa, Asia, and the Americas.

AI Infrastructure & Platform Operations Engineer (remote in the EU)

3 days ago
Full-time
Europe
$60,000 to $67,000 per year
Key requirements: 3 years of experience, NVIDIA GPU infrastructure, Kubernetes platform operations, AI infrastructure or HPC environments, Infrastructure automation technologies, Large-scale distributed systems, Observability platforms
Mirantis

Mirantis is a Campbell, California-based B2B company specializing in open source cloud computing and Kubernetes-native AI infrastructure, serving diverse sectors including automotive, healthcare, and financial services.

Remote policy: Mirantis supports flexible remote work, primarily within the EU, and offers the option to work from their Helsinki hub, fostering a remote-first culture.

Senior Engineer, SRE

3 days ago
Full-time
Worldwide
$120,000 to $160,000 per year
Key requirements: Kubernetes, GCP, Cloud networking, Infrastructure-as-code, Distributed systems, CI/CD, Automated deployment, Operational reliability, Cost optimization, Debugging across systems
Zencoder

Zencoder is a B2B SaaS platform headquartered in the US, specializing in AI coding assistance to enhance developer productivity and integrate generative coding workflows.

Remote policy: Zencoder embraces a global and flexible remote work culture, hiring talent from various locations and allowing employees to work from wherever they are most productive.

Senior Engineer, Infrastructure

3 days ago
Full-time
Worldwide
$100,000 to $150,000 per year
Key requirements: GCP, AWS, Kubernetes, Infrastructure-as-code, Cloud networking, Cloud security, OpenSearch, Elasticsearch, PostgreSQL, CI/CD, Automated deployment
Zencoder

Zencoder is a B2B SaaS platform headquartered in the US, specializing in AI coding assistance to enhance developer productivity and integrate generative coding workflows.

Remote policy: Zencoder embraces a global and flexible remote work culture, hiring talent from various locations and allowing employees to work from wherever they are most productive.

Cloud Infrastructure Network Engineer

3 days ago
Full-time
United States
$100,000 to $150,000 per year
Key requirements: 6 years of experience, Cloud networking, VPC/VNet design, Hybrid connectivity, Infrastructure-as-code (Terraform), Cloud security, Kubernetes networking, Multi-cloud networking, SD-WAN familiarity, eBPF-based networking tools
Bright Vision Technologies

Bright Vision Technologies is a New Jersey-based IT staffing firm specializing in placing technical professionals in software development roles across the U.S. government and enterprise sectors.

Monitoring Engineer

3 days ago
Full-time
United States
$100,000 to $150,000 per year
Key requirements: 6 years of experience, Prometheus, Grafana, Datadog, New Relic, Splunk, OpenTelemetry, distributed tracing, structured logging, Go, Python, Java, high-cardinality metrics, SLOs, error budgets, SRE principles, CI/CD integration, Linux internals, networking, container platforms, Thanos, Mimir, Cortex, Loki, Tempo, eBPF, observability cost optimization, audit-grade logging
Bright Vision Technologies

Bright Vision Technologies is a New Jersey-based IT staffing firm specializing in placing technical professionals in software development roles across the U.S. government and enterprise sectors.

Reliability Engineer

3 days ago
Full-time
United States
$100,000 to $150,000 per year
Key requirements: 6 years of experience, Python, Go, Java, Kubernetes, Linux at scale, Observability tooling, CI/CD pipelines, Distributed system design, SLOs and error budgets, Chaos engineering, Cloud platforms (AWS, Azure, GCP)
Bright Vision Technologies

Bright Vision Technologies is a New Jersey-based IT staffing firm specializing in placing technical professionals in software development roles across the U.S. government and enterprise sectors.

Site Reliability Engineer (SRE)

3 days ago
Full-time
United States
$100,000 to $180,000 per year
Key requirements: 6 years of experience, Kubernetes, Python, Go, Prometheus, Grafana, CI/CD pipelines, Chaos engineering, SLOs and error budgets, Linux at scale, Observability tooling
Bright Vision Technologies

Bright Vision Technologies is a New Jersey-based IT staffing firm specializing in placing technical professionals in software development roles across the U.S. government and enterprise sectors.

Cloud Systems Engineer II

4 days ago
Full-time
India
Key requirements: 5 years of experience, AWS, GCP, Terraform, Python, Linux administration, Networking fundamentals, CI/CD principles, High availability architecture, Mentoring, Technical documentation
taketwo

Take-Two Interactive Software, Inc. is a New York City-based B2C video game developer and publisher, known for franchises like Grand Theft Auto and NBA 2K, serving global markets across console, PC, and mobile platforms.

Senior Debug System Engineer, Datacenter

4 days ago
Full-time
United States
$168,000 to $258,750 per year
Key requirements: 8 years of experience, Failure analysis on GPU products, Debugging motherboards and servers, DFx enabling knowledge, Hardware and Software debugging, Familiarity with oscilloscopes and analyzers, Strong problem-solving skills, Effective communication with partners
NVIDIA

NVIDIA Corporation is a Santa Clara-based technology company specializing in designing GPUs and AI solutions for gaming, professional visualization, and cloud services, operating in both B2B and B2C markets globally.

Remote policy: NVIDIA supports flexible remote work arrangements and hires from various regions globally, including the Americas, Europe, Asia, and the Middle East, with roles that may require collaboration across time zones.

Observability Engineer

4 days ago
Full-time
United States
$100,000 to $160,000 per year
Key requirements: 12 years of experience, Prometheus, Grafana, Datadog, OpenTelemetry, Distributed tracing, High-cardinality metrics, SLOs, CI/CD integration, Linux internals, eBPF-based observability
Bright Vision Technologies

Bright Vision Technologies is a New Jersey-based IT staffing firm specializing in placing technical professionals in software development roles across the U.S. government and enterprise sectors.

OpenShift Platform Engineer

4 days ago
Full-time
United States
$100,000 to $150,000 per year
Key requirements: 6 years of experience, OpenShift, Kubernetes internals, Linux administration, Infrastructure-as-code (Ansible, Terraform, Helm), CI/CD pipelines (Tekton, Jenkins, Argo CD), Scripting (Bash, Python, Go), Cluster monitoring and logging, Container image security, Disaster recovery strategies, GitOps workflows (Argo CD, Flux)
Bright Vision Technologies

Bright Vision Technologies is a New Jersey-based IT staffing firm specializing in placing technical professionals in software development roles across the U.S. government and enterprise sectors.

Cloud Operations Manager

5 days ago
Full-time
Worldwide
Key requirements: 5 years of experience, AWS (EC2, RDS, S3, ECS, Fargate, IAM), CI/CD tools (GitHub Actions, Jenkins, GitLab CI, CircleCI), People management, Operational process implementation, Cloud security best practices
Panoptyc

Panoptyc is a New York-based B2B startup specializing in AI and computer vision technology for retail loss prevention, providing automated security systems to retail operators and micro markets.

Remote policy: Panoptyc is a fully distributed company, hiring remotely from various locations, with team members working from different regions including the United States.

Software Engineer III, Site Reliability

6 days ago
Full-time
United States
$120,000 to $165,000 per year
Key requirements: 5 years of experience, SLI/SLO frameworks, Incident response leadership, Datadog, Infrastructure as Code (Terraform), Kubernetes management, CI/CD pipeline ownership, SAST/DAST/SCA integration, Cloud security fundamentals, Policy-as-code frameworks, B2C/mobile backend experience
MyFitnessPal

MyFitnessPal is a health and fitness tracking mobile application (SaaS) based in the health and wellness technology industry, serving individual users globally with tools for nutrition and fitness management.

Site Reliability Engineer (US - Central/Eastern time)

6 days ago
Full-time
United States
$100,000 to $150,000 per year
Key requirements: Kubernetes (EKS), AWS multi-account management, Terraform/Terragrunt automation, Linux systems, Stateful systems support, Performance debugging, End-to-end system ownership
PostHog

PostHog is a San Francisco-based B2B SaaS platform offering an integrated suite of tools for product engineers to build, test, and analyze software products, with a focus on the global market.

Remote policy: PostHog is a fully remote company with a globally distributed team, currently hiring in time zones between GMT-8 and GMT+2.

Site Reliability Engineer - ClickHouse

6 days ago
Full-time
Europe
$100,000 to $150,000 per year
Key requirements: ClickHouse, AWS, VM-based systems, Infrastructure automation, Linux systems, Stateful systems support, Performance debugging, End-to-end system ownership
PostHog

PostHog is a San Francisco-based B2B SaaS platform offering an integrated suite of tools for product engineers to build, test, and analyze software products, with a focus on the global market.

Remote policy: PostHog is a fully remote company with a globally distributed team, currently hiring in time zones between GMT-8 and GMT+2.

Staff Platform Engineer

6 days ago
Full-time
Worldwide
$190,000 to $230,000 per year
Key requirements: 5 years of experience, Rust, Python, Kubernetes, AWS, Terraform, GitOps, Incident response, Data systems fluency, Agent fluency, Guardrail engineering
Postscript

Postscript is a remote-based SaaS company specializing in SMS marketing solutions for eCommerce brands, targeting the B2B market with a focus on enhancing customer engagement and driving sales.

Remote policy: Postscript is a fully remote organization, hiring from various locations globally, allowing team members to work from anywhere.

Senior / Expert Engineer, Site Reliability Engineering (Garena)

7 days ago
Full-time
Asia
Key requirements: 3 years of experience, Linux, Kubernetes, TCP/IP, Bash, Python, Go, Analytical skills, Problem-solving skills
Garena

Garena is a Singapore-based B2C gaming and esports company specializing in free-to-play games across mobile and PC platforms, with a strong presence in Southeast Asia and Taiwan.

Managed Services Engineer I

7 days ago
Full-time
Worldwide
Key requirements: 1 years of experience, Microsoft Exchange, SQL, Windows Server, ConnectWise, LAN/WAN troubleshooting, Microsoft Active Directory, Azure AD, Basic PowerShell, Strong organization skills, Strong interpersonal skills
Logically

Logically is a Brighouse, England-based B2B Managed Security Services Provider (MSSP) specializing in cybersecurity solutions and IT services for organizations across various industries.

Remote policy: Logically supports remote work and hires from various regions, including the Raleigh–Durham–Chapel Hill area in North Carolina, while encouraging a collaborative team environment.

AWS Cloud Engineer

7 days ago
Full-time
North America
Key requirements: 5 years of experience, AWS core services, Multi-cloud transformation, Basic Azure and Oracle Cloud exposure, Windows Server, IIS, DNS, Active Directory, MS SQL Server, Basic Linux administration, Docker, Shell scripting, PowerShell scripting, Troubleshooting skills, Information security best practices, US time zone flexibility
PairSoft

PairSoft is a Miami-based fintech company offering B2B SaaS solutions for procure-to-pay automation, targeting mid-market and enterprise organizations across North America and Europe.

Platform Engineer II

8 days ago
Full-time
United States
$118,000 to $151,000 per year
Key requirements: AWS, Docker, Kubernetes, Python, CI/CD, Terraform, Observability, Security best practices, Production systems operation
Plüm énergie

Plüm énergie is a French B2C startup supplying green electricity with a focus on energy reduction and customer savings, operating in the renewable energy sector.

Senior Cloud Operations Engineer (Hybrid) Bellevue, WA; Newtown Square, PA; New York, New York, or San Francisco, CA

8 days ago
Full-time
United States
Key requirements: 4 years of experience, Azure administration, Azure DevOps, Identity management, Zero Trust architecture, Scripting/automation, Monitoring tools, Cost optimization, Cloud-native application design, CI/CD pipeline management, Disaster recovery strategies
Arctiq

Arctiq is a Toronto-based B2B DevOps and cloud solution integrator specializing in professional IT services and managed services for enterprise organizations across North America.

Staff Infrastructure Engineer – Kubernetes Platform

9 days ago
Full-time
United States
$120,000 to $160,000 per year
Key requirements: 7 years of experience, Kubernetes at scale, Multi-tenant cluster models, Control plane architecture design, CNI plugins (Cilium preferred), Multi-region scaling, Deep troubleshooting across Kubernetes, Networking stack expertise, Observability platforms (Prometheus, Grafana), Virtual cluster technologies (vcluster, Kamaji)
TensorWave

TensorWave is a Las Vegas-based B2B cloud computing provider specializing in AI and high-performance computing infrastructure, utilizing AMD Instinct GPUs to deliver scalable solutions for enterprises and AI researchers.

Site Reliability Engineer

9 days ago
Full-time
Europe
€58,000 to €97,000 per year
Key requirements: Kubernetes, AWS, GCP, Python, C++, ClickHouse, SQL, Observability, Distributed systems, On-call participation, EU timezone
Tinybird

Tinybird is a serverless real-time analytics platform headquartered in Madrid, specializing in B2B solutions for developers and data teams to build scalable analytics applications using ClickHouse.

Remote policy: Tinybird is a fully remote company with a remote-first culture, hiring from various locations including Europe and the Americas, with team members in cities such as Madrid and New York City.