Remote Site Reliability Engineer jobs open to India.

India? Yes. $$$? When they say. Stack? Before the essay.

11 openings. None older than 30 days.

Newest first

Senior Staff SRE – Compute Platform

3 days ago
salary not disclosedopen to IndiaStaffFull-time

Python · Kubernetes

10 years of experience, Kubernetes administration, Bare-metal infrastructure, Python or Go, Infrastructure as Code, SRE and observability, HPC or AI infrastructure, VMware vSphere, Generative AI applications, Secure operational platforms, High-impact infrastructure projects

NVIDIANVIDIA Corporation is a Santa Clara-based technology company specializing in designing GPUs and AI solutions for gaming, professional visualization, and cloud services, operating in both B2B and B2C markets globally.

Site Reliability Engineer

9 days ago
salary not disclosedopen to IndiaFull-time

AWS · Azure · GCP

Kubernetes, Cloud platforms (AWS, Azure, GCP), Infrastructure-as-code (Terraform, AWS CDK), Observability tools (OpenTelemetry, Prometheus, Grafana), CI/CD pipelines, AI/ML exposure, Scripting for automation, Container orchestration

NVIDIANVIDIA Corporation is a Santa Clara-based technology company specializing in designing GPUs and AI solutions for gaming, professional visualization, and cloud services, operating in both B2B and B2C markets globally.

Senior Staff Site Reliability Engineer

9 days ago
salary not disclosedopen to IndiaStaffFull-time

Kubernetes

8 years of experience, Kubernetes-based platforms, AI inference services, Control planes, Platform APIs, GPU scheduling, Vector databases, GitOps, CI/CD, Observability tools, Self-service platforms, Inference-serving frameworks, Open-source contributions

NVIDIANVIDIA Corporation is a Santa Clara-based technology company specializing in designing GPUs and AI solutions for gaming, professional visualization, and cloud services, operating in both B2B and B2C markets globally.

Senior Site Reliability Engineer

12 days ago
salary not disclosedopen to IndiaSeniorFull-time

TypeScript · Python · Kubernetes

5 years of experience, Kubernetes (EKS, GKE), Terraform, Pulumi, GitOps (ArgoCD, Flux), Datadog, Prometheus, Grafana, SLI/SLO/error budget fluency, Go, Python, TypeScript, Linux internals, TCP/IP networking, Incident response leadership, Live-service game experience, Service mesh (Istio, Cilium), FinOps, Cloud certifications

2k2K is a Novato, California-based video game publisher specializing in a diverse range of genres including sports, action, and role-playing, primarily operating in the B2C market.

Site Reliability Engineer

12 days ago
salary not disclosedopen to WorldwideFull-time

Python · Azure · Kubernetes

3 years of experience, SLIs/SLOs definition, Multi-tenant SaaS platforms, Datadog, Grafana, Elastic Stack, High-availability architectures, Kubernetes, Python, Bash, Incident response, Cloud experience (Azure)

HostPapaHostPapa is a Canadian-based web hosting company offering B2B and B2C solutions, including shared, reseller, and VPS hosting services, with a focus on small businesses and a global presence.

Staff Platform Engineer, Core Cloud Platform

15 days ago
$223,100 to $305,000 per yearopen to India + United States + Canada + United Kingdom + Singapore + Ireland + FinlandStaffFull-time

Python · Go · AWS

Kubernetes expertise, AWS cloud proficiency, Production SaaS systems experience, Networking and service mesh knowledge, Operational troubleshooting skills, Python and Golang proficiency, CI/CD practices advocacy, Technical roadmap design and leadership

AlphaSenseAlphaSense is a New York City-based B2B fintech platform specializing in AI-driven market intelligence and search solutions for financial institutions and top companies globally.

Senior Staff Site Reliability Engineer

17 days ago
$150,000 to $200,000 per yearopen to IndiaStaffFull-time

Python · AWS · Kubernetes

10 years of experience, Kubernetes, AI/ML infrastructure, OpenTelemetry, AWS CDK, Terraform, Python, Go, Distributed systems debugging, Agentic AI platforms, Automation systems, Technical strategy steering

NVIDIANVIDIA Corporation is a Santa Clara-based technology company specializing in designing GPUs and AI solutions for gaming, professional visualization, and cloud services, operating in both B2B and B2C markets globally.

Site Reliability Engineer

19 days ago
salary not disclosedopen to WorldwideFull-time

AWS · Kubernetes

7 years of experience, Event-driven architecture, Deep AWS, Infrastructure as Code, Kubernetes, Observability and SLOs, Chaos engineering, Distributed systems debugging, Proven technical leadership, AI / MLOps infrastructure, Experience in payments industry

YunoYuno is a global payment orchestration platform headquartered in [location], specializing in B2B payment infrastructure that enables businesses to integrate over 1,000 payment methods through a single API.

Principal Site Reliability Engineer

22 days ago
salary not disclosedopen to IndiaPrincipalFull-time

Python

10 years of experience, Openshift, Nutanix AHV, VMware vSphere, RedHat OpenShift, Python, Go, Ansible, GitHub, 24x7 operations

DellDell Technologies designs and develops hardware and software for infrastructure, enterprise solutions, and data management.

Senior HPC Cluster Engineer - AI, ML

23 days ago
salary not disclosedopen to IndiaSeniorFull-time

Python · Docker

5 years of experience, AI/HPC advanced job schedulers, Slurm, Centos/RHEL, Ubuntu Linux, Cluster configuration management, Docker, Python, MPI, NVIDIA GPUs, CUDA Programming, InfiniBand, Lustre

NVIDIANVIDIA Corporation is a Santa Clara-based technology company specializing in designing GPUs and AI solutions for gaming, professional visualization, and cloud services, operating in both B2B and B2C markets globally.

Operations Support Engineer

24 days ago
salary not disclosedopen to WorldwideFull-time

5 years of experience, ITIL-aligned frameworks, Second-level support for mission-critical systems, Observability tools, Linux/Unix environments, Container orchestration, CI/CD pipelines, Middleware and integration technologies, Database performance monitoring, System hardening and patching, English fluency

EUROPEAN DYNAMICSEuropean Dynamics is a Greece-based B2G SaaS provider specializing in ICT services and software development for e-government, operating internationally across 47 countries.