Remote Site Reliability Engineer jobs open to Asia.
Asia? Yes. $$$? When they say. Stack? Before the essay.
Can you be more specific?
16 openings. None older than 30 days.
Newest first
Senior Staff Site Reliability Engineer
7 days agoPython · Java · AWS
10 years of experience, Incident Commander experience, Distributed systems expertise, Kubernetes and cloud-native infrastructure, Automation for incident management, AI/ML applied to operations, Infrastructure-as-code tooling, Observability tooling expertise, Strong programming skills (Python, Go, Java), Experience with public cloud platforms (AWS, Azure, GCP), Scaling reliability across distributed teams
Software Systems Engineer III,
9 days agoPython · AWS · Azure
8 years of experience, Kubernetes, CI/CD, GitOps, Infrastructure-as-code, Azure, AWS, GCP, Argo CD, Helm, Terraform, Python, MLOps, Kubeflow, AI/ML workloads, Monitoring tools, SRE, Platform engineering
NCX Senior Engineer
9 days agoKubernetes
8 years of experience, NVIDIA technologies, Kubernetes, GPU infrastructure management, Infrastructure observability tools, Automation for lifecycle management, Linux-based distributed systems, Cloud infrastructure operations, Collaboration with cloud partners
Senior Staff Site Reliability Engineer
10 days agoPython
15 years of experience, eBPF, XDP, Containerization architectures, Distributed Systems Infrastructure, Terraform, Config Management tools, Go, Python, Linux Kernel Internals, Network Protocols (VLAN/VxLAN/SDN/BGP/Anycast), DNS, LDAP, Microservices architecture, Infrastructure as Code (IaC)
Datadog Administration and Operations (Servicenow)
10 days agoPython · SQL
5 years of experience, Datadog expertise, ServiceNow integration, Observability strategy design, Scripting (Python/PowerShell/Bash), Cloud cost optimization, CI/CD integration, Network performance monitoring, Database monitoring (Postgres/SQL Server/Oracle/MySQL), Automation of monitoring processes, ITIL v4 Foundation certification
L2 Technician - #35272
14 days agoAzure
2 years of experience, Microsoft technologies, Azure Support, Active Directory, Exchange, Microsoft 365, Windows Servers, Excellent troubleshooting skills, MSP experience
Senior Staff SRE – Compute Platform
16 days agoPython · Kubernetes
10 years of experience, Kubernetes administration, Bare-metal infrastructure, Python or Go, Infrastructure as Code, SRE and observability, HPC or AI infrastructure, VMware vSphere, Generative AI applications, Secure operational platforms, High-impact infrastructure projects
Senior R&D Site Reliability Engineer
18 days agoPython · AWS
Terraform, AWS architecture, CI/CD (Gitlab runner), Linux system management, Python scripting, Observability tools (Prometheus, Grafana, ELK Stack), Automation tools development, Analytical and problem-solving skills, English proficiency (IELTS 6.5), Mandarin Chinese (plus)
Site Reliability Engineer
22 days agoAWS · Azure · GCP
Kubernetes, Cloud platforms (AWS, Azure, GCP), Infrastructure-as-code (Terraform, AWS CDK), Observability tools (OpenTelemetry, Prometheus, Grafana), CI/CD pipelines, AI/ML exposure, Scripting for automation, Container orchestration
Senior Staff Site Reliability Engineer
22 days agoKubernetes
8 years of experience, Kubernetes-based platforms, AI inference services, Control planes, Platform APIs, GPU scheduling, Vector databases, GitOps, CI/CD, Observability tools, Self-service platforms, Inference-serving frameworks, Open-source contributions
Senior Site Reliability Engineer
25 days agoTypeScript · Python · Kubernetes
5 years of experience, Kubernetes (EKS, GKE), Terraform, Pulumi, GitOps (ArgoCD, Flux), Datadog, Prometheus, Grafana, SLI/SLO/error budget fluency, Go, Python, TypeScript, Linux internals, TCP/IP networking, Incident response leadership, Live-service game experience, Service mesh (Istio, Cilium), FinOps, Cloud certifications
Site Reliability Engineer
25 days agoPython · Azure · Kubernetes
3 years of experience, SLIs/SLOs definition, Multi-tenant SaaS platforms, Datadog, Grafana, Elastic Stack, High-availability architectures, Kubernetes, Python, Bash, Incident response, Cloud experience (Azure)
Senior Site Reliability Engineer in Test, SDET
25 days agoPython · Kubernetes
8 years of experience, GitLab CI, ArgoCD, Kubernetes, Linux, Python, SRE principles, SBOM tooling, AI/ML techniques, test environment management, chaos engineering
Staff Platform Engineer, Core Cloud Platform
28 days agoPython · Go · AWS
Kubernetes expertise, AWS cloud proficiency, Production SaaS systems experience, Networking and service mesh knowledge, Operational troubleshooting skills, Python and Golang proficiency, CI/CD practices advocacy, Technical roadmap design and leadership
WFH Sr. Network and Systems Administrator (Firewall) - #35234
29 days agoAzure
5 years of experience, Firewall management, Active Directory, Network security, VPN configuration, Enterprise-scale software deployments, Advanced PowerShell scripting, Hospitality industry experience, Microsoft Azure, Intrusion detection/prevention, Automation frameworks
Senior Staff Site Reliability Engineer
about 1 month agoPython · AWS · Kubernetes
10 years of experience, Kubernetes, AI/ML infrastructure, OpenTelemetry, AWS CDK, Terraform, Python, Go, Distributed systems debugging, Agentic AI platforms, Automation systems, Technical strategy steering