Work Whale
Find JobsFind WorkersPost a JobHire TalentPricingHow It WorksAbout
Log InSign Up Free
Back to jobs
Work Whale

Connecting great employers with talented remote workers worldwide.

For Workers

  • Browse Jobs
  • Create Profile
  • How It Works

For Employers

  • Post a Job
  • Hire Talent
  • How It Works

Company

  • About Us
  • Pricing
  • Contact
  • Privacy Policy
  • Terms of Service

© 2026 Work Whale. All rights reserved.

Built for remote work, made for everyone.

GR

Site Reliability / Cloud Platform Engineer

Global Recruitment and Consultancy OPC

Remote Posted Aug 23, 2026
Full TimeDevOps & Infrastructure

Job Description

We are looking for an experienced Site Reliability / Cloud Platform Engineer to support and improve the reliability, performance, and scalability of enterprise technology platforms. The role involves troubleshooting complex technical issues, performing root cause analysis, improving system reliability, and developing automation and monitoring solutions. You will work closely with engineering, infrastructure, and support teams to resolve critical issues and implement long-term improvements. Key Responsibilities Investigate and resolve complex production, infrastructure, and platform issues through detailed root cause analysis. Troubleshoot issues across Kubernetes, Linux, networking, cloud infrastructure, configurations, and underlying systems. Identify recurring incidents and system weaknesses and implement solutions to prevent future occurrences. Develop automation, scripts, tools, dashboards, and monitoring solutions to improve operational efficiency and supportability. Support and optimize Kubernetes and cloud-native environments , including service mesh technologies. Participate in critical incident response, outage investigations, and post-incident reviews. Analyze system metrics, logs, and performance data to identify reliability and availability risks. Provide technical recommendations for improving platform resilience, scalability, observability, and performance. Required Qualifications 4–8+ years of experience in Site Reliability Engineering, DevOps, Platform Engineering, Systems Engineering, or a related field. Experience supporting enterprise-scale platforms and production environments. Strong hands-on experience with Kubernetes and Linux. Strong understanding of networking and infrastructure. Experience with cloud-native technologies, automation, and observability. Experience or exposure to AI / Machine Learning technologies. Required Technical Skills: Kubernetes Linux Networking Service Mesh / Istio PKI / Public Key Infrastructure AI / Machine Learning Google Cloud Compute Prometheus Grafana Loki Splunk

Requirements

  • Communication — 1 year
  • Remote Work — 1 year
Competitive salary

Competitive compensation

Apply Now

Sign in to submit your application

About Global Recruitment and Consultancy OPC

GR

Global Recruitment and Consultancy OPC

CategoryDevOps & Infrastructure
TypeFull Time
LocationRemote
ExpiresNo expiry