All Remote Jobs
Loading the current view.
The page is checking the latest available data.
All Remote Jobs
The page is checking the latest available data.
Concentrix
Bulgaria · remote · Salary not listed
We’re looking for a hands-on Senior Site Reliability Engineer to help build, transition, and operate a multi-cloud production platform supporting GPO products. This role is ideal for someone who enjoys working across cloud infrastructure, Kubernetes, GitOps, reliability, and production improvement—bringing strong engineering fundamentals, sound operational judgment, and a passion for building resilient, scalable platforms.
Partner with GPO architects and planners to ensure target platform designs are deployable, secure, observable, supportable, and recoverable in production.
Assess and document the current multi-cloud and Kubernetes estate across Azure, GCP, AWS, and GitOps environments, identifying dependencies, migration needs, risks, and operational gaps.
Build, maintain, and enhance cloud infrastructure and shared platform services using infrastructure as code, peer-reviewed workflows, and safe production change practices.
Apply on source siteListed via Himalayas
Operate and improve AKS, GKE, and EKS environments, including networking, identity, upgrades, scalability, availability, recovery, and shared platform capabilities.
Develop and support GitOps and CI/CD workflows that make platform changes repeatable, reviewable, observable, and easy to validate and roll back.
Troubleshoot production issues across cloud, network, Kubernetes, GitOps, database, and shared platform layers, while collaborating effectively across team boundaries.
Reduce operational toil through automation and continuously improve SLOs, alerts, dashboards, runbooks, disaster recovery procedures, and cost controls.
Complete all assigned, mandatory training within the timeframe provided.
Conduct and/or participate in regularly scheduled 1:1 meetings with your direct manager and/or direct reports.
Strong Linux and networking fundamentals, with practical understanding of how distributed production systems behave and fail.
Hands-on experience in at least one key area such as Azure, GCP, AWS, Kubernetes, infrastructure as code, GitOps, CI/CD, observability, or reliability engineering.
Ability to build, test, review, and maintain infrastructure automation using tools or languages such as Terraform/OpenTofu, Terragrunt, Ansible, Python, Go, shell scripting, or Kubernetes configuration.
Experience with Git-based workflows, peer review, CI validation, rollback planning, and safe production change management.
Demonstrated troubleshooting and incident response skills, with a structured, risk-aware approach to solving unfamiliar technical problems.
Familiarity with managed Kubernetes platforms such as AKS, GKE, or EKS, and exposure to shared platform services including secrets, certificates, private connectivity, messaging, caching, or observability tooling.
Experience supporting brownfield platform environments, staged migrations, and database technologies such as MongoDB, PostgreSQL, or MySQL is preferred.
Strong written and verbal English communication skills, with the ability to collaborate effectively across global teams and meet requirements for privileged production access.
Originally posted on Himalayas