CNX Logo

CNX

Senior Site Reliability Engineer

Posted 13 Days Ago
In-Office or Remote
Hiring Remotely in Direct, TX
Senior level
In-Office or Remote
Hiring Remotely in Direct, TX
Senior level
Build, transition, operate, and improve a multi-cloud production platform across Azure, GCP, AWS, and Kubernetes. Responsibilities include infrastructure as code, GitOps and CI/CD, managed Kubernetes operations, troubleshooting, incident response, observability, reliability improvements, disaster recovery, automation, and cost control. The role partners with architects and global teams to ensure platforms are secure, scalable, recoverable, and supportable.
The summary above was generated by AI

Job Title:

Senior Site Reliability Engineer

Job Description

We're Concentrix. The intelligent transformation partner. Solution-focused. Tech-powered. Intelligence-fueled.
The global technology and services leader that powers the world’s best brands, today and into the future. We’re solution-focused, tech-powered, intelligence-fueled. With unique data and insights, deep industry expertise, and advanced technology solutions, we’re the intelligent transformation partner that powers a world that works, helping companies become refreshingly simple to work, interact, and transact with. We shape new game-changing careers in over 70 countries, attracting the best talent.
In our Information Technology and Global Security team, you will deliver the latest technology infrastructure, transformative software solutions and industry-leading global security for our staff and clients. You will work with the best in the world to design, implement and strategize IT, security, application development, innovation, and solutions in today’s hyperconnected world. You will be part of the technology team that is core to our vision of develop, build and run the future of Integrated Services.
Our game-changers around the world have devoted their careers to ensuring every relationship is exceptional. And we’re proud to be recognized with awards such as "World's Best Workplaces," “Best Companies for Career Growth,” and “Best Company Culture,” year after year.
We embrace our game-changers with open arms, people from diverse backgrounds, who are curious and willing to learn. Your natural talent to help others and go beyond WOW for our customers will fit right in with what we do and who we are.
Join us and be part of this journey towards greater opportunities and brighter futures.

We’re looking for a hands-on Senior Site Reliability Engineer to help build, transition, and operate a multi-cloud production platform supporting GPO products. This role is ideal for someone who enjoys working across cloud infrastructure, Kubernetes, GitOps, reliability, and production improvement—bringing strong engineering fundamentals, sound operational judgment, and a passion for building resilient, scalable platforms.

Responsibilities

  • Partner with GPO architects and planners to ensure target platform designs are deployable, secure, observable, supportable, and recoverable in production.

  • Assess and document the current multi-cloud and Kubernetes estate across Azure, GCP, AWS, and GitOps environments, identifying dependencies, migration needs, risks, and operational gaps.

  • Build, maintain, and enhance cloud infrastructure and shared platform services using infrastructure as code, peer-reviewed workflows, and safe production change practices.

  • Operate and improve AKS, GKE, and EKS environments, including networking, identity, upgrades, scalability, availability, recovery, and shared platform capabilities.

  • Develop and support GitOps and CI/CD workflows that make platform changes repeatable, reviewable, observable, and easy to validate and roll back.

  • Troubleshoot production issues across cloud, network, Kubernetes, GitOps, database, and shared platform layers, while collaborating effectively across team boundaries.

  • Reduce operational toil through automation and continuously improve SLOs, alerts, dashboards, runbooks, disaster recovery procedures, and cost controls.

  • Complete all assigned, mandatory training within the timeframe provided.

  • Conduct and/or participate in regularly scheduled 1:1 meetings with your direct manager and/or direct reports.

Qualifications

  • Strong Linux and networking fundamentals, with practical understanding of how distributed production systems behave and fail.

  • Hands-on experience in at least one key area such as Azure, GCP, AWS, Kubernetes, infrastructure as code, GitOps, CI/CD, observability, or reliability engineering.

  • Ability to build, test, review, and maintain infrastructure automation using tools or languages such as Terraform/OpenTofu, Terragrunt, Ansible, Python, Go, shell scripting, or Kubernetes configuration.

  • Experience with Git-based workflows, peer review, CI validation, rollback planning, and safe production change management.

  • Demonstrated troubleshooting and incident response skills, with a structured, risk-aware approach to solving unfamiliar technical problems.

  • Familiarity with managed Kubernetes platforms such as AKS, GKE, or EKS, and exposure to shared platform services including secrets, certificates, private connectivity, messaging, caching, or observability tooling.

  • Experience supporting brownfield platform environments, staged migrations, and database technologies such as MongoDB, PostgreSQL, or MySQL is preferred.

  • Strong written and verbal English communication skills, with the ability to collaborate effectively across global teams and meet requirements for privileged production access.

#li-remote

Location:

BGR Work-at-Home

Language Requirements:

Time Type:

Full time

Similar Jobs

7 Days Ago
Easy Apply
Remote
USA
Easy Apply
191K-226K Annually
Senior level
191K-226K Annually
Senior level
Big Data • Healthtech • HR Tech • Machine Learning • Software • Telehealth • Big Data Analytics
Own the reliability, performance, resilience, observability, and security of AWS and Kubernetes infrastructure supporting products and AI/ML workloads. Define SLOs, lead incident response and root-cause analysis, build Terraform automation, optimize cloud costs, reduce operational toil, and establish deployment standards that help engineers ship reliably. Participate in on-call rotations and maintain HIPAA-compliant infrastructure.
Top Skills: AWSClaudeDatadogGitlabGoHipaaIstioKubernetesNatsPostgresPythonSoc 2TerraformTypescript
Yesterday
Easy Apply
Remote
United States
Easy Apply
Senior level
Senior level
Cloud • Security • Software • Cybersecurity • Automation
Build and operate reliable, scalable production infrastructure for GitLab’s user-facing services. Responsibilities include developing infrastructure automation and tooling, managing Kubernetes deployments, maintaining infrastructure as code, supporting CI/CD and GitOps, participating in on-call and incident response, improving observability and SLOs, troubleshooting production systems, and documenting operational practices. The role spans Intermediate through Senior Staff levels and requires strong software engineering, cloud, reliability, and asynchronous collaboration skills.
Top Skills: AlertingAWSCi/CdGCPGitopsGoInfrastructure As CodeKubernetesLoggingMetricsRubySlisSlosTerraform
Senior level
Agency • Cloud • Professional Services • Software
Improve AWS production infrastructure reliability, observability, performance, and operational maturity. Build Terraform infrastructure, enhance CI/CD, automate operational work, manage incident response and on-call operations, lead postmortems, improve application resilience, support capacity planning and database reliability, and collaborate on security hardening and compliance. Mentor engineers and promote reliability practices across the organization.
Top Skills: AWSCi/CdCircleCIDatadogGithub ActionsGitlab CiLinuxNew RelicPostgresRubyRuby On RailsSlisSlosTerraform

What you need to know about the Charlotte Tech Scene

Ranked among the hottest tech cities in 2024 by CompTIA, Charlotte is quickly cementing its place as a major U.S. tech hub. Home to more than 90,000 tech workers, the city’s ecosystem is primed for continued growth, fueled by billions in annual funding from heavyweights like Microsoft and RevTech Labs, which has created thousands of fintech jobs and made the city a go-to for tech pros looking for their next big opportunity.

Key Facts About Charlotte Tech

  • Number of Tech Workers: 90,859; 6.5% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Lowe’s, Bank of America, TIAA, Microsoft, Honeywell
  • Key Industries: Fintech, artificial intelligence, cybersecurity, cloud computing, e-commerce
  • Funding Landscape: $3.1 billion in venture capital funding in 2024 (CED)
  • Notable Investors: Microsoft, Google, Falfurrias Management Partners, RevTech Labs Foundation
  • Research Centers and Universities: University of North Carolina at Charlotte, Northeastern University, North Carolina Research Campus

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account