Lead day-to-day Redis and OpenShift platform operations, including cluster builds, maintenance, upgrades, monitoring, troubleshooting, incident response, and decommissioning. Develop automation and remediation workflows using Python, Bash, GitOps, and AI-assisted tools. Build observability solutions, improve reliability and resilience, enforce security and compliance controls, conduct root-cause analysis, and drive operational improvements across engineering, SRE, security, and development teams. Participate in on-call rotations and work onsite.
About this role:
We are seeking a highly skilled and forward-thinking Lead Engineer to join our Technology Operations team. This role is ideal for someone who excels in Kubernetes and OpenShift platform operations, drives operational excellence, and leads initiatives that improve stability, automation, and service reliability. You will play a key role in operating and improving our cloud-native platforms, reducing operational toil, and ensuring the resilience and compliance of critical infrastructure services.
In this role, you will:
4 Sep 2026
*Job posting may come down early due to volume of applicants.
We Value Equal Opportunity
Wells Fargo is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, status as a protected veteran, or any other legally protected characteristic.
Employees support our focus on building strong customer relationships balanced with a strong risk mitigating and compliance-driven culture which firmly establishes those disciplines as critical to the success of our customers and company. They are accountable for execution of all applicable risk programs (Credit, Market, Financial Crimes, Operational, Regulatory Compliance), which includes effectively following and adhering to applicable Wells Fargo policies and procedures, appropriately fulfilling risk and compliance obligations, timely and effective escalation and remediation of issues, and making sound risk decisions. There is emphasis on proactive monitoring, governance, risk identification and escalation, as well as making sound risk decisions commensurate with the business unit's risk appetite and all risk and compliance program requirements.
Candidates applying to job openings posted in Canada: Applications for employment are encouraged from all qualified candidates, including women, persons with disabilities, aboriginal peoples and visible minorities. Accommodation for applicants with disabilities is available upon request in connection with the recruitment process.
Applicants with Disabilities
To request a medical accommodation during the application or interview process, visit Disability Inclusion at Wells Fargo .
Drug and Alcohol Policy
Wells Fargo maintains a drug free workplace. Please see our Drug and Alcohol Policy to learn more.
Wells Fargo Recruitment and Hiring Requirements:
a. Third-Party recordings are prohibited unless authorized by Wells Fargo.
b. Wells Fargo requires you to directly represent your own experiences during the recruiting and hiring process.
We are seeking a highly skilled and forward-thinking Lead Engineer to join our Technology Operations team. This role is ideal for someone who excels in Kubernetes and OpenShift platform operations, drives operational excellence, and leads initiatives that improve stability, automation, and service reliability. You will play a key role in operating and improving our cloud-native platforms, reducing operational toil, and ensuring the resilience and compliance of critical infrastructure services.
In this role, you will:
- Platform Operations Leadership: Lead day-to-day REDIS, OpenShift platform operations, including cluster maintenance, upgrades, performance monitoring, and troubleshooting.
- Incident Response & Problem Management: Serve as an operational lead during incidents, driving rapid diagnosis, resolution, root-cause analysis, and long-term corrective actions.
- Operational Automation: Develop or enhance automation (Python, Bash, GitOps workflows, or AI-assisted tools), build AI Agent,
- REDIS Platform Readiness: Lead REDIS lifecycle activities, including new cluster builds, configuration, onboarding, upgrades, and cluster decommissioning, ensuring consistency, reliability, and compliance across environments.
- Collaboration & Enablement: Partner with engineering, SRE, security, and development teams to implement repeatable operational patterns, guardrails, and platform readiness standards.
- Security, Compliance & Governance: Ensure platform operations follow organizational policies, security standards, audit controls, and regulatory requirements.
- Continuous Improvement: Identify operational gaps, recurring issues, or inefficiencies and lead initiatives to enhance reliability, resiliency, and operational maturity.
- 5+ years of Systems Engineering, Technology Architecture experience, or equivalent demonstrated through one or a combination of the following: work experience, training, military experience, education
- 5+ years of Systems Operations, Cloud Operations, or Technology Architecture experience
- 5+ years of hands-on experience supporting REDIS, Python platform operations
- 3 + years of experience supporting enterprise level complex applications and platforms in Production
- 5 + years of designing and building complex observability solutions leveraging industry standard toolset and or custom-built solutions
- 5+ years working with configuration and monitoring technologies such as Ansible, Grafana, Elastic, Splunk, Prometheus.
- 2+ years Deep expertise with REDISincludes building clusters with pipelines, diagnosing, debugging, remediation, upgrades, patching, and RCA.
- 2+ years' experience building automated remediation workflows and operational tools.
- 2+ years of Linux system operations experience
- Strong analytical and operational problem-solving skills
- Experience with Open shift, Kubernetes
- Hands-on experience with operational tooling such as Grafana, Splunk, Prometheus, Jira, or GitHub, SDLC
- Demonstrated ability to influence operational improvements across teams
- AI development (Agents, MCP, Tools, Skills)
- Ability to work on-site at approved location listed
- This position is not available for visa sponsorship
- Relocation assistance is not available for this position
- Participation in on-call rotations
4 Sep 2026
*Job posting may come down early due to volume of applicants.
We Value Equal Opportunity
Wells Fargo is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, status as a protected veteran, or any other legally protected characteristic.
Employees support our focus on building strong customer relationships balanced with a strong risk mitigating and compliance-driven culture which firmly establishes those disciplines as critical to the success of our customers and company. They are accountable for execution of all applicable risk programs (Credit, Market, Financial Crimes, Operational, Regulatory Compliance), which includes effectively following and adhering to applicable Wells Fargo policies and procedures, appropriately fulfilling risk and compliance obligations, timely and effective escalation and remediation of issues, and making sound risk decisions. There is emphasis on proactive monitoring, governance, risk identification and escalation, as well as making sound risk decisions commensurate with the business unit's risk appetite and all risk and compliance program requirements.
Candidates applying to job openings posted in Canada: Applications for employment are encouraged from all qualified candidates, including women, persons with disabilities, aboriginal peoples and visible minorities. Accommodation for applicants with disabilities is available upon request in connection with the recruitment process.
Applicants with Disabilities
To request a medical accommodation during the application or interview process, visit Disability Inclusion at Wells Fargo .
Drug and Alcohol Policy
Wells Fargo maintains a drug free workplace. Please see our Drug and Alcohol Policy to learn more.
Wells Fargo Recruitment and Hiring Requirements:
a. Third-Party recordings are prohibited unless authorized by Wells Fargo.
b. Wells Fargo requires you to directly represent your own experiences during the recruiting and hiring process.
Wells Fargo Charlotte, North Carolina, USA Office
355 W Martin Luther King, Jr BLVD, Charlotte, NC, United States, 28202
Similar Jobs at Wells Fargo
Fintech • Financial Services
Leads Site Reliability Engineering and systems operations for critical consumer-facing platforms. Responsibilities include improving reliability, resilience, observability, scalability, and operational readiness; establishing SLOs, SLIs, and error budgets; leading major incident response and root-cause analysis; developing monitoring, automation, self-healing, and recovery capabilities; managing vendor dependencies; and mentoring engineering teams across a regulated financial-services environment.
Top Skills:
AppdynamicsAutomationCi/CdCloud-Native ArchitecturesDistributed SystemsDynatraceError BudgetsGcp MonitoringGrafanaInfrastructure As CodeObservabilitySite Reliability EngineeringSlisSlosSplunk
Fintech • Financial Services
Leads systems operations and site reliability engineering for modernized, cloud-native payment platforms. Responsibilities include defining non-functional requirements, capacity and performance testing, observability, SLOs, resilience validation, production readiness, cutovers, runbooks, on-call support, chaos engineering, root-cause analysis, and service stabilization. The role collaborates with engineering, architecture, and service operations teams to build resilient, scalable, and supportable event-driven payment systems.
Top Skills:
BlazemeterChaos MonkeyCi/CdCloud-Native ArchitectureDistributed TracingEvent-Driven ArchitectureKubernetesLoggingMetricsMongoDBRedisResilience4JSlo ToolingSpring BootSpring Webflux
Fintech • Financial Services
Leads systems operations and reliability initiatives across platform, application, and engineering teams. Responsibilities include infrastructure planning, production issue resolution, technical change decisions, observability, incident and change management, automation, cloud-native architecture, Kubernetes and OpenShift operations, resiliency engineering, and technical debt remediation. The role drives self-service and operational toil reduction through scripting, infrastructure automation, SRE practices, and Generative AI or agent development, with occasional on-call coverage.
Top Skills:
Ai AgentsAnsibleAPIsAppdynamicsAzureBashBigpandaElkGCPGenerative AiGitGrafanaKubernetesPowershellPrometheusPythonRed Hat Enterprise LinuxRed Hat Openshift Container PlatformSplunk ObservabilityTerraformThousandeyes
What you need to know about the Charlotte Tech Scene
Ranked among the hottest tech cities in 2024 by CompTIA, Charlotte is quickly cementing its place as a major U.S. tech hub. Home to more than 90,000 tech workers, the city’s ecosystem is primed for continued growth, fueled by billions in annual funding from heavyweights like Microsoft and RevTech Labs, which has created thousands of fintech jobs and made the city a go-to for tech pros looking for their next big opportunity.
Key Facts About Charlotte Tech
- Number of Tech Workers: 90,859; 6.5% of overall workforce (2024 CompTIA survey)
- Major Tech Employers: Lowe’s, Bank of America, TIAA, Microsoft, Honeywell
- Key Industries: Fintech, artificial intelligence, cybersecurity, cloud computing, e-commerce
- Funding Landscape: $3.1 billion in venture capital funding in 2024 (CED)
- Notable Investors: Microsoft, Google, Falfurrias Management Partners, RevTech Labs Foundation
- Research Centers and Universities: University of North Carolina at Charlotte, Northeastern University, North Carolina Research Campus

