Wells Fargo & Co logo

Site Reliability Engineer (SRE) / Production Support Engineer

Wells Fargo & Co
  • Charlotte, NC
    10 days ago

    Job Description

    Title: Site Reliability Engineer (SRE) / Production Support Engineer

    Location: 300 S Brevard St Charlotte, NC

    Duration: 18 months, possibility of extending or converting

    Work Engagement: W2

    Work Schedule: Hybrid

    Benefits on offer for this contract position: Health Insurance, Life insurance, 401K and Voluntary Benefits

    Overview

    We are seeking a highly motivated Site Reliability Engineer (SRE) / Production Support Engineer to join a team supporting critical Small Business Lending and Deposit platforms, including digital banking applications used across a nationwide branch network. This team plays a key role in modernizing enterprise applications, improving platform reliability, and supporting the migration of workloads to OpenShift Container Platform (OCP).

    The ideal candidate will bring strong production support and SRE expertise while helping drive automation, observability, monitoring enhancements, and continuous improvement initiatives within a large-scale enterprise environment.

    Key Responsibilities

    Production Support & Operations

    • Monitor production environments and maintain overall platform health.

    • Participate in daily production support standups and operational reviews.

    • Review and address overnight handoffs from offshore support teams.

    • Investigate, troubleshoot, and resolve incidents, alerts, and service disruptions.

    • Support production deployments, releases, and infrastructure changes.

    • Collaborate across application, infrastructure, and operations teams to ensure service stability and performance.

    Site Reliability Engineering & Platform Modernization

    • Enhance monitoring, observability, and alerting capabilities across supported applications.

    • Build and maintain dashboards that provide actionable operational insights.

    • Implement proactive alerting strategies to improve incident detection and response.

    • Support onboarding and migration of applications to OpenShift Container Platform (OCP).

    • Drive reliability, scalability, and operational excellence initiatives.

    • Identify opportunities for automation and eliminate manual operational tasks.

    Automation & Engineering

    • Develop scripts and tools using Python, Shell Scripting, and Linux/Unix technologies.

    • Leverage Infrastructure as Code (IaC) tools to streamline deployments and operational processes.

    • Create automation solutions using Terraform and Ansible.

    • Contribute to continuous improvement efforts focused on efficiency, stability, and supportability.

    Required Qualifications

    Applicants must be authorized to work for ANY employer in the U.S. This position is not eligible for visa sponsorship.

    • 5+ years of experience in Site Reliability Engineering, Production Support, DevOps, or Infrastructure Engineering roles.

    • Strong experience with:

    • Monitoring and observability practices

    • Incident management and production support

    • Alerting strategy design and dashboard creation

    • Root cause analysis and problem resolution

    • Hands-on experience with:

    • OpenShift Container Platform (OCP)

    • Containerization technologies

    • Linux/Unix environments

    • Python and Shell scripting

    • Experience leveraging monitoring and observability tools such as:

    • Grafana

    • Splunk

    • AppDynamics

    • ThousandEyes

    • Experience with Infrastructure as Code and automation tools including:

    • Terraform

    • Ansible

    • Strong understanding of production deployment processes and operational best practices.

    Preferred Qualifications

    Business Applications

    • Experience supporting Salesforce environments.

    • Experience with nCino is highly desirable.

    AI & Emerging Technologies

    • Experience utilizing Microsoft Copilot or similar AI-enabled tools.

    • Knowledge of AI-driven monitoring, observability, or operational automation solutions.

    • Experience applying AI to non-functional requirements, security monitoring, vulnerability management, or platform operations.

    Industry Experience

    • Financial Services industry experience preferred.

    • Experience supporting large enterprise environments preferred.

    Soft Skills & Success Profile

    We are looking for someone who brings more than technical expertise. Successful candidates will demonstrate:

    • Strong ownership and accountability.

    • Self-starter mentality with the ability to work independently.

    • Quick learner with a passion for new technologies.

    • Ability to identify problems and propose innovative solutions.

    • Strong communication and collaboration skills.

    • Process improvement mindset and continuous learning approach.

    • Ability to influence positive change within the organization.

    • Team-oriented attitude with a strong focus on knowledge sharing and collaboration.

    Numbers & Facts

    LocationCharlotte, NC
    IndustryFinancial Services
    Company Size10,000 employees or more
    Year Founded1852

    About Company

    We believe in our vision and values just as strongly today as we did the first time we put them on paper more than 20 years ago. Staying true to them will guide us toward continued growth and success for decades to come. As you read more about our vision and values, you will learn about who we are, where we’re headed and how every Wells Fargo team member can help us get there.

    Skills

    • Ansibleunmatched
    • Artificial Intelligence (AI)unmatched
    • Automationunmatched
    • Banking Servicesunmatched
    • Best Practicesunmatched
    • Business Loansunmatched
    • Business Operationsunmatched
    • Communication Skillsunmatched
    • Computer Securityunmatched
    • Continuous Improvementunmatched
    • DevOpsunmatched
    • Emerging Technologyunmatched
    • Enterprise Applicationsunmatched
    • Identify Issuesunmatched
    • Incident Managementunmatched
    • Incident Responseunmatched
    • Linux Operating Systemunmatched
    • Microsoft Product Familyunmatched
    • Offshoringunmatched
    • Operational Auditunmatched
    • Operational Supportunmatched
    • Operations Processesunmatched
    • Process Improvementunmatched
    • Production Controlunmatched
    • Production Supportunmatched
    • Production Systemsunmatched
    • Python Programming/Scripting Languageunmatched
    • Reliability Engineeringunmatched
    • Reporting Dashboardsunmatched
    • Salesforce.comunmatched
    • Scripting (Scripting Languages)unmatched
    • Security Monitoringunmatched
    • Software Administrationunmatched
    • Standup Meetingsunmatched
    • Team Playerunmatched
    • Technical Supportunmatched
    • Unix Operating Systemsunmatched
    • Unix Shell Programmingunmatched

    Be found by employers

    5,500+ employers search our resume database daily. Add yours to get found by recruiters looking for candidates like you.

    Level up your application

    Professional resume templates

    Browse dozens of recruiter approved resume templates, layouts and formats. Choose your favorite and make it your own in minutes.

    Free resume templates

    Free resume builder

    Improve your existing resume or start from scratch and create a standout, ATS-friendly resume. Add job-specific content, download and apply.

    Free resume builder