Production Support Analyst/SRE Reliability engineer

Veterans Sourcing Group

  • South Jordan, UT
  • 30+ days ago
  • $53 Per Hour
Want to know if you’re a fit?
Upload your resume and let our AI show you.

Skills

  • Apache Kafkaunmatched
  • Artificial Intelligence (AI)unmatched
  • Automationunmatched
  • C++ Programming Languageunmatched
  • Cloud Computingunmatched
  • Communication Skillsunmatched
  • Consultingunmatched
  • Continuous Deployment/Deliveryunmatched
  • Continuous Integrationunmatched
  • Debugging Skillsunmatched
  • Distributed Computingunmatched
  • Dockerunmatched
  • Go Programming Language (Golang)unmatched
  • ITIL (IT Infrastructure Library)unmatched
  • Identify Issuesunmatched
  • Incident Managementunmatched
  • Javaunmatched
  • Linux Operating Systemunmatched
  • On Callunmatched
  • Operating Systemsunmatched
  • Operational Supportunmatched
  • Problem Solving Skillsunmatched
  • Production Controlunmatched
  • Production Supportunmatched
  • Production Systemsunmatched
  • Python Programming/Scripting Languageunmatched
  • Reliability Analysisunmatched
  • Reliability Engineeringunmatched
  • Root Cause Analysisunmatched
  • SQL (Structured Query Language)unmatched
  • Scala Programming Languageunmatched
  • Scripting (Scripting Languages)unmatched
  • ServiceNowunmatched
  • Splunkunmatched
  • Standard Operating Procedures (SOP)unmatched
  • Systems Reliabilityunmatched
  • Unix Operating Systemsunmatched

Description

Production Support Analyst / SRE Reliability engineer
Location: South Jordan, UT (Hybrid)
12 months Contract-to-Hire | 
Pay rate: $53/hr W2

Interview Process:

  • 1st Round: Zoom
  • 2nd Round: Onsite

Experience & Education:

  • 2–5 years of relevant experience
  • Bachelor's Degree required

Shifts:

  • Morning: 8:00 AM – 5:00 PM
  • Evening: 12:30 PM – 8:00 AM
  • Weekend: On-call (Remote)

Key Responsibilities:

  • Monitor and support production systems across OS, applications, and network
  • Troubleshoot incidents, perform root cause analysis, and resolve live issues
  • Collaborate with Dev teams to reduce recurring issues and improve system reliability
  • Automate repetitive tasks using Python/scripting
  • Maintain SOPs and support operational readiness activities
  • Participate in on-call rotation and critical event support

Must-Have Skills:

  • Strong hands-on experience with  Linux/Unix (OS-level troubleshooting)
  • Production support experience (incident management, debugging live systems)
  • Python scripting (automation-focused, not development-heavy)
  • SQL knowledge
  • Experience with  ServiceNow (ticketing)
  • Understanding of  ITIL principles
  • Excellent communication skills

Nice to Have:

  • Exposure to Java, Go, C++, Scala
  • Monitoring tools (Grafana, Splunk, Dynatrace, etc.)
  • Cloud experience
  • Snowflake knowledge
  • CI/CD, Kafka, Docker, or distributed systems exposure
  • Awareness of SRE concepts or Agentic AI

Role Overview:

This role supports  Production Support / SRE (Reliability Engineering) functions—focused on system stability, incident resolution, automation, and improving platform reliability in a large-scale Linux environment.

Numbers & Facts

LocationSouth Jordan, UT

Similar Jobs

See more jobs