APPLICATION ADMINISTRATOR LEAD (SITE RELIABILITY ENGINEER) - OPEN SHIFT - 07212026-79339

State of Tennessee

Nashville, TN

JOB DETAILS
SKILLS
Ansible, Automation, Computer Networks, Configuration Management, Continuous Deployment/Delivery, Continuous Integration, Cost Benefit Analysis, Cross-Functional, Customer Relations, DevOps, Enterprise Applications, Finance, High Availability, Jenkins, Knowledge Base, Machine Tool, Mentoring, Metrics, Performance Tuning/Optimization, Puppet (Configuration Management), Reliability Engineering, Root Cause Analysis, Software Administration, Software Patches, Strategic Planning, Systems Administration/Management, Systems Reliability, Team Lead/Manager
LOCATION
Nashville, TN
POSTED
3 days ago

Job Information

State of Tennessee Job Information

Opening Date/Time07/21/2026 12:00AM Central TimeClosing Date/Time08/03/2026 11:59PM Central TimeSalary (Monthly)$7,458.00 - $9,697.00Salary (Annually)$89,496.00 - $116,364.00Job TypeFull-TimeCity, State LocationNashville, TNDepartmentFinance and Administration

LOCATION OF (1) POSITION(S) TO BE FILLED: DEPARTMENT OF FINANCE & ADMINISTRATION, DAVIDSON COUNTY

The Department of Finance & Administration does not sponsor applicants for work visas.

This position is designed as Hybrid.

This position requires a criminal background check and CJIS/FTI Fingerprints. Therefore, you may be required to provide information about your criminal history in order to be considered for this position.

Qualifications

Education and Experience: Bachelor's degree and five years of relevant experience in system administration, infrastructure, or application support. Associate degree with equivalent experience may be substituted. Graduate coursework may replace up to two years of experience.

Overview

The Application Administrator ¿ Lead is responsible for ensuring the reliability, availability, and performance of critical enterprise applications and infrastructure. This role supervises and leads cross-functional engineering teams, drives automation and observability initiatives, enforces operational excellence, and collaborates across IT and business units to sustain and improve service-level objectives (SLOs).

Responsibilities

  • Lead the design, automation, and operation of scalable infrastructure and application deployments.
  • Resolve complex incidents involving compute, networks, and application layers, with root cause analysis and follow-up.
  • Implement monitoring, alerting, and metrics to maintain high service availability and reduce MTTR.
  • Coordinate and automate application releases, environment migrations, and patching using CI/CD pipelines.
  • Mentor team members in engineering, DevOps practices, and tooling.
  • Enforce system reliability, security, and compliance using infrastructure-as-code and configuration management.
  • Maintain and improve runbooks, postmortems, and knowledge bases for operational continuity.
  • Collaborate with vendors and internal teams for third-party integrations, support, and lifecycle management.
  • Contribute to strategic planning with reliability-focused cost-benefit analysis and technology roadmaps.

Competencies (KSA's)

Competencies:

  1. Business Insight

  2. Decision Quality

  3. Self-Development

  4. Customer Focus

  5. Instills Trust

Knowledges:

  1. Reliability Engineering & Automation

  2. Incident Response & Root Cause Analysis

  3. Performance Tuning & Scalability

  4. Infrastructure as Code (IaC)

  5. Operational Excellence

Skills:

  1. Observability (Metrics, Logging, Tracing)

  2. Communication & Cross-Team Collaboration

  3. Security & Compliance Awareness

Abilities:

  1. Perseverance

  2. Logical Thought

Tools & Equipment

  1. Observability platforms (Datadog, Prometheus)

  2. Configuration management (Ansible, Puppet)

  3. CI/CD tools (Jenkins, GitLab)

  4. Cloud services (AWS, Azure, GCP)

  5. Container orchestration (Kubernetes, Docker)

About the Company

S

State of Tennessee