APPLICATION ADMINISTRATOR LEAD (SITE RELIABILITY ENGINEER) - OPEN SHIFT - 07212026-79339
Tennessee Department of Finance & Administration · Nashville, TN, US
Apply directly on Tennessee Department of Finance & Administration’s careers site — no account needed.
About the role
Job Information
State of Tennessee Job Information
LOCATION OF (1) POSITION(S) TO BE FILLED: DEPARTMENT OF FINANCE & ADMINISTRATION, DAVIDSON COUNTY
The Department of Finance & Administration does not sponsor applicants for work visas.
This position is designed as Hybrid.
This position requires a criminal background check and CJIS/FTI Fingerprints. Therefore, you may be required to provide information about your criminal history in order to be considered for this position.
Qualifications
Education and Experience: Bachelor's degree and five years of relevant experience in system administration, infrastructure, or application support. Associate degree with equivalent experience may be substituted. Graduate coursework may replace up to two years of experience.
Overview
The Application Administrator ' Lead is responsible for ensuring the reliability, availability, and performance of critical enterprise applications and infrastructure. This role supervises and leads cross-functional engineering teams, drives automation and observability initiatives, enforces operational excellence, and collaborates across IT and business units to sustain and improve service-level objectives (SLOs).
Responsibilities
- Lead the design, automation, and operation of scalable infrastructure and application deployments.
- Resolve complex incidents involving compute, networks, and application layers, with root cause analysis and follow-up.
- Implement monitoring, alerting, and metrics to maintain high service availability and reduce MTTR.
- Coordinate and automate application releases, environment migrations, and patching using CI/CD pipelines.
- Mentor team members in engineering, DevOps practices, and tooling.
- Enforce system reliability, security, and compliance using infrastructure-as-code and configuration management.
- Maintain and improve runbooks, postmortems, and knowledge bases for operational continuity.
- Collaborate with vendors and internal teams for third-party integrations, support, and lifecycle management.
- Contribute to strategic planning with reliability-focused cost-benefit analysis and technology roadmaps.
Competencies (KSA's)
Competencies:
1. Business Insight
2. Decision Quality
3. Self-Development
4. Customer Focus
5. Instills Trust
Knowledges:
1. Reliability Engineering & Automation
2. Incident Response & Root Cause Analysis
3. Performance Tuning & Scalability
4. Infrastructure as Code (IaC)
5. Operational Excellence
Skills:
1. Observability (Metrics, Logging, Tracing)
2. Communication & Cross-Team Collaboration
3. Security & Compliance Awareness
Abilities:
1. Perseverance
2. Logical Thought
Tools & Equipment
1. Observability platforms (Datadog, Prometheus)
2. Configuration management (Ansible, Puppet)
3. CI/CD tools (Jenkins, GitLab)
4. Cloud services (AWS, Azure, GCP)
5. Container orchestration (Kubernetes, Docker)
Description sourced from the public Indeed listing — this role isn't indexed from the company's career page yet.
Skills
- Datadog
- Prometheus
- Ansible
- Jenkins
- AWS
- Azure
- GCP
- Docker
- Kubernetes
Never be applicant #200 again
Every job here is indexed straight from company career pages — often hours after it opens, before it reaches the big boards. Create a free account and get your best matches in a twice-daily digest.
- Your best matches, twice a day
- No duplicates, no ghost jobs, no recruiter spam
- Every job free to browse — pay only when you apply
Free account — no card required
93 151 live jobs · 17 788 companies tracked · 5 869 added today