he platform is mature and operational. Your primary focus will be maintaining reliability, performing upgrades, managing compliance, and improving automation. You will provide US timezone coverage alongside existing team members, ensuring round-the-clock operational resilience for this critical platform.
What You'll Do
Operate and maintain production and pre-production GitLab environments
Perform GitLab version upgrades through the Stage-to-Production pipeline
Execute system patching, vulnerability remediation, and compliance tasks
Manage GitLab Shared Runner infrastructure
Manage GitLab Geo replication across primary and secondary sites
Conduct and maintain disaster recovery exercises and documentation
Automate secret rotation via Ansible Automation Platform (AAP)
Maintain and improve Infrastructure as Code (Ansible/Terraform + GitLab CI)
Handle SNOW tickets access requests, pipeline issues, configuration changes
Monitor service health using Prometheus, Grafana, and Datadog
Participate in on-call rotation with peer engineers
Create and maintain runbooks, documentation, and post-incident reviews
Required Qualifications
4+ years of experience in Platform Engineering, DevOps, or production platform operations
Strong GitLab administration experience installation, configuration, upgrades, Geo replication, backup/restore at scale
Linux systems administration (RHEL/CentOS)
Infrastructure as Code proficiency Ansible, Terraform, and CI/CD pipelines
Monitoring and observability experience Prometheus, Grafana, or equivalent
Containerization and orchestration Docker/Podman and Kubernetes/OpenShift
Incident management experience on-call, incident response, root cause analysis
Networking fundamentals DNS, load balancing, VPN, firewall rules
Strong documentation skills
Preferred Qualifications
GitLab Geo replication operations and troubleshooting
Enterprise compliance frameworks (SOC2, ISO 27001, or equivalent; CLIENT ESS/PIA/SIA a plus)
IAM integration (SSO/SAML, LDAP)
High-availability architecture design and operations
CLIENT or IBM enterprise environment experience
Datadog monitoring platform
AWS infrastructure operations
Key Stakeholders
ALM/DEP Platform ownership, priority alignment
InfoSec SOC monitoring, incident response, vulnerability remediation
IT-IAM User provisioning, SSO integration
Engineering teams Thousands of users relying on platform availability
PCO/Ops Infrastructure, networking, AWS account management
Platform State
GitLab 10k reference architecture with high availability
Geo replication across multiple AWS regions
Automated deployment via Ansible/Terraform + GitLab CI
Monitoring: Prometheus + Grafana + Datadog
C1 Mission-Critical service
This Role Is NOT
A build-from-scratch project the infrastructure is mature and well-documented
A pure development role this is infrastructure operations
A solo position you join an existing team of engineers
A user support role you manage the platform, not individual project workflows
Interview Process
TA Screen Recruiter conversation, logistics, and fit
Manager Interview Motivation, expectations, team fit, and competency evaluation
Technical Interview (1 1.5h) Deep dive into relevant technical skills with scenario-based questions
Competencies Evaluated
Motivation What drives you? What do you expect from this role?
RH Multiplier How do you share knowledge and lift others?
Influence How do you drive alignment across teams?
EEO:
Mindlance is an Equal Opportunity Employer and does not discriminate in employment on the basis of Minority/Gender/Disability/Religion/LGBTQI/Age/Veterans.
| Location | Raleigh, NC |
Browse dozens of recruiter approved resume templates, layouts and formats. Choose your favorite and make it your own in minutes.
Free resume templatesImprove your existing resume or start from scratch and create a standout, ATS-friendly resume. Add job-specific content, download and apply.
Free resume builder