• HOUSTON, TX
    30+ days ago

    Job Description

    Classification: Admin/Prof

    Exemption Status/Test: Exempt/Computer

    Job Grade: 5

    Department: Information Technology

    Reports to: Department Director

    Job Goal:

    Design, implement, and continuously improve scalable, secure, and highly available infrastructure and deployment pipelines that enable rapid software delivery and reliable system performance. Drive automation across development and operations processes, leverage cloud technologies to optimize cost and efficiency, and collaborate closely with cross-functional teams to enhance system reliability, accelerate innovation, and support seamless, high-quality product releases.

    Qualifications:

    Education:

    • Bachelor's degree within a technical field such as Computer Science or Information Technology, preferred

    Experience:

    • 3 years of experience in DevOps, SRE, or related engineering roles
    • Strong experience with AWS cloud platform (EC2, EKS, VPC, IAM, Auto Scaling Groups, Route53, Lambda, CloudWatch)
    • Hands-on experience with ArgoCD or other GitOps tools for continuous deployment
    • Experience with Helm charts for packaging and deploying Kubernetes applications
    • Solid understanding of AWS IAM and IRSA (IAM Roles for Service Accounts) for secure pod-level permissions
    • Experience with CI/CD pipelines (GitHub Actions, GitLab CI, Jenkins, CircleCI, etc.)

    Special Knowledge and Skills:

    • Deep knowledge of Kubernetes including deployment strategies, resource management, and cluster operations
    • Proficiency with containerization (Docker, container registries, multi-stage builds)
    • Strong Infrastructure-as-Code skills with Terraform for provisioning and managing AWS resources
    • Knowledge of MongoDB deployment and management in containerized environments
    • Familiarity with Linux systems, networking, and security fundamentals
    • Strong communication and collaboration skills
    • Problem-solving mindset with a focus on reliability and scalability
    • Ability to work in fast-paced, agile environments
    • Ownership mentality and proactive approach to automation

    Major Responsibilities

    • Build and maintain CI/CD pipelines to support automated testing, integration, and deployments.
    • Manage cloud infrastructure using Infrastructure-as-Code tools such as Terraform, CloudFormation, or Pulumi.
    • Monitor, troubleshoot, and optimize system performance across distributed applications and microservices.
    • Implement automation for configuration management, provisioning, and workflow orchestration using tools like Ansible, Chef, Puppet, or similar.
    • Enhance system reliability & availability through observability, log management, alerting, and incident response best practices.
    • Collaborate closely with development teams to streamline build processes, improve deployment strategies, and ensure production readiness.
    • Maintain containerized environments using Docker and orchestration tools (Kubernetes, ECS, EKS).
    • Ensure security best practices in infrastructure, pipelines, secrets management, and access control.
    • Document systems, standards, and procedures to improve maintainability and knowledge sharing.

    Supervisory Responsibilities:

    None

    Physical Demands/Environmental Factors/Mental Demands:

    Frequent use of standard office equipment; prolonged sitting; occasional bending/stooping, pushing/pulling, and twisting; repetitive hand motions (keyboarding and use of mouse); occasional light lifting and carrying (less than 15 pounds); may work prolonged and irregular hours; work with frequent interruptions; maintain emotional control under pressure.

    Numbers & Facts

    LocationHOUSTON, TX

    Skills

    • AWS Lambdaunmatched
    • Access Controlunmatched
    • Amazon Elastic Compute Cloud (EC2)unmatched
    • Amazon Web Services (AWS)unmatched
    • Ansibleunmatched
    • Automationunmatched
    • Autoscalingunmatched
    • Best Practicesunmatched
    • Chef (Configuration Management)unmatched
    • Cloud Computingunmatched
    • Communication Skillsunmatched
    • Computer Scienceunmatched
    • Configuration Managementunmatched
    • Continuous Deployment/Deliveryunmatched
    • Continuous Improvementunmatched
    • Continuous Integrationunmatched
    • Cost Controlunmatched
    • Cross-Functionalunmatched
    • DevOpsunmatched
    • Distributed Applicationsunmatched
    • Dockerunmatched
    • Documentation Standardsunmatched
    • GitHubunmatched
    • High Availabilityunmatched
    • Identify Issuesunmatched
    • Incident Responseunmatched
    • Information Technology & Information Systemsunmatched
    • Integration Testingunmatched
    • Jenkinsunmatched
    • Linux Operating Systemunmatched
    • Microservicesunmatched
    • MongoDBunmatched
    • Network Securityunmatched
    • Operations Processesunmatched
    • Performance Tuning/Optimizationunmatched
    • Problem Solving Skillsunmatched
    • Process Developmentunmatched
    • Process Improvementunmatched
    • Puppet (Configuration Management)unmatched
    • Resource Managementunmatched
    • Sales Pipelineunmatched
    • Systems Reliabilityunmatched
    • Team Playerunmatched
    • Test Automationunmatched

    Be found by employers

    5,500+ employers search our resume database daily. Add yours to get found by recruiters looking for candidates like you.

    Level up your application

    Professional resume templates

    Browse dozens of recruiter approved resume templates, layouts and formats. Choose your favorite and make it your own in minutes.

    Free resume templates

    Free resume builder

    Improve your existing resume or start from scratch and create a standout, ATS-friendly resume. Add job-specific content, download and apply.

    Free resume builder