Senior DevOps Engineer

Iconma LLC
  • AR
    21 days ago

    Job Description

    Our Client, a Business Maunufacturing and Supply company, is looking for a Senior DevOps Engineer for their Remote location.

    Responsibilities:

    • Design, implement, and operate cloud infrastructure and automation using AWS and Terraform across services such as EC2, S3, EKS, ECR, Route 53, IAM, and core networking components.
    • Own EKS and Kubernetes platform operations, including cluster lifecycle management, upgrades, node group strategy, autoscaling, workload scheduling, capacity planning, performance tuning, and cost optimization.
    • Build and improve core Kubernetes platform capabilities such as networking, ingress, DNS integration, service discovery, RBAC, policy enforcement, secrets management, and multi-environment consistency.
    • Develop and maintain deployment workflows using Helm, Kustomize, GitLab CI/CD, and GitOps approaches with tools such as Argo CD or Flux to improve release speed, consistency, traceability, and rollback safety.
    • Deploy and evolve observability capabilities across infrastructure and Kubernetes environments, including metrics, logging, alerting, dashboards, and incident diagnostics.
    • Support and optimize platform and application workloads on Kubernetes, improving deployment patterns, scaling behavior, runtime efficiency, resilience, and day-to-day operational support.
    • Strengthen security across infrastructure, pipelines, and Kubernetes environments through IAM least-privilege access, secrets management, image and artifact scanning, admission controls, policy-as-code, workload isolation, and practical runtime hardening.
    • Support the reliability, backup integrity, availability, and operational performance of PostgreSQL/RDS environments in partnership with application teams.
    • Work closely with development and platform teams to improve deployment strategies, runtime reliability, developer experience, and operational standards for cloud-native systems.
    • Monitor system health, investigate incidents, troubleshoot infrastructure and application issues, and drive timely resolution through strong root cause analysis and preventative improvements.
    • Lead technical decision-making within the scope of the role by prioritizing work, evaluating tradeoffs, integrating inputs from multiple stakeholders, and driving issues through to completion with limited oversight.
    • Use modern AI tools to accelerate infrastructure design exploration, Terraform authoring, Kubernetes troubleshooting, CI/CD workflow drafting, observability analysis, and operational documentation, while rigorously validating outputs for correctness, security, maintainability, and production readiness.
    • Document and continually refine DevOps methodologies, infrastructure standards, deployment workflows, operational procedures, and support runbooks.

    Requirements:

    • 7+ years of experience in DevOps, platform engineering, SRE, or related roles, with a proven track record in designing and operating scalable cloud infrastructure.
    • Strong expertise with AWS services, including EC2, S3, EKS, ECR, Route 53, IAM, and core networking and security concepts.
    • Advanced hands-on experience with Kubernetes, including cluster operations, upgrades, autoscaling, networking, ingress, RBAC, policy enforcement, and production workload management.
    • Strong experience with Terraform for infrastructure-as-code, environment management, and repeatable platform provisioning.
    • Experience with Kubernetes packaging and deployment tools such as Helm and Kustomize.
    • Experience with CI/CD systems such as GitLab CI/CD and familiarity with GitOps approaches using Argo CD, Flux, or similar tools.
    • Strong experience with observability and monitoring tools such as Prometheus, Grafana, Loki, or comparable tooling.
    • Experience evaluating or operating Kubernetes ecosystem tools such as ingress controllers, operators, Cluster Autoscaler, or Karpenter.
    • Strong experience with PostgreSQL/RDS, including performance tuning, reliability, and operational support.
    • Solid understanding of cloud and container security practices, including IAM least privilege, secrets management, image scanning, admission controls, and policy enforcement.
    • Strong scripting and automation skills in Python, Bash, or similar languages.
    • Practical fluency with AI-assisted engineering tools such as LLMs, code assistants, and workflow automation tools, along with strong judgment to evaluate and refine AI-generated outputs responsibly.
    • Excellent problem-solving, teamwork, and communication skills.
    • Ability to work independently in a remote-friendly environment and make sound technical decisions in ambiguous situations.
    • Bachelor's Degree or higher in Computer Science, Information Technology, or a related field, or equivalent practical experience.
    • Certifications in AWS, Kubernetes, Terraform, or related platform technologies.
    • Experience with additional cloud platforms such as GCP or Azure.
    • Experience improving incident response, operational readiness, and platform standards for distributed engineering teams.
    • Experience designing internal platform capabilities that improve developer self-service and reduce operational toil.
    • Familiarity with compliance, governance, and risk management requirements in regulated or enterprise environments.

    Why Should You Apply?

    • Health Benefits
    • Referral Program
    • Excellent growth and advancement opportunities

    Numbers & Facts

    LocationAR

    Skills

    • Amazon Elastic Compute Cloud (EC2)unmatched
    • Amazon Simple Storage Service (S3)unmatched
    • Amazon Web Services (AWS)unmatched
    • Artificial Intelligence (AI)unmatched
    • Automationunmatched
    • Autoscalingunmatched
    • Bash Scriptingunmatched
    • Calendar Managementunmatched
    • Capacity and Performance Managementunmatched
    • Cloud Computingunmatched
    • Communication Skillsunmatched
    • Computer Scienceunmatched
    • Continuous Deployment/Deliveryunmatched
    • Continuous Integrationunmatched
    • Cost Controlunmatched
    • DNS (Domain Name System)unmatched
    • DevOpsunmatched
    • Documentationunmatched
    • Ecosystemsunmatched
    • Environmental Managementunmatched
    • Establish Prioritiesunmatched
    • GCP (Good Clinical Practices)unmatched
    • Health Planunmatched
    • Identify Issuesunmatched
    • Incident Responseunmatched
    • Information Technology & Information Systemsunmatched
    • Machine Toolunmatched
    • Metricsunmatched
    • Microsoft Windows Azureunmatched
    • Network Securityunmatched
    • Operational Auditunmatched
    • Operational Supportunmatched
    • Operations Processesunmatched
    • Performance Tuning/Optimizationunmatched
    • PostgreSQLunmatched
    • Problem Solving Skillsunmatched
    • Python Programming/Scripting Languageunmatched
    • Reporting Dashboardsunmatched
    • Risk Managementunmatched
    • Root Cause Analysisunmatched
    • Scripting (Scripting Languages)unmatched
    • Security Infrastructureunmatched
    • Software Engineeringunmatched
    • Team Playerunmatched
    • Technical Leadershipunmatched
    • Time Managementunmatched
    • Traceabilityunmatched
    • Trade-Off Analysisunmatched

    Be found by employers

    5,500+ employers search our resume database daily. Add yours to get found by recruiters looking for candidates like you.

    Level up your application

    Professional resume templates

    Browse dozens of recruiter approved resume templates, layouts and formats. Choose your favorite and make it your own in minutes.

    Free resume templates

    Free resume builder

    Improve your existing resume or start from scratch and create a standout, ATS-friendly resume. Add job-specific content, download and apply.

    Free resume builder