Our Client, an IT Services and Consultant company, is looking for an AWS Cloud DevOps / Kubernetes Engineer for their Lowell, MA/Troy MI location.
Responsibilities:
Design and manage infrastructure using AWS, Kubernetes, and Terraform, with a focus on Terraform module design, remote state management, and policy-as-code.
Implement GitOps practices using tools like Flux or Argo CD for declarative, version-controlled environment management and drift detection.
Take ownership of Kubernetes as a Service (KaaS) platforms, including multi-cluster lifecycle management, version upgrades, and advanced autoscaling strategies.
Troubleshoot complex traffic flows and optimize cloud networking, including CNI configurations, ingress controllers, and load balancing.
Build and support CI/CD pipelines to enable efficient and reliable deployments.
Automate infrastructure provisioning and configuration using tools like Ansible, Puppet, or Chef.
Write and maintain infrastructure-as-code and configuration-as-code for all environments using Terraform.
Monitor, troubleshoot, and optimize infrastructure performance, availability, and cost.
Integrate and manage third-party tools, plug-ins, and internal scripts as needed.
Create and maintain technical documentation, including runbooks and architecture diagrams.
Participate in on-call rotations and contribute to improving incident response processes.
Collaborate cross-functionally to ensure infrastructure meets application and business needs.
Requirements:
Bachelor’s degree in computer engineering / Software Engineering / Electrical Engineering / Computer Science or equivalent
Experience
7+ years of experience working as DevOps, SRE, PE or infrastructure engineering experience.
5+ years of hands-on experience with cloud platforms, especially AWS (GCP or Azure a plus).
Advanced Linux systems expertise, including deep knowledge of systemd, cgroups, namespaces, OS tuning, and production hardening.
Deep Kubernetes and cloud networking knowledge beyond fundamentals (e.g., CNI plugins, Service Mesh, and advanced ingress patterns).
Experience managing Kubernetes clusters and containerized workloads (EKS / Docker).
Expertise in Terraform and infrastructure automation practices.
Proficiency in scripting with Python and Bash.
Experience with tools like Git, Jenkins, and artifact repositories.
Familiarity with networking fundamentals such as DNS, DHCP, VPN, and LDAP.
Experience with monitoring and logging systems such as DataDog, Prometheus, Grafana, ELK, or equivalent.
Strong understanding of cloud networking, VPCs, security groups, and IAM policies.
Comfortable working with Agile teams and tools like JIRA and Confluence
Nice To Have
Experience with Helm and managing Kubernetes via Helm charts.
Exposure to hybrid environments (cloud + on-prem).