Data Platform Infrastructure Engineer

PeopleNTech LLC
  • Santa Clara, CA
  • $65 Per Hour
  • Quick Apply
30+ days ago

Job Description


Role : Data Platform Infrastructure Engineer
Location : Scottsdale AZ (100% onsite)
Rate : $65
Indent :
SF_OP_204640-8-2

Preferred Skill Combination (Must-Have Exposure to One or Both):
  • Cloudera + Databricks + Terraform + AWS

Role Overview
We are seeking a highly skilled Data Platform Infrastructure Engineer to design, build, and manage scalable data platforms across on-premise and cloud environments. The role involves working with cluster technologies, infrastructure automation, and modern data ecosystems to enable reliable and high-performing data platforms.

Key Responsibilities
  • Design, deploy, and manage data platform infrastructure across on-prem (Cloudera) and cloud (AWS, Databricks) environments
  • Build and maintain distributed data clusters ensuring high availability, scalability, and performance
  • Automate infrastructure provisioning using Terraform and Ansible
  • Manage and optimize Cloudera Hadoop ecosystems (HDFS, Hive, Spark, YARN, etc.)
  • Deploy and manage Databricks workspaces, clusters, and integrations on AWS
  • Implement infrastructure-as-code (IaC) and configuration management best practices
  • Monitor cluster performance, troubleshoot issues, and ensure system reliability
  • Collaborate with data engineers, architects, and DevOps teams to support data pipelines and analytics workloads
  • Ensure security, compliance, and governance across data platforms
  • Support migration from on-prem to cloud-based data platforms

Technical Skills Required
Core Technologies
  • Strong experience in Cloudera (CDH/CDP) cluster setup and administration
  • Hands-on experience with Databricks (cluster management, jobs, notebooks)
  • Strong exposure to AWS (EC2, S3, IAM, VPC, EMR, networking concepts)
Infrastructure & Automation
  • Expertise in Terraform (mandatory) for infrastructure provisioning
  • Proficiency in Ansible for configuration management and automation
  • Experience with CI/CD pipelines for infrastructure deployments
Cluster & Data Technologies
  • Experience managing distributed systems / cluster technologies
  • Strong understanding of:
    • Hadoop ecosystem (HDFS, Hive, Spark, Kafka, etc.)
    • Spark performance tuning and cluster optimization
  • Knowledge of containerization (Docker/Kubernetes) is a plus


Numbers & Facts

LocationSanta Clara, CA

Skills

  • Amazon Elastic Compute Cloud (EC2)unmatched
  • Amazon Simple Storage Service (S3)unmatched
  • Amazon Web Services (AWS)unmatched
  • Ansibleunmatched
  • Apache Hadoopunmatched
  • Apache Hiveunmatched
  • Apache Sparkunmatched
  • Automationunmatched
  • Best Practicesunmatched
  • Cloud Computingunmatched
  • Clouderaunmatched
  • Configuration Managementunmatched
  • Continuous Deployment/Deliveryunmatched
  • Continuous Integrationunmatched
  • Data Analysisunmatched
  • Data Clusteringunmatched
  • Data Managementunmatched
  • DevOpsunmatched
  • Distributed Computingunmatched
  • Dockerunmatched
  • Ecosystemsunmatched
  • Electronic Medical Recordsunmatched
  • HDFS (Hadoop Distributed File System)unmatched
  • High Availabilityunmatched
  • Identify Issuesunmatched
  • Maintain Complianceunmatched
  • Multiplatform/Cross-Platformunmatched
  • Performance Analysisunmatched
  • Performance Tuning/Optimizationunmatched
  • Scalable System Developmentunmatched
  • Systems Administration/Managementunmatched
  • Systems Reliabilityunmatched

Be found by employers

5,500+ employers search our resume database daily. Add yours to get found by recruiters looking for candidates like you.

Level up your application

Professional resume templates

Browse dozens of recruiter approved resume templates, layouts and formats. Choose your favorite and make it your own in minutes.

Free resume templates

Free resume builder

Improve your existing resume or start from scratch and create a standout, ATS-friendly resume. Add job-specific content, download and apply.

Free resume builder