Senior Data Engineer

Tata Consultancy Services Ltd
  • Irvine, CA
  • $120,000–$180,000 Per Year
24 days ago

Job Description

Senior Data Engineer

We are seeking a hands-on Senior Data Engineer to design, build, optimize, and support enterprise-scale data solutions on the Databricks Lakehouse Platform. The role will develop reliable batch and streaming pipelines, modernize legacy ETL workloads, implement governed data models, and deliver trusted data products for analytics, reporting, risk, regulatory, and investment-management use cases.

The ideal candidate has deep experience with Databricks, Apache Spark, PySpark, SQL, Python, dbt, Apache Airflow, Delta Lake, Unity Catalog, cloud storage, data quality, CI/CD, and production support.

Key Responsibilities

Data Engineering and Development

  • Design, build, test, deploy, and maintain scalable ETL and ELT pipelines using Databricks, PySpark, Spark SQL, Python, and SQL.
  • Develop reusable ingestion and transformation frameworks for structured, semi-structured, and streaming data.
  • Implement batch, incremental, change-data-capture, and streaming processing patterns.
  • Build and maintain Delta Lake tables using medallion architecture across Bronze, Silver, and Gold layers.
  • Develop dbt models, tests, macros, packages, documentation, and incremental processing patterns.
  • Create and maintain Apache Airflow DAGs and Databricks Workflows with dependency management, retries, alerting, and operational controls.
  • Integrate data from APIs, databases, files, event streams, and cloud data services.
  • Produce technical designs, mapping specifications, lineage documentation, deployment instructions, and operational runbooks.

Performance, Reliability, and Data Quality

  • Tune Spark workloads, joins, partitioning, file sizes, caching, cluster configurations, and query plans.
  • Apply Delta Lake optimization techniques, including compaction, data skipping, clustering, retention, and vacuum controls.
  • Implement automated data quality, reconciliation, schema validation, observability, and freshness checks.
  • Monitor pipeline health and resolve failures, performance degradation, data defects, and service-level breaches.
  • Perform root-cause analysis and implement durable preventive measures.
  • Support release readiness, production cutover, incident resolution, and ongoing platform operations.
  • Improve compute utilization and cost efficiency across batch and streaming workloads.

Governance, Security, and Delivery Practices

  • Apply Unity Catalog standards for catalogs, schemas, tables, views, lineage, classification, and controlled access.
  • Implement secure handling of credentials, secrets, personally identifiable information, and regulated data.
  • Contribute to CI/CD pipelines, automated testing, code-quality checks, and environment promotion.
  • Use Git-based development, peer reviews, branching standards, and release-management practices.
  • Collaborate with platform engineers to deploy data assets through Terraform and Databricks Asset Bundles where applicable.
  • Follow enterprise architecture, security, data-governance, and regulatory requirements.

Collaboration and Mentoring

  • Partner with architects, product owners, analysts, data scientists, governance teams, and business stakeholders.
  • Translate business requirements into scalable data models, pipelines, and technical work packages.
  • Conduct code reviews and enforce engineering, documentation, testing, and support standards.
  • Mentor junior and mid-level engineers and share reusable patterns and best practices.
  • Communicate delivery status, risks, dependencies, and technical trade-offs clearly.

Required Qualifications

  • Typically 7-10 years of data engineering, data warehousing, or distributed data-processing experience.
  • Strong hands-on experience with Databricks, Apache Spark, PySpark, Delta Lake, Python, and advanced SQL.
  • Experience building production-grade ETL and ELT pipelines for large datasets.
  • Experience with dbt Core or dbt Cloud, including models, macros, tests, documentation, and incremental processing.
  • Experience with Apache Airflow, Astronomer, Databricks Workflows, or comparable orchestration platforms.
  • Experience with Unity Catalog, data lineage, role-based access, and data-governance controls.
  • Experience with cloud data services on AWS, Azure, or Google Cloud.
  • Working knowledge of Git, CI/CD, automated testing, monitoring, and production-support practices.
  • Strong troubleshooting, communication, collaboration, and technical-documentation skills.

Preferred Qualifications

  • Experience in banking, financial services, insurance, asset management, risk, compliance, or another regulated industry.
  • Experience modernizing Hadoop, legacy data warehouses, or traditional ETL platforms.
  • Experience with Kafka, Structured Streaming, Auto Loader, Delta Live Tables, or Lakeflow Declarative Pipelines.
  • Familiarity with Terraform, Databricks Asset Bundles, cloud networking, IAM, secrets management, and infrastructure automation.
  • Databricks Data Engineer Associate or Professional certification.
  • Experience delivering data reconciliation, regulatory reporting, test automation, and audit-ready controls.

Salary Range- $120,000-$180,000 a year

Numbers & Facts

LocationIrvine, CA
Salary$120,000–$180,000 Per Year

Skills

  • Access Controlunmatched
  • Amazon Web Services (AWS)unmatched
  • Apacheunmatched
  • Apache Hadoopunmatched
  • Apache Sparkunmatched
  • Application Programming Interface (API)unmatched
  • Asset Managementunmatched
  • Automationunmatched
  • Banking Servicesunmatched
  • Best Practicesunmatched
  • Cachingunmatched
  • Cisco Unityunmatched
  • Cloud Computingunmatched
  • Cloud Storageunmatched
  • Code Reviewsunmatched
  • Communication Skillsunmatched
  • Continuous Deployment/Deliveryunmatched
  • Continuous Integrationunmatched
  • Data Analysisunmatched
  • Data Clusteringunmatched
  • Data Collectionunmatched
  • Data Modelingunmatched
  • Data Qualityunmatched
  • Data Scienceunmatched
  • Data Warehousingunmatched
  • Database Extract Transform and Load (ETL)unmatched
  • Documentationunmatched
  • Financial Servicesunmatched
  • Gitunmatched
  • Identify Issuesunmatched
  • Insuranceunmatched
  • Investment Managementunmatched
  • Mentoringunmatched
  • Microsoft Windows Azureunmatched
  • Production Controlunmatched
  • Production Supportunmatched
  • Python Programming/Scripting Languageunmatched
  • Reconciliationunmatched
  • Regulationsunmatched
  • Regulatory Reportsunmatched
  • Release Management/Engineeringunmatched
  • Requirements Managementunmatched
  • Riskunmatched
  • Root Cause Analysisunmatched
  • SQL (Structured Query Language)unmatched
  • Structured Dataunmatched
  • Technical Writingunmatched
  • Technical/Engineering Designunmatched
  • Test Automationunmatched
  • Testingunmatched
  • Use Casesunmatched

Be found by employers

5,500+ employers search our resume database daily. Add yours to get found by recruiters looking for candidates like you.

Level up your application

Professional resume templates

Browse dozens of recruiter approved resume templates, layouts and formats. Choose your favorite and make it your own in minutes.

Free resume templates

Free resume builder

Improve your existing resume or start from scratch and create a standout, ATS-friendly resume. Add job-specific content, download and apply.

Free resume builder