Design, build, and operationalize scalable data pipelines and data products using a Python-based, code-first engineering approach, aligned to enterprise architecture, governance, and AI-readiness goals.
Key Responsibilities
Develop pipelines and transformations using Python (PySpark / notebooks / scripts)
Work within VS Code + GitHub development workflows
Build ingestion, transformation, and curated data layers aligned to medallion architecture
Integrate data from ERP, MES, OT, historian, and operational systems
Convert notebooks into production-grade, reusable components
Implement unit testing, logging, monitoring, and observability frameworks
Follow branching strategies, pull request processes, and CI/CD pipelines
Package reusable logic into shared modules and libraries
Enable creation of certified, semantic-ready data products
Optimize performance, troubleshoot failures, and ensure production reliability
Maintain technical documentation and operational runbooks
Required Skills
Strong proficiency in Python and PySpark-based data engineering
Experience with VS Code, GitHub, and code-based pipeline development
Strong experience with Azure Data Factory, Synapse, Microsoft Fabric, SQL
Understanding of:
Notebook vs. production pipeline design
Code modularization and reuse
Strong debugging, optimization, and problem-solving capabilities
Preferred Background
Manufacturing data experience
Exposure to AI-ready datasets, feature engineering, and data observability
We are committed to fostering a diverse, inclusive, and equitable workplace where individuals from all backgrounds feel valued and empowered to contribute their unique perspectives. We strongly encourage applications from candidates of all genders, races, ethnicities, abilities, and experiences to join our team and help us build a culture of belonging.
Numbers & Facts
Location
Monmouth Junction, New Jersey
Job Type
Contractor
Skills
Artificial Intelligence (AI)unmatched
Continuous Deployment/Deliveryunmatched
Continuous Integrationunmatched
Data Managementunmatched
Data Setsunmatched
Debugging Skillsunmatched
ERP (Enterprise Resource Planning)unmatched
Enterprise Architectureunmatched
GitHubunmatched
Identify Issuesunmatched
Manufacturingunmatched
Microsoft Product Familyunmatched
Microsoft Windows Azureunmatched
Performance Tuning/Optimizationunmatched
Problem Solving Skillsunmatched
Python Programming/Scripting Languageunmatched
SQL (Structured Query Language)unmatched
Scalable System Developmentunmatched
Scripting (Scripting Languages)unmatched
Software Engineeringunmatched
Software Reuseunmatched
Technical Writingunmatched
Unit Testunmatched
🎯
Be found by employers
5,500+ employers search our resume database daily. Add yours to get found by recruiters looking for candidates like you.
Level up your application
Professional resume templates
Browse dozens of recruiter approved resume templates, layouts and formats. Choose your favorite and make it your own in minutes.