Job Title: - Senior Data Scientist - Manufacturing Operations
Location: - Austin, TX (Onsite)
Contract (W2)
Description:
We are hiring a Senior Data Scientist to work with factory and industrial data: find what is going wrong (or about to), build models that hold up against plant reality, and help the team act on them.
Most of the job is in the data - understanding how a process behaves, cleaning noisy and incomplete signals, defining "normal" vs "abnormal" with people who run the line, creating features, validating against real outcomes, and explaining limits when the data cannot support a model. You will partner with manufacturing, quality, maintenance, data engineering, and software. You will not own the data platform. This is not a research lab role and not a platform-engineering role.
What we are hiring for
Someone who can walk a real example: this was the grain of the data, this is what I found, this is the model, this is how I knew it was wrong or right, this is what operations did with it.
Typical problems: process drift, abnormal machine behavior, quality prediction, equipment health, bottlenecks, downtime, scrap/rework, root-cause support. Methods follow the problem (statistical limits, clustering, isolation forest, time series, autoencoders, supervised models when labels exist) - we do not hire to a method list.
Manufacturing experience is a plus. We will also consider people from industrial IoT, equipment, quality, automotive, semiconductor, energy, telecom/ops, or similar operational environments who have done this loop on messy sensor or process data.
What this role is not
Please do not submit (and we will not screen) profiles whose recent work is mainly:
Data engineering / data modeling (lakehouse, medallion, star schema, Unity Catalog, ADF, "pipelines for the DS team")
MLOps / ML platform (Airflow, SageMaker plumbing, FastAPI services, CI/CD) with little analysis of a dataset
GenAI product work (RAG chatbots, LangChain agents, Copilot apps) as the primary story
BI/reporting (Power BI/Tableau KPI apps) without model work
Resumes that list MES, SCADA, PLC, JPH, Databricks, and anomaly detection but never name a dataset, a finding, or a validation result
Those skills exist on the team or in partner teams. We need the person who works the data.
Responsibilities:
Frame manufacturing problems with plant and engineering partners; push back when labels, ground truth, or "accuracy" expectations are not real.
Explore, clean, and join fragmented operational data (machines, sensors, quality, maintenance, production, MES/historian extracts - you do not need to have used every acronym).
Build and validate statistical and machine-learning models for anomaly, quality, health, and process monitoring; report false positives/negatives and business cost, not only a leaderboard metric.
Hand usable outputs to engineers and operators (thresholds, explanations, "what to do when this fires"), and support models after they are in use.
Work with data engineering and software on pipelines, Databricks, and production - you are the customer of the platform, not the person hired to build it.
Required qualifications:
Bachelor's or master's in a quantitative or engineering field (data science, CS, statistics, industrial/mechanical/manufacturing engineering, OR, applied math, or related).
6+ years of applied data science (analysis, feature work, statistical or ML modeling on real operational or business datasets). Count data-science years, not total years in IT, DBA, or software engineering.
Strong Python and SQL; evidence of working large, messy tables - not only notebooks on clean extracts.
Production of models you can defend: classification, regression, clustering, anomaly detection, or time series, with a clear target and validation approach.
Experience creating features from machine, sensor, process, quality, maintenance, or other operational data (industrial preferred; high-volume ops data from adjacent domains is acceptable).
Comfort telling stakeholders when a model should not ship.
Ability to learn an unfamiliar plant process quickly.
Preferred qualifications:
Time in manufacturing, industrial IoT, semiconductor, automotive, aerospace, energy, or equipment-heavy operations.
Databricks, Spark/PySpark, or similar cloud analytics (we use Databricks; we do not require you to have been the lakehouse owner).
Familiarity with MLOps (tracking, monitoring, drift) as a partner to platform teams.
SPC, explainability, or prior work with historians /MES data.
Typical Day in the Role:
Compelling Story & Candidate Value Proposition | |
| Building AI application from ground up Impacting all of GM Manufacturing Not limited to any technologies |
Candidate Requirements | ||
Education - Bachelor's degree in a technical field such as computer science, computer engineering or related field required Years of experience at least 8-10 years of experience Must have experience with data ***very important*** Application AI platform skill set is a nice to have, not required This is a true Data Scientist role | ||
1 | Data modeling at least 8-10 years of experience | |
2 | Data pipeline at least 8-10 years of experience | |
3 | Data analytics able to build something out of messy data at least 8-10 years of experience | |
| Location | Austin, TX |
Browse dozens of recruiter approved resume templates, layouts and formats. Choose your favorite and make it your own in minutes.
Free resume templatesImprove your existing resume or start from scratch and create a standout, ATS-friendly resume. Add job-specific content, download and apply.
Free resume builder