

Insight Global
Senior Data Scientist
β - Featured Role | Apply direct with Data Freelance Hub
This role is for a Senior Data Scientist with a contract length of "unknown" and a pay rate of "unknown." It requires strong applied machine learning, Python, SQL skills, and 6+ years of experience, preferably in manufacturing or related industries. The work location is on-site.
π - Country
United States
π± - Currency
$ USD
-
π° - Day rate
432
-
ποΈ - Date
August 8, 2026
π - Duration
Unknown
-
ποΈ - Location
On-site
-
π - Contract
Unknown
-
π - Security
Unknown
-
π - Location detailed
Austin, Texas Metropolitan Area
-
π§ - Skills detailed
#Statistics #Cloud #Version Control #Logging #Regression #Storage #SQL (Structured Query Language) #Data Science #Anomaly Detection #PySpark #Model Deployment #Computer Science #Documentation #React #Spark (Apache Spark) #ML (Machine Learning) #Data Engineering #Supervised Learning #AI (Artificial Intelligence) #Python #Data Processing #Code Reviews #Monitoring #Clustering #Scala #Databricks #Deployment #Data Pipeline #Forecasting #Classification #"ETL (Extract #Transform #Load)" #Datasets #Mathematics #Data Quality
Role description
Senior Data Scientist
We are looking for a Senior Data Scientist to help build advanced analytics, machine learning, and AI solutions for manufacturing operations.
β’ This role will focus on using manufacturing factory data to detect anomalies, improve quality, reduce downtime, optimize throughput, and support reusable data models that connect fragmented manufacturing systems into a common intelligence layer.
β’ The ideal candidate has strong applied machine learning skills, practical experience working with complex operational data, and the ability to partner with manufacturing, data engineering, platform, and software teams to move analytical solutions toward production.
β’ This is not a pure research role. We are looking for someone who can move from problem framing to data understanding, model development, validation, stakeholder alignment, and production support. The candidate should be able to learn unfamiliar domains quickly, challenge assumptions constructively, and push back when requirements, data quality, or model expectations are not realistic.
β’ Manufacturing experience is strongly preferred, but we are also open to candidates from adjacent industrial, operations, quality, aerospace, semiconductor, supply chain, or equipment-heavy environments who can learn the manufacturing domain quickly.
Summary of Data Science Work in a Manufacturing Environment
β’ A Data Scientist in manufacturing works at the intersection of factory operations, engineering, quality, maintenance, data platforms, and machine learning.
β’ The work is not only about building models. It includes understanding how the plant operates, identifying where data is generated, defining what βnormalβ and βabnormalβ look like, creating reliable features from machine and process signals, validating model outputs against real-world outcomes, and delivering insights that plant teams can act on.
β’ Typical manufacturing data science work includes detecting process drift, identifying abnormal machine behavior, predicting quality issues, improving equipment health visibility, supporting root cause analysis, and helping teams move from reactive firefighting to proactive detection, triage, and prevention.
β’ Success requires technical depth, manufacturing curiosity, practical judgment, and the ability to build solutions that work with messy, incomplete, noisy, and high-frequency industrial data.
Key Responsibilities
Applied Machine Learning & Analytics
β’ Develop machine learning and statistical models to support manufacturing use cases such as anomaly detection, quality prediction, equipment health, process monitoring, throughput improvement, and decision support.
β’ Apply supervised, unsupervised, and semi-supervised learning methods, including classification, regression, clustering, anomaly detection, time-series analysis, statistical process control, and model explainability.
β’ Build anomaly detection solutions using methods such as control limits, isolation forests, clustering, Mahalanobis distance, autoencoders, time-series models, and supervised classification where labeled defects are available.
β’ Develop models for manufacturing use cases such as assembly issues, predictive maintenance, bottleneck detection, process optimization, and quality prediction.
β’ Evaluate model performance using appropriate metrics, ground truth definitions, validation strategies, false positive and false negative analysis, and business impact measures.
β’ Identify when data is insufficient, labels are unreliable, ground truth is weak, or a machine learning approach is not appropriate, and communicate those limitations clearly.
Manufacturing Data & Feature Engineering
β’ Analyze real-time and historical factory data from sources such as PLCs, sensors, machines, MES, SCADA, historians, quality systems, maintenance systems, production logs, and enterprise platforms.
β’ Create features from manufacturing signals such as cycle time, pressure, temperature, torque, vibration, current, force, cushion pressure, line speed, JPH, FTT, FRC, scrap, rework, downtime, and fault codes.
β’ Work with noisy, incomplete, high-frequency, or fragmented industrial data to create reliable analytical datasets.
β’ Partner with plant teams and domain experts to understand process behavior, validate assumptions, and determine whether model outputs reflect real operating conditions.
Cloud, Data Pipelines & MLOps
β’ Use cloud data platforms, preferably Databricks, to support scalable analytics and machine learning workflows.
β’ Develop and partner with Data Engineering to build data pipelines that ingest, transform, and prepare manufacturing data for analysis, modeling, monitoring, and reporting.
β’ Work with tools such as Cloud Storage, databricks, bigdata processing (pyspark, spark)
β’ Support real-time and near-real-time analytics use cases by working with streaming data or event-driven architectures.
β’ Partner with platform and software engineering teams to move models and analytical workflows from prototype to production-ready solutions.
β’ Follow MLOps practices such as experiment tracking, model versioning, model deployment, model monitoring, drift detection, retraining workflows, and production documentation.
β’ Monitor model performance after deployment, including false positives, false negatives, data drift, model drift, latency, uptime, pipeline failures, and changing manufacturing conditions.
Productization, Communication & Delivery
β’ Collaborate with data engineers, platform engineers, software engineers, manufacturing engineers, quality teams, and plant stakeholders to move data science prototypes into production-ready workflows.
β’ Follow software engineering best practices, including version control, modular code, code reviews, testing, logging, documentation, reusable packages, and reproducible environments.
β’ Document model logic, assumptions, input features, thresholds, limitations, operational dependencies, and recommended actions for business and plant-floor users.
β’ Distinguish between exploratory research, prototype development, and production-ready delivery, with focus on prototype development and production-ready delivery
Required Qualifications
β’ Bachelorβs or Masterβs degree in Data Science, Computer Science, Statistics, Industrial Engineering, Mechanical Engineering, Manufacturing Engineering, Operations Research, Applied Mathematics, or a related technical field.
β’ 6+ years of experience applying data science, machine learning, statistical modeling, optimization, or advanced analytics in a professional environment.
β’ Strong Python skills.
β’ Strong SQL skills and experience working with large, complex datasets.
β’ Experience with supervised and unsupervised machine learning methods, including classification, regression, clustering, anomaly detection, time-series analysis, forecasting, or process optimization.
β’ Experience building features from machine, sensor, process, quality, maintenance, production, or operational datasets.
β’ Experience working with cloud-based data and analytics platforms such as databricks.
β’ Experience working with data engineering, software engineering, or platform teams to move analytical solutions toward production.
β’ Understanding of MLOps concepts such as experiment tracking, model deployment, model monitoring, CI/CD, version control, testing, model registry, and retraining.
β’ Ability to work with noisy, incomplete, high-frequency, or fragmented operational data.
β’ Ability to communicate technical findings clearly to plant teams, engineers, leaders, and non-technical stakeholders.
β’ Ability to operate in ambiguous environments where requirements, data quality, and success criteria may need to be clarified.
β’ Professional confidence to challenge assumptions, push back constructively, and influence stakeholders with evidence.
β’ Demonstrated ability to learn new technical and business domains quickly.
Senior Data Scientist
We are looking for a Senior Data Scientist to help build advanced analytics, machine learning, and AI solutions for manufacturing operations.
β’ This role will focus on using manufacturing factory data to detect anomalies, improve quality, reduce downtime, optimize throughput, and support reusable data models that connect fragmented manufacturing systems into a common intelligence layer.
β’ The ideal candidate has strong applied machine learning skills, practical experience working with complex operational data, and the ability to partner with manufacturing, data engineering, platform, and software teams to move analytical solutions toward production.
β’ This is not a pure research role. We are looking for someone who can move from problem framing to data understanding, model development, validation, stakeholder alignment, and production support. The candidate should be able to learn unfamiliar domains quickly, challenge assumptions constructively, and push back when requirements, data quality, or model expectations are not realistic.
β’ Manufacturing experience is strongly preferred, but we are also open to candidates from adjacent industrial, operations, quality, aerospace, semiconductor, supply chain, or equipment-heavy environments who can learn the manufacturing domain quickly.
Summary of Data Science Work in a Manufacturing Environment
β’ A Data Scientist in manufacturing works at the intersection of factory operations, engineering, quality, maintenance, data platforms, and machine learning.
β’ The work is not only about building models. It includes understanding how the plant operates, identifying where data is generated, defining what βnormalβ and βabnormalβ look like, creating reliable features from machine and process signals, validating model outputs against real-world outcomes, and delivering insights that plant teams can act on.
β’ Typical manufacturing data science work includes detecting process drift, identifying abnormal machine behavior, predicting quality issues, improving equipment health visibility, supporting root cause analysis, and helping teams move from reactive firefighting to proactive detection, triage, and prevention.
β’ Success requires technical depth, manufacturing curiosity, practical judgment, and the ability to build solutions that work with messy, incomplete, noisy, and high-frequency industrial data.
Key Responsibilities
Applied Machine Learning & Analytics
β’ Develop machine learning and statistical models to support manufacturing use cases such as anomaly detection, quality prediction, equipment health, process monitoring, throughput improvement, and decision support.
β’ Apply supervised, unsupervised, and semi-supervised learning methods, including classification, regression, clustering, anomaly detection, time-series analysis, statistical process control, and model explainability.
β’ Build anomaly detection solutions using methods such as control limits, isolation forests, clustering, Mahalanobis distance, autoencoders, time-series models, and supervised classification where labeled defects are available.
β’ Develop models for manufacturing use cases such as assembly issues, predictive maintenance, bottleneck detection, process optimization, and quality prediction.
β’ Evaluate model performance using appropriate metrics, ground truth definitions, validation strategies, false positive and false negative analysis, and business impact measures.
β’ Identify when data is insufficient, labels are unreliable, ground truth is weak, or a machine learning approach is not appropriate, and communicate those limitations clearly.
Manufacturing Data & Feature Engineering
β’ Analyze real-time and historical factory data from sources such as PLCs, sensors, machines, MES, SCADA, historians, quality systems, maintenance systems, production logs, and enterprise platforms.
β’ Create features from manufacturing signals such as cycle time, pressure, temperature, torque, vibration, current, force, cushion pressure, line speed, JPH, FTT, FRC, scrap, rework, downtime, and fault codes.
β’ Work with noisy, incomplete, high-frequency, or fragmented industrial data to create reliable analytical datasets.
β’ Partner with plant teams and domain experts to understand process behavior, validate assumptions, and determine whether model outputs reflect real operating conditions.
Cloud, Data Pipelines & MLOps
β’ Use cloud data platforms, preferably Databricks, to support scalable analytics and machine learning workflows.
β’ Develop and partner with Data Engineering to build data pipelines that ingest, transform, and prepare manufacturing data for analysis, modeling, monitoring, and reporting.
β’ Work with tools such as Cloud Storage, databricks, bigdata processing (pyspark, spark)
β’ Support real-time and near-real-time analytics use cases by working with streaming data or event-driven architectures.
β’ Partner with platform and software engineering teams to move models and analytical workflows from prototype to production-ready solutions.
β’ Follow MLOps practices such as experiment tracking, model versioning, model deployment, model monitoring, drift detection, retraining workflows, and production documentation.
β’ Monitor model performance after deployment, including false positives, false negatives, data drift, model drift, latency, uptime, pipeline failures, and changing manufacturing conditions.
Productization, Communication & Delivery
β’ Collaborate with data engineers, platform engineers, software engineers, manufacturing engineers, quality teams, and plant stakeholders to move data science prototypes into production-ready workflows.
β’ Follow software engineering best practices, including version control, modular code, code reviews, testing, logging, documentation, reusable packages, and reproducible environments.
β’ Document model logic, assumptions, input features, thresholds, limitations, operational dependencies, and recommended actions for business and plant-floor users.
β’ Distinguish between exploratory research, prototype development, and production-ready delivery, with focus on prototype development and production-ready delivery
Required Qualifications
β’ Bachelorβs or Masterβs degree in Data Science, Computer Science, Statistics, Industrial Engineering, Mechanical Engineering, Manufacturing Engineering, Operations Research, Applied Mathematics, or a related technical field.
β’ 6+ years of experience applying data science, machine learning, statistical modeling, optimization, or advanced analytics in a professional environment.
β’ Strong Python skills.
β’ Strong SQL skills and experience working with large, complex datasets.
β’ Experience with supervised and unsupervised machine learning methods, including classification, regression, clustering, anomaly detection, time-series analysis, forecasting, or process optimization.
β’ Experience building features from machine, sensor, process, quality, maintenance, production, or operational datasets.
β’ Experience working with cloud-based data and analytics platforms such as databricks.
β’ Experience working with data engineering, software engineering, or platform teams to move analytical solutions toward production.
β’ Understanding of MLOps concepts such as experiment tracking, model deployment, model monitoring, CI/CD, version control, testing, model registry, and retraining.
β’ Ability to work with noisy, incomplete, high-frequency, or fragmented operational data.
β’ Ability to communicate technical findings clearly to plant teams, engineers, leaders, and non-technical stakeholders.
β’ Ability to operate in ambiguous environments where requirements, data quality, and success criteria may need to be clarified.
β’ Professional confidence to challenge assumptions, push back constructively, and influence stakeholders with evidence.
β’ Demonstrated ability to learn new technical and business domains quickly.






