Apolis logo

Data Scientist

Apolis
  • Raleigh, NC
  • $1–$59 Per Hour
  • Quick Apply
30+ days ago

Job Description

Data Scientist
Raleigh, North Carolina (North Hills) (4 days onsite)
6+ months CTH
Pay - $55-58/Hour on W2

Ideal Candidate Profile Summary:
Must have: Python & Pyspark experience
Preferred: AWS, GCP or other machine learning certifications
Preferred: XGBoos, Timeseries, Pytorch, or Tensorflow experience

Remote or On-site: This position is 4 days in office, 1 day remote per week, based at our corporate headquarters in Raleigh, North Carolina (North Hills) Time Zone Requirements.
Interview Process: MS Teams interviews scheduled through Beeline.
__________________________________________________


Data Scientist
We are seeking an experienced Data Scientist with strong expertise in Data Science, machine learning engineering with hands on experience in designing and deploying ML solutions in production. This role focuses on building scalable ML solutions, productionizing models, and enabling robust ML platforms for enterprise-grade deployments.
This position is 4 days in office, 1 day remote per week, based at our corporate headquarters in Raleigh, North Carolina (North Hills)
Key Responsibilities
" Build ML Models: Design and implement predictive and prescriptive models for regression, classification, and optimization problems.Apply advanced techniques such as structural time series modeling and boosting algorithms (e.g., XGBoost, LightGBM).
" Train and Tune Models: Develop and tune machine learning models using Python, PySpark, TensorFlow, and PyTorch.
" Collaboration & Communication: Work closely with stakeholders to understand business challenges and translate them into data science solutions and work in the end-to-end solutioning. Collaborate with cross-functional teams to ensure successful integration of models into business processes.
" Monitoring & Visualization: Rapidly prototype and test hypotheses to validate model approaches. Build automated workflows for model monitoring and performance evaluation. Create dashboards using tools like Databricks and Palantir to visualize key model metrics like model drift, Shapley values etc.
" Productionize ML: Build repeatable paths from experimentation to deployment (batch, streaming, and low-latency endpoints), including feature engineering, training, evaluation,
" Own ML Platform: Stand up and operate core platform components model registry, feature store, experiment tracking, artifact stores, and standardized CI/CD for ML.
" Pipeline Engineering: Author robust data/ML pipelines (orchestrated with Step Functions / Airflow / Argo) that train, validate, and release models on schedules or events.
" Observability & Quality: Implement end-to-end monitoring, data validation, model/drift checks, and alerting SLA/SLOs.
" Governance & Risk: Enforce model/version lineage, reproducibility, approvals, rollback plans, auditability, and cost controls aligned to enterprise policies.
" Partner & Mentor: Collaborate with on-shore/off-shore teams; coach data scientists on packaging, testing, and performance; contribute to standards and reviews.
" Hands-on Delivery: Prototype new patterns; troubleshoot production issues across data, model, and infrastructure layers.

Required Qualifications
" Education: Bachelor s degree in Computer Science, Information Technology, Data Science, or related field.
" Programming: 5+ years experience with Python (pandas, PySpark, scikit-learn; familiarity with PyTorch/TensorFlow helpful), bash, experience with Docker.
" ML Experimentation: Design and implement predictive and prescriptive models for regression, classification, and optimization problems. Apply advanced techniques such as structural time series modeling and boosting algorithms (e.g., XGBoost, LightGBM).
" ML Tooling: 5+ years experience with SageMaker (training, processing, pipelines, model registry, endpoints) or equivalents (Kubeflow, MLflow/Feast, Vertex, Databricks ML).
" Pipelines & Orchestration: 5+ years experience with Databricks DABS or Airflow or Step Functions, e-driven designs with EventBridge/SQS/Kinesis.
" Cloud Foundations: 3+ years experience with AWS/Azure/GCP on various services like ECR/ECS, Lambda, API Gateway, S3, Glue/Athena/EMR, RDS/Aurora (PostgreSQL/MySQL), DynamoDB, CloudWatch, IAM, VPC, WAF.
" Snowflake Foundations: Warehouses, databases, schemas, stages, Snowflake SQL, RBAC, UDF, Snowpark.
" CI/CD: 3+ years hands-on experience with CodeBuild/Code Pipeline or GitHub Actions/GitLab; blue/green, canary, and shadow deployments for models and services.
" Feature Pipelines: Proven experience with batch/stream pipelines, schema management, partitioning, performance tuning; parquet/iceberg best practices.
" Testing & Monitoring: Unit/integration tests for data and models, contract tests for features, reproducible training; data drift/performance monitoring.
" Operational Mindset: Incident response for model services, SLOs, dashboards, runbooks; strong debugging across data, model, and infra layers.
" Soft Skills: Clear communication, collaborative mindset, and a bias to automate & document.

Additional Qualification:
" Experience in retail/manufacturing is preferred.
Data Scientist
Role Summary
We are seeking an experienced Data Scientist with strong expertise in Data Science, machine learning engineering with hands on experience in designing and deploying ML solutions in production. This role focuses on building scalable ML solutions, productionizing models, and enabling robust ML platforms for enterprise-grade deployments.
This position is 4 days in office, 1 day remote per week, based at our corporate headquarters in Raleigh, North Carolina (North Hills)

Key Responsibilities
" Build ML Models: Design and implement predictive and prescriptive models for regression, classification, and optimization problems.Apply advanced techniques such as structural time series modeling and boosting algorithms (e.g., XGBoost, LightGBM).
" Train and Tune Models: Develop and tune machine learning models using Python, PySpark, TensorFlow, and PyTorch.
" Collaboration & Communication: Work closely with stakeholders to understand business challenges and translate them into data science solutions and work in the end-to-end solutioning. Collaborate with cross-functional teams to ensure successful integration of models into business processes.
" Monitoring & Visualization: Rapidly prototype and test hypotheses to validate model approaches. Build automated workflows for model monitoring and performance evaluation. Create dashboards using tools like Databricks and Palantir to visualize key model metrics like model drift, Shapley values etc.
" Productionize ML: Build repeatable paths from experimentation to deployment (batch, streaming, and low-latency endpoints), including feature engineering, training, evaluation,
" Own ML Platform: Stand up and operate core platform components model registry, feature store, experiment tracking, artifact stores, and standardized CI/CD for ML.
" Pipeline Engineering: Author robust data/ML pipelines (orchestrated with Step Functions / Airflow / Argo) that train, validate, and release models on schedules or events.
" Observability & Quality: Implement end-to-end monitoring, data validation, model/drift checks, and alerting SLA/SLOs.
" Governance & Risk: Enforce model/version lineage, reproducibility, approvals, rollback plans, auditability, and cost controls aligned to enterprise policies.
" Partner & Mentor: Collaborate with on-shore/off-shore teams; coach data scientists on packaging, testing, and performance; contribute to standards and reviews.
" Hands-on Delivery: Prototype new patterns; troubleshoot production issues across data, model, and infrastructure layers.

Required Qualifications
" Education: Bachelor s degree in Computer Science, Information Technology, Data Science, or related field.
" Programming: 5+ years experience with Python (pandas, PySpark, scikit-learn; familiarity with PyTorch/TensorFlow helpful), bash, experience with Docker.
" ML Experimentation: Design and implement predictive and prescriptive models for regression, classification, and optimization problems. Apply advanced techniques such as structural time series modeling and boosting algorithms (e.g., XGBoost, LightGBM).
" ML Tooling: 5+ years experience with SageMaker (training, processing, pipelines, model registry, endpoints) or equivalents (Kubeflow, MLflow/Feast, Vertex, Databricks ML).
" Pipelines & Orchestration: 5+ years experience with Databricks DABS or Airflow or Step Functions, e-driven designs with EventBridge/SQS/Kinesis.
" Cloud Foundations: 3+ years experience with AWS/Azure/GCP on various services like ECR/ECS, Lambda, API Gateway, S3, Glue/Athena/EMR, RDS/Aurora (PostgreSQL/MySQL), DynamoDB, CloudWatch, IAM, VPC, WAF.
" Snowflake Foundations: Warehouses, databases, schemas, stages, Snowflake SQL, RBAC, UDF, Snowpark.
" CI/CD: 3+ years hands-on experience with CodeBuild/Code Pipeline or GitHub Actions/GitLab; blue/green, canary, and shadow deployments for models and services.
" Feature Pipelines: Proven experience with batch/stream pipelines, schema management, partitioning, performance tuning; parquet/iceberg best practices.
" Testing & Monitoring: Unit/integration tests for data and models, contract tests for features, reproducible training; data drift/performance monitoring.
" Operational Mindset: Incident response for model services, SLOs, dashboards, runbooks; strong debugging across data, model, and infra layers.
" Soft Skills: Clear communication, collaborative mindset, and a bias to automate & document.
Additional Qualification:
" Experience in retail/manufacturing is preferred.

Numbers & Facts

LocationRaleigh, NC
IndustryComputer/IT Services
Salary$1–$59 Per Hour
Company Size500 to 999 employees
Websitehttps://www.apolisrises.com/

Benefits

Paid Sick Days, Employee Referral Program, Employee Events, Retirement / Pension Plans

About Company

Since 1996, RJT has provided successful SAP, Oracle, and IT consulting solutions and staffing services to clients around the world. The new Apolis brings you the same personalized service fortified with a greater array of IT solutions, global expertise, and cost-management strategies.

We are a global IT consultancy that seamlessly integrates experts and leading-edge solutions into your organization so you can focus on what really matters.

Skills

  • AWS Lambdaunmatched
  • Algorithmsunmatched
  • Amazon Simple Storage Service (S3)unmatched
  • Amazon Web Services (AWS)unmatched
  • Application Programming Interface (API)unmatched
  • Bash Scriptingunmatched
  • Best Practicesunmatched
  • Business Modelunmatched
  • Business Processesunmatched
  • Calendar Managementunmatched
  • Cloud Computingunmatched
  • Coachingunmatched
  • Communication Skillsunmatched
  • Computer Scienceunmatched
  • Continuous Deployment/Deliveryunmatched
  • Continuous Integrationunmatched
  • Cost Controlunmatched
  • Cross-Functionalunmatched
  • Data Managementunmatched
  • Data Modelingunmatched
  • Data Qualityunmatched
  • Data Scienceunmatched
  • Data Warehousingunmatched
  • Database Technologyunmatched
  • Debugging Skillsunmatched
  • Dockerunmatched
  • Electronic Medical Recordsunmatched
  • GCP (Good Clinical Practices)unmatched
  • GitHubunmatched
  • Identify Issuesunmatched
  • Incident Responseunmatched
  • Information Technology & Information Systemsunmatched
  • Integration Testingunmatched
  • Machine Learningunmatched
  • Machine Toolunmatched
  • Manufacturingunmatched
  • Mentoringunmatched
  • Metricsunmatched
  • Microsoft Windows Azureunmatched
  • Model Reviewunmatched
  • Model Validationunmatched
  • MySQLunmatched
  • Offshoringunmatched
  • Performance Analysisunmatched
  • Performance Modelingunmatched
  • Performance Reviewsunmatched
  • Performance Testingunmatched
  • Performance Tuning/Optimizationunmatched
  • PostgreSQLunmatched
  • Predictive Modelingunmatched
  • Process Modelingunmatched
  • Prototypingunmatched
  • Python Programming/Scripting Languageunmatched
  • Reporting Dashboardsunmatched
  • Retailunmatched
  • Risk Modelingunmatched
  • SQL (Structured Query Language)unmatched
  • Sales Pipelineunmatched
  • Scalable System Developmentunmatched
  • Service Level Agreement (SLA)unmatched
  • Simple Queue Service (SQS)unmatched
  • Snowflake Schemaunmatched
  • Team Playerunmatched
  • Testingunmatched
  • Unit Testunmatched
  • Writing Skillsunmatched

Be found by employers

5,500+ employers search our resume database daily. Add yours to get found by recruiters looking for candidates like you.

Level up your application

Professional resume templates

Browse dozens of recruiter approved resume templates, layouts and formats. Choose your favorite and make it your own in minutes.

Free resume templates

Free resume builder

Improve your existing resume or start from scratch and create a standout, ATS-friendly resume. Add job-specific content, download and apply.

Free resume builder