Client: global biotech company Job: Junior Data Scientist Location: Remote, but candidates must live in California Duration: 1-year contract, open to extensions Pay: 40-45 per hour W2 Job Description:
You will build data pipelines that integrate structured and unstructured data into executive dashboards and analytical products. You will also prototype and productionize AI agents that synthesize information, identify trends, answer business questions, and support executive decision-making.
Responsibilities Include:
Design and maintain scalable data pipelines, curated datasets, and reporting layers using SQL, Python, Databricks, APIs, and cloud-based platforms.
Integrate and process structured and unstructured data from databases, enterprise systems, documents, presentations, collaboration platforms, and external sources.
Create trusted data models that support executive dashboards, scorecards, and recurring business reviews.
Partner with business and technical stakeholders to translate business questions into practical data, analytics, and AI solutions.
Prototype and deploy AI agents using large language models, retrieval-augmented generation, semantic search, and workflow orchestration.
Establish data-quality controls, AI evaluation methods, monitoring, access controls, and responsible-AI guardrails.
Apply software engineering best practices, including automated testing, version control, CI/CD, documentation, and production monitoring.
Communicate technical recommendations and business insights clearly while providing effective documentation and knowledge transfer.
Basic Qualifications:
Masters degree OR Bachelors degree and 2 years of experience
Preferred Qualifications:
Experience working within a cross-functional digital product team and collaborating with product managers, designers, engineers, data scientists, and business stakeholders to deliver and continuously improve data and AI products.
Advanced proficiency in SQL and Python for data engineering, automation, analytics, and AI development.
Hands-on experience with Databricks or a comparable platform, including ETL/ELT workflows, Delta Lake, Apache Spark or PySpark, and cloud data architecture.
Experience developing analytical data models and executive dashboards using Power BI, Tableau, or similar tools.
Experience processing unstructured data and developing AI agents using large language models, retrieval-augmented generation, embeddings, vector search, and tool calling.
Experience moving data and AI solutions from rapid prototype to secure, scalable, and monitored production deployment.
Familiarity with Azure, AWS, or Google Cloud, as well as Git, automated testing, CI/CD, and data governance practices.
Strong communication and data-storytelling skills, with the ability to work independently, manage ambiguity, and engage executive audiences.
Experience in pharmaceutical, biotechnology, healthcare, or another regulated environment is preferred.
Numbers & Facts
Location
Thousand Oaks, CA (Remote)
Salary
$40–$45 Per Hour
Skills
Access Controlunmatched
Amazon Web Services (AWS)unmatched
Analysis Skillsunmatched
Apache Sparkunmatched
Application Programming Interface (API)unmatched
Artificial Intelligence (AI)unmatched
Artificial Intelligence (AI) Agentsunmatched
Automationunmatched
Best Practicesunmatched
Biotech and Pharmaceuticalunmatched
Cloud Architectureunmatched
Cloud Computingunmatched
Communication Skillsunmatched
Continuous Deployment/Deliveryunmatched
Continuous Improvementunmatched
Continuous Integrationunmatched
Cross-Functionalunmatched
Data Analysisunmatched
Data Managementunmatched
Data Modelingunmatched
Data Processingunmatched
Data Qualityunmatched
Data Scienceunmatched
Data Setsunmatched
Database Extract Transform and Load (ETL)unmatched