Automation Engineer - Scientific Data, AI/ML Pipelines & Integration Dev

Zifo

  • Boston, MA
  • 2 days ago
    Want to know if you’re a fit?
    Upload your resume and let our AI show you.

    Skills

    • AWS Lambdaunmatched
    • Agile Programming Methodologiesunmatched
    • Amazon Elastic Compute Cloud (EC2)unmatched
    • Amazon Relational Database Service (RDS)unmatched
    • Amazon Simple Storage Service (S3)unmatched
    • Amazon Web Services (AWS)unmatched
    • Application Programming Interface (API)unmatched
    • Artificial Intelligence (AI)unmatched
    • Assaysunmatched
    • Atlassian JIRAunmatched
    • Automation Engineeringunmatched
    • Biologyunmatched
    • Biotech and Pharmaceuticalunmatched
    • Code Reviewsunmatched
    • Code of Federal Regulationsunmatched
    • Computer Scienceunmatched
    • Continuous Deployment/Deliveryunmatched
    • Continuous Integrationunmatched
    • Cross-Functionalunmatched
    • Data Analysisunmatched
    • Data Managementunmatched
    • Data Modelingunmatched
    • Data Qualityunmatched
    • Data Scienceunmatched
    • Database Designunmatched
    • Database Extract Transform and Load (ETL)unmatched
    • Debugging Skillsunmatched
    • Diversityunmatched
    • Djangounmatched
    • Dockerunmatched
    • Flaskunmatched
    • Gitunmatched
    • GitHubunmatched
    • GxPunmatched
    • Informaticsunmatched
    • Information/Data Security (InfoSec)unmatched
    • Insuranceunmatched
    • Jenkinsunmatched
    • Laboratory Information Management System (LIMS)unmatched
    • Laboratory Operationsunmatched
    • Manufacturingunmatched
    • Microservicesunmatched
    • MongoDBunmatched
    • MySQLunmatched
    • NoSQLunmatched
    • Onboardingunmatched
    • Oracleunmatched
    • PostgreSQLunmatched
    • Pytestunmatched
    • Python Programming/Scripting Languageunmatched
    • REST (Representational State Transfer)unmatched
    • Realtime Transport Protocolunmatched
    • Regulatory Complianceunmatched
    • Regulatory Requirementsunmatched
    • Research & Development (R&D)unmatched
    • SQL (Structured Query Language)unmatched
    • Scientific Data Management System (SDMS)unmatched
    • Scrum Project Management and Software Developmentunmatched
    • Software Developmentunmatched
    • Software Development Lifecycle (SDLC)unmatched
    • System Integration (SI)unmatched
    • Team Lead/Managerunmatched
    • Team Playerunmatched
    • Technical Deliveryunmatched
    • Technical Writingunmatched
    • Test Automationunmatched
    • Test Driven Development (TDD)unmatched
    • Test Plan/Scheduleunmatched
    • Training Data Setsunmatched

    Description

    Location: Boston, Indianapolis, RTP North Carolina & ChicagoZifo is a global specialist scientific and process informatics services company supporting life sciences, biotech, and pharmaceutical organizations. We enable digital transformation across R&D, manufacturing, and quality by delivering data-driven, scalable, and compliant software solutions.Zifo is seeking a passionate Software Developer who can work at the intersection of science, data, and technology. The role requires strong expertise in Benchling, Python, SQL/NoSQL, AWS and FastAPI, along with the ability to work directly with scientists performing assay-based experiments. The successful candidate will translate experimental workflows into robust data components, scientific system integrations, AI-enabled insights, and next-generation data pipelines.RequirementsCollaborate with scientists, assay teams, and lab operations to capture end-to-end assay and experimental workflows, from sample onboarding and execution through data ingestion, validation, and downstream analyticsTranslate scientific and operational requirements into well-defined functional, technical, and data requirements for laboratory platforms, system integrations, and next-generation data pipelinesDesign, develop, and maintain Python-based backend services, APIs, microservices, and data pipelines on AWS using FastAPI and supporting frameworks such as Flask or Django, including integrations with scientific systems such as Benchling, Signals, LIMS, ELN, CDS, and SDMS.Design and optimize SQL and NoSQL data models and build ETL/ELT and next-generation data pipelines to support structured, semi-structured, and high-volume scientific data, analytics, and AI/ML workloads, including dataset preparation, feature engineering, and model integration into pipelines and applications.Implement and maintain CI/CD pipelines for automated build, testing and deploymentEnsure solutions meet performance, data integrity, security, and regulatory compliance requirements (e.g., GxP, 21 CFR Part 11)Perform code reviews, debugging, and performance optimizationCoordinate across cross-functional and geographically distributed teams, managing dependencies and ensuring delivery alignmentCreate ready to deliver technical documentation and track deliverables using JIRA and ConfluenceRequired QualificationsBachelor's or master's degree in computer science, Engineering, Life Sciences with 3–8 years of hands‑on experience in Python development with FastAPIProficiency in SQL, including schema design, complex queries, and performance optimizationRelational databases such as PostgreSQL, MySQL, Oracle, AWS RDS/Aurora, NoSQL databases such as DynamoDB, MongoDB, or equivalentExperience with scientific data and laboratory informatics, including familiarity with Benchling or similar scientific data platforms ELN like Benchling, LIMS, ELN, SDMS, CDS, within the life sciences or pharmaceutical industry (Preferred)AWS experience, including S3, EC2, Lambda, Step Functions, RDS / Aurora, IAM, monitoring, and loggingProficiency with Git‑based collaborative development, including branch management, pull requests, code reviews, and integration with CI/CD pipelines (GitHub Actions, GitLab CI, Jenkins, AWS CodePipeline) to ensure reliable and traceable software deliveryHands‑on experience with Test‑Driven Development and Python testing frameworks such as pytest, unittest, and mocking librariesWorking knowledge of AI/ML concepts, including data preparation, feature engineering, model integration, and inference workflowsExposure to the data and ML libraries such as pandas, NumPy, and scikit‑learn (exposure to TensorFlow or PyTorch is a plus)Ability to design data models aligned to scientific and assay workflows & integrating scientific or enterprise systems and working directly with scientists or lab usersKnowledge of containerization (Docker) and modern deployment best practicesFamiliarity with Agile/Scrum & SDLC development methodologies & Solid understanding of REST APIs, microservices, and integration patternsStrong communication, stakeholder engagement, and cross‑team coordination skillsAdditional PreferencesWillingness to travel/ relocate based on project or business needsAbility to work in a fast‑paced, client‑focused environmentComfortable managing cross‑team coordination and dependency management, particularly across globally distributed teams and user groupsBenefitsWe offer a competitive compensation package including accrued vacation, medical, dental, vision, 401k with company matching, life insurance, and flexible spending accounts.Zifo is an equal opportunity employer, and we value diversity at our company. We do not discriminate on the basis of race, religion, color, national origin, gender, sexual orientation, age, marital status, veteran status, or disability status.#J-18808-Ljbffr

    Numbers & Facts

    LocationBoston, MA

    Similar Jobs

    See more jobs