Research Engineer - Language Model Pre-Training

Zyphra Technologies

  • San Francisco, CA
  • 30+ days ago
    Want to know if you’re a fit?
    Upload your resume and let our AI show you.

    Skills

    • Artificial Intelligence (AI)unmatched
    • Communication Skillsunmatched
    • Computer Scienceunmatched
    • Data Managementunmatched
    • Data Modelingunmatched
    • Data Processingunmatched
    • Data Setsunmatched
    • Flexible Spending Accountsunmatched
    • GPU (Graphics Processing Unit)unmatched
    • Machine Learningunmatched
    • Mathematicsunmatched
    • Modeling Languagesunmatched
    • Parallel Computingunmatched
    • Performance Tuning/Optimizationunmatched
    • Physicsunmatched
    • Team Playerunmatched
    • Vision Planunmatched

    Description

    Zyphra is an artificial intelligence company based in San Francisco, California.

    The Role:

    As a Research Engineer - Language Model Pre-Training, youll shape our language model roadmap through end-to-end pretraining development. You will work extremely closely with our pretraining team, who will integrate your insights into our next-generation models.

    Youll Work Across:

    • Large-scale training runs and model parallelization

    • Performance optimization of our pretraining stack

    • Dataset collection, processing, and evaluation

    • Architecture and methodology research, including optimizer ablations

    What Were Looking For / Requirements:

    • Strong engineering aptitude for rapidly implementing reliable and robust systems

    • Can rapidly learn new fields and are excited to implement new ideas

    • Excellent communication and collaboration skills, and can work effectively on both research and engineering implementation at scale

    Qualifications / Additional Skills:

    • Deep expertise and intuition for solving machine learning problems and training models

    • Experience with training on large-scale (multi-node) GPU clusters

    • Deep understanding of model training pipelines - including model/data parallelism, distributed optimizers, etc.

    • Strong grasp of proper experimental methodology for running rigorous ablations and other hypothesis testing

    • Understanding of large-scale, highly parallel data processing pipelines

    • High proficiency with PyTorch and Python.

    • Strong ability to dive into large pre-existing codebases and rapidly get up to speed

    • Published machine learning research in well-respected venues is a plus

    • Postgraduate degree in a scientific subject (Computer Science, EE/EECS, Math, Physics)

    Why Work at Zyphra:

    • Our research methodology is to make grounded, methodical steps toward ambitious goals. Both deep research and engineering excellence are equally valued

    • We strongly value new and crazy ideas and are very willing to bet big on new ideas

    • We move as quickly as we can; we aim to minimize the bar to impact as low as possible

    • We all enjoy what we do and love discussing AI

    Benefits and Perks:

    • Comprehensive medical, dental, vision, and FSA plans

    • Competitive compensation and 401(k) plan

    • Relocation and immigration support on a case-by-case basis

    • In-office snacks and meals provided

    • Unlimited PTO and company holidays

    • In-person team in San Francisco with a collaborative, high-energy environment

    Numbers & Facts

    LocationSan Francisco, CA

    Similar Jobs

    • Jobot logo

      Principal Engineer Jobot

      Burlingame, CA7 days ago
      • $230,000–$250,000 Per Year
      • Instant Apply
    • Jobot logo

      Quant Engineer Jobot

      San Francisco, CA7 days ago
      • $300,000–$375,000 Per Year
      • Instant Apply
    • Jobot logo
      New!

      Senior AI Engineer Jobot

      San Francisco, CA4 days ago
      • $225,000–$280,000 Per Year
      • Instant Apply
    • Jobot logo
      New!

      Senior Software Engineer Jobot

      San Francisco, CA3 days ago
      • $225,000–$450,000 Per Year
      • Instant Apply
    See more jobs