ML Engineer - Inference & Model Deployment HiringCafeML Engineer - Inference & Model DeploymentCupertino, CaliforniaYou will own the bridge between model development and real user-facing infrastructure: deploying models, optimizing inference latency and throughput, scaling serving systems, and making sure our models run efficiently in production. Implement optimization techniques such as quantization, pruning, batching, caching, efficient attention, and precision trade-offs while preserving model quality.
Research Scientist (Model Evaluation) SanasResearch Scientist (Model Evaluation)Palo Alto, CaliforniaDevelop reference-based and reference-free metrics calibrated to Sanas's specific model tasks: SI-SDR, PESQ, STOI, DNSMOS, speaker similarity, WER delta, COMET, and task-specific custom metrics where off-the-shelf measures fall short. Deep familiarity with speech and audio quality metrics — perceptual (MOS, MUSHRA, PESQ, STOI), signal-level (SI-SDR, SNR), and task-specific (WER, speaker similarity, DNSMOS) — and an understanding of when each is and isn't the right tool.
Applied Research Scientist - Foundation Models Ambient.aiApplied Research Scientist - Foundation ModelsRedwood City, CaliforniaPowered by Ambient Pulsar, the first reasoning Vision-Language Model purpose-built for physical security, our platform seamlessly integrates with existing security cameras and physical access control systems to unify monitoring, access control, threat assessment, response, and investigations through an always-on reasoning layer that augments security operators with superhuman capabilities. The momentum speaks for itself: we doubled new ARR in FY26, we process 200M+ video hours per day, and have delivered results for world-class customers including Cisco, ServiceNow, SentinelOne, TikTok, Bayer, and MoMA.
Product Designer, AI Models FIGMAProduct Designer, AI ModelsSan Francisco, CA$169,000–$303,000 / yearJob level and actual compensation will be decided based on factors including, but not limited to, individual qualifications objectively assessed during the interview process (including skills and prior relevant experience, potential impact, and scope of role), market demands, and specific work location. Figma offers equity to employees, as well as a competitive package of additional benefits, including health, dental, and vision coverage; retirement benefits with company contributions; parental leave and reproductive or family planning support; mental health and wellness benefits; and paid time off.
Senior Director, Device And Spice Modeling NvidiaSenior Director, Device And Spice ModelingSanta Clara, CAThe "Device & SPICE modeling" group is responsible to co-develop with foundries in advanced device technology in the following three areas: Achieve device performance targets in Speed, Leakage, and Variation, Release accurate SPICE model based on test chip data, Tape-out device/RO test structures in test chip for SPICE model validation, process readiness & product scribe line monitors. The Advanced Technology Group (ATG) at NVIDIA is an organization of process, CAD, and design engineers that works closely with key foundry partners and internal design groups.
Staff Software Engineer, Foundation Model Inference DataBricksStaff Software Engineer, Foundation Model InferenceSan Francisco, CA$190,000–$265,000 / yearThe impact you will have: Build LLM infrastructure powering large-scale inference workloads for customers through partner models (OpenAI, Anthropic, Gemini) and self-hosted models (Qwen, GPT-OSS, Llama). More than 20,000 organizations worldwide - including adidas, AT&T, Bayer, Block, Mastercard, Rivian, Unilever, and 70% of the Fortune 500 - rely on the Databricks Data + AI Platform to build and scale data and AI apps, analytics and agents.
Senior Software Engineer, Model Serving DataBricksSenior Software Engineer, Model ServingSan Francisco, CA$166,000–$225,000 / yearContribute directly to key components across the serving infrastructure - from model container builds and deployment workflows to runtime systems like routing, caching, observability, and intelligent autoscaling - ensuring smooth and efficient operations at scale. You will design and build systems that enable high-throughput, low-latency inference across CPU and GPU workloads, influence architectural direction, and collaborate closely across platform, product, infrastructure, and research teams to deliver a world-class serving platform.
Staff Software Engineer, Model Serving DataBricksStaff Software Engineer, Model ServingSan Francisco, CA$192,000–$260,000 / yearYou will design and build systems that enable high-throughput, low-latency inference across CPU and GPU workloads, influence architectural direction, and collaborate closely across platform, product, infrastructure, and research teams to deliver a world-class serving platform. Contribute directly to key components across the serving infrastructure - from model container builds and deployment workflows to runtime systems like routing, caching, observability, and intelligent autoscaling - ensuring smooth and efficient operations at scale.
Software Engineer - Voice Model TwitterSoftware Engineer - Voice ModelPalo Alto, CA$150,000–$450,000 / yearWork on pre-training and post-training of speech-language models, with targeted enhancements through supervised fine-tuning, reinforcement learning, and other techniques to ensure Grok Voice responses are accurate, factually grounded, natural and idiomatic in spoken style, conversational in tone, and fluent across multiple languages. Build and iterate a comprehensive evaluation framework covering objective metrics (accuracy, quality, latency, expressiveness), human preference studies, content factuality assessments, real-time interaction quality, and experimentation infrastructure to measure and improve performance.
Research Scientist / Engineer – Foundation Model: Core Research LumaResearch Scientist / Engineer – Foundation Model: Core ResearchRedwood City, CaliforniaUnified Modeling & Efficiency Drive the core research that powers all of Luma's products — co-designing multimodal representations, advancing core algorithms for long-context training, and establishing rigorous scaling laws to predict performance across compute budgets. This role offers the chance to bridge frontier research with magical, shipped products like Dream Machine and Ray3, solving novel problems where no playbook exists.
Research Engineer / Research Scientist - Personal Agi, Personality And Model Behavior OpenAIResearch Engineer / Research Scientist - Personal Agi, Personality And Model BehaviorSan Francisco, CAFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. We're looking for individuals with strong ML engineering skills and research experience, especially with novel and highly capable models, and in areas like reinforcement learning and reward modeling.
ML Research Scientist, Foundation Models (Senior / Staff / Principal) Genesis TherapeuticsML Research Scientist, Foundation Models (Senior / Staff / Principal)San Mateo, CAOur generative and predictive AI platform, GEMS (Genesis Exploration of Molecular Space), integrates AI and physics into industry-leading models to generate and optimize drug molecules, including the breakthrough generative diffusion model Pearl for structure prediction. Genesis has also signed category-leading AI-pharma deals, the most recent of which was a significant expansion with Incyte (see coverage in Forbes and GEN) with a total potential deal value of several billion dollars.
Research Engineer / Research Scientist - Personal AGI, Personality and Model Behavior OpenAIResearch Engineer / Research Scientist - Personal AGI, Personality and Model BehaviorSan Francisco, CaliforniaFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. We're looking for individuals with strong ML engineering skills and research experience, especially with novel and highly capable models, and in areas like reinforcement learning and reward modeling.
Research Engineer/Research Scientist- Personal AGI, Model Experience OpenAIResearch Engineer/Research Scientist- Personal AGI, Model ExperienceSan Francisco, CaliforniaFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity.
Research Engineer/Research Scientist- Personal Agi, Model Experience OpenAIResearch Engineer/Research Scientist- Personal Agi, Model ExperienceSan Francisco, CAFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity.
Product Manager, Model Gateway MercorProduct Manager, Model GatewaySan Francisco, CaliforniaOwn the LLM gateway end-to-end : Set the roadmap and policy for the shared layer that routes every product surface to the models it depends on, from priority queueing to cost attribution to model access, and own its lifecycle from conception through execution. You'll own this product area with significant autonomy, working directly with an infrastructure team that owns the underlying shared services and with leadership to turn ambiguous, cross-cutting tradeoffs into durable systems and clear policy.
Senior Software Engineer, Data & Model DexmateSenior Software Engineer, Data & ModelFremont, CaliforniaOur mission is to democratize robotics by lowering the barrier to entry, delivering a plug-and-play platform for developers, researchers, and enterprises, and cultivating an open ecosystem that accelerates the evolution of physical AI. Today, robotics is fragmented, slow, and closed: most builders are forced to reinvent the same stack again and again, and most ideas never make it past the prototype stage.
Technical Program Manager, Model Alignment And Deployment Character AITechnical Program Manager, Model Alignment And DeploymentRedwood City, CATogether, these groups are responsible for transforming powerful pretrained language models into intelligent, engaging, safely aligned, and highly scalable products-working across data, compute, algorithms, infrastructure, and user insights to improve model performance and ensure reliable delivery. Program ownership: Lead planning and execution of cross-functional programs spanning data collection, annotation pipelines, alignment workflows (RLHF, DPO, Constitutional AI), safety guardrails (adversarial testing, red-teaming), and model serving.
NewProduct Manager, Model Management TapestryProduct Manager, Model ManagementMountain View, CaliforniaIn this role, you will collaborate across technical and commercial teams and alongside subject-matter experts to define, build, and deliver essential grid network model management capabilities to both internal and external partners. Originally born at X, Alphabet’s moonshot factory, Tapestry brings together experts in energy, AI, software engineering, and products to build tools that help the electricity ecosystem plan smarter, move faster, and operate more efficiently.
Research Engineering Manager - Model Training PerplexityResearch Engineering Manager - Model TrainingSan Francisco, CaliforniaExperience leading or managing research or engineering teams working on large-scale AI model development, including driving complex projects from idea to production. Lead a team of researchers and engineers focused on training SotA models for Perplexity-relevant use cases, leveraging the latest supervised and reinforcement learning techniques.