Research Engineer, Interactive World Models NvidiaResearch Engineer, Interactive World ModelsSanta Clara, CAContributions to an open-source ML project or developer platform, such as implementing model support, improving performance, building tests and benchmarks, fixing difficult issues, writing documentation, or helping users adopt the technology. Advance the production-ready world model frontier by working with researchers on few-step distillation, causal or autoregressive generation, reward fine-tuning, action conditioning, and long-horizon spatiotemporal memory and consistency.
Software Engineer - Model Performance Systems BasetenSoftware Engineer - Model Performance SystemsSan Francisco, CaliforniaBy uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. You will not just be building the automated "speedometer and diagnostic" suite for our next-generation AI infrastructure; you will be defining the roadmap, driving key technical decisions, and taking full ownership of the future of this work.
Software Engineer - Model Products BasetenSoftware Engineer - Model ProductsSan Francisco, CaliforniaDesign, build, and operate the Model APIs surface with focus on advanced inference capabilities: structured outputs (JSON mode, grammar-constrained generation), tool/function calling and multi-modal serving. Profile and optimize TensorRT-LLM kernels, analyze CUDA kernel performance, implement custom CUDA operators, tune memory allocation patterns for maximum throughput and optimize communication patterns across multi-GPU setups.
Machine Learning Research Engineer, Model Evaluation WindBorne SystemsMachine Learning Research Engineer, Model EvaluationPalo Alto, CaliforniaWindBorne Systems is supercharging weather forecasts with a proprietary data source: a global constellation of next-generation smart weather balloons targeting critical atmospheric data. Evaluation strategy — Work with our Meteorology team to develop a rigorous, meteorologically valid strategy for comparing WeatherMesh with leading AI and physics-based models.
Engineering Manager - Model Performance BasetenEngineering Manager - Model PerformanceSan Francisco, CaliforniaBy uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. Drive the development and deployment of large-scale optimization techniques for various ML models, especially large language models (LLMs).
VP, Bioassays, Cell Models, And Genomic Screening InsitroVP, Bioassays, Cell Models, And Genomic ScreeningSouth San Francisco, CA$290,000–$326,000 / yearLead innovation: Provide strategic, technical, and operational leadership in the development, validation, and implementation of disease-relevant cell models, genomic screening (phenotypic pooled optical screening and arrayed screening), and supporting assay development for drug discovery. In this leadership role, your primary responsibilities will be to: Drive insitro's discovery engine: Set the strategic vision for cell models, image based and genomic screens, as well as bioassay development and execution across all therapeutic areas, fueling our target and drug discovery ML engine.
Software Engineer - Model Performance BasetenSoftware Engineer - Model PerformanceSan Francisco, CaliforniaBy uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. Implement, refine, and productionize cutting-edge techniques (quantization, speculative decoding, kv cache reuse, chunked prefill and LoRA) for ML model inference and infrastructure.
ML Engineer - Inference & Model Deployment HiringCafeML Engineer - Inference & Model DeploymentCupertino, CaliforniaYou will own the bridge between model development and real user-facing infrastructure: deploying models, optimizing inference latency and throughput, scaling serving systems, and making sure our models run efficiently in production. Implement optimization techniques such as quantization, pruning, batching, caching, efficient attention, and precision trade-offs while preserving model quality.
Senior Software Engineer - Model Performance InferenceSenior Software Engineer - Model PerformanceSan Francisco, CaliforniaYour work spans from implementing known optimization techniques to experimenting with novel approaches, always with the goal of serving models faster and cheaper at scale. If you love squeezing every last drop of performance out of GPUs, diving deep into CUDA kernels, and turning optimization techniques into production systems, we'd love to meet you.
Machine Learning: Multimodal Foundation Models The Bot CompanyMachine Learning: Multimodal Foundation ModelsSan Francisco, CaliforniaImprove Cross-Modal Reasoning: Research and implement methods to ensure the model doesn't just "associate" modalities but actually reasons through them (e.g., grounding visual physics in kinematic constraints). Ship and Iterate on Real Systems: Integrate models into real robotic stacks, build on robot code to deploy your models, and optimize performance for edge inference.
Applied Research Scientist - Foundation Models Ambient.aiApplied Research Scientist - Foundation ModelsRedwood City, CaliforniaPowered by Ambient Pulsar, the first reasoning Vision-Language Model purpose-built for physical security, our platform seamlessly integrates with existing security cameras and physical access control systems to unify monitoring, access control, threat assessment, response, and investigations through an always-on reasoning layer that augments security operators with superhuman capabilities. The momentum speaks for itself: we doubled new ARR in FY26, we process 200M+ video hours per day, and have delivered results for world-class customers including Cisco, ServiceNow, SentinelOne, TikTok, Bayer, and MoMA.
Product Designer, AI Models FIGMAProduct Designer, AI ModelsSan Francisco, CA$169,000–$303,000 / yearJob level and actual compensation will be decided based on factors including, but not limited to, individual qualifications objectively assessed during the interview process (including skills and prior relevant experience, potential impact, and scope of role), market demands, and specific work location. Figma offers equity to employees, as well as a competitive package of additional benefits, including health, dental, and vision coverage; retirement benefits with company contributions; parental leave and reproductive or family planning support; mental health and wellness benefits; and paid time off.
Research Scientist (Model Evaluation) SanasResearch Scientist (Model Evaluation)Palo Alto, CaliforniaDevelop reference-based and reference-free metrics calibrated to Sanas's specific model tasks: SI-SDR, PESQ, STOI, DNSMOS, speaker similarity, WER delta, COMET, and task-specific custom metrics where off-the-shelf measures fall short. Deep familiarity with speech and audio quality metrics — perceptual (MOS, MUSHRA, PESQ, STOI), signal-level (SI-SDR, SNR), and task-specific (WER, speaker similarity, DNSMOS) — and an understanding of when each is and isn't the right tool.
Research Scientist-Model Efficiency Bitdeer Technologies GroupResearch Scientist-Model Efficiencysan jose, CAA culture that values authenticity and diversity of thoughts and backgrounds; An inclusive and respectable environment with open workspaces and exciting start-up spirit; Fast-growing company with the chance to network with industrial pioneers and enthusiasts; Ability to contribute directly and make an impact on the future of the digital asset industry; Involvement in new projects, developing processes/systems; Personal accountability, autonomy, fast growth, and learning opportunities; Attractive welfare benefits and developmental opportunities such as training and mentoring. Genuine implementation-level depth in at least one area of model efficiency, such as quantization, sparsity and pruning, speculative decoding and MTP, or serving-time attention and KV-cache methods.
NewSenior Engineer, State Estimation and Modeling - (SJ2026DV) ArcherSenior Engineer, State Estimation and Modeling - (SJ2026DV)San Jose, CA$205,379–$215,647.95 / yearIntegrate state estimation solutions with various hardware components, including air data systems, Inertial Navigation Systems (INS), Global Navigation Satellite Systems (GNSS), and other aviation-grade sensors. Collaborate with cross-functional teams, including control laws, hardware, software, and test engineers, to solve complex aircraft-level problems and ensure seamless integration.
Member of Research Staff (Machine Learning for Neural Circuit Modeling Postdoc) EpistemeMember of Research Staff (Machine Learning for Neural Circuit Modeling Postdoc)San Francisco, California$120,000–$140,000 / yearWorking closely with both computational and experimental scientists, you will develop and deploy machine learning models for imaging data, neural activity traces, and predictive and causal modeling of neural circuits. Explore advanced deep learning methods (e.g., symbolic regression, GNNs, Transformers, generative models) to inform neural models from data and enhance interpretability.
Model Policy, Frontier Cyber Risk OpenAIModel Policy, Frontier Cyber RiskSan Francisco, CaliforniaFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. Our open-plan offices have height-adjustable desks, conference rooms, phone booths, well-stocked kitchens full of snacks and drinks, three in-house prepared meals daily, a private outdoor space for working in the sun or socializing, nap rooms, private bike storage, and more.
NewStaff ML Engineer, Search Ads Shopping Relevance Models Google LLCStaff ML Engineer, Search Ads Shopping Relevance ModelsMountain View, CAWe"re looking for engineers who bring fresh ideas from all areas, including information retrieval, distributed computing, large-scale system design, networking and data storage, security, artificial intelligence, natural language processing, UI design and mobile; the list goes on and is growing every day. Collaborate on user journey understanding, metric and label formulation, feature and model improvements, live traffic experiments, data analysis, tools and infrastructure, and more, to predict and improve user experience on search ads.
NewSoftware Developer Optics and Electromagnetic Modeling for Lithography Siemens AGSoftware Developer Optics and Electromagnetic Modeling for LithographySanta Clara, CA$90,000–$162,000 / yearResponsibilities Include:Simulate and theoretically analyze electromagnetic and optical physics applicable to lithography modelingUnderstand enhancement requests and bugs and discuss solutions with the teamImplement production ready software in C++Document new software in functional specificationsPerform detailed, quantitative analysis of performance and quality of the softwareCandidates must be able to:Understand electromagnetic and optical physics and how it applies to practical lithography modelingImplement object oriented software in C++Analyze data in Matlab or PythonWork comfortably in Linux and with Microsoft Office programs Word, Power Point, and OutlookCompile information into clear documents or slides and present to peers and managementNice to have:Experience in the semiconductor industry, specifically with lithography and optical proximity correction softwareBSEE, MSEE/ MS or Ph. Experience level Experienced Professional Job type Full-time Work mode Hybrid (Remote/Office) Employment type Permanent Location(s) Santa Clara - California - United States of America Wilsonville - Oregon - United States of America Siemens EDA is a global technology leader in Electronic Design Automation software.
ML Engineer, Apple Foundation Models AppleML Engineer, Apple Foundation ModelsCupertino, CAMinimum Qualifications** + Demonstrated expertise in LLM or Multi-modal LLM with a publication record in relevant conferences (e.g., NeurIPS, ICML, ICLR, CVPR, ICCV, ECCV, KDD, ACL, ICASSP, InterSpeech) or a track record in applying deep learning techniques to products + Proficient programming skills in Python and one of the deep learning toolkits such as JAX, PyTorch, or Tensorflow + Ability to work in a collaborative environment + Ph. You will work closely with researchers, engineers, and product teams to identify capability gaps, design data-centric solutions, and create high-quality training signals for reasoning, agentic behavior, multimodal understanding, tool use, and alignment.