Workload / Performance Model Lead SiFiveWorkload / Performance Model LeadBerkeley, CA$231,444–$282,876 / yearSiFive's unrivaled compute platforms are continuing to enable leading technology companies around the world to innovate, optimize and deliver the most advanced solutions of tomorrow across every market segment of chip design, including artificial intelligence, machine learning, automotive, data center, mobile, and consumer. Any offer of employment for this position is also contingent on the Company verifying that you are a authorized for access to export-controlled technology under applicable export control laws or, if you are not already authorized, our ability to successfully obtain any necessary export license(s) or other approvals.
Software Engineer, Model Hardware CoDesign Tesla IncSoftware Engineer, Model Hardware CoDesignPalo Alto, CADesign train and iterate on neural network architectures for autonomous driving and robotics with a focus on efficiency-aware model design architecture search distillation pruning quantization-aware training. Our team designs trains and deploys large-scale neural networks optimized for inference on compute-constrained edge devices CPU GPU custom AI ASIC.
Human Interactive Driving Intern - World Models Toyota Research InstituteHuman Interactive Driving Intern - World ModelsLos Altos, CA$45–$65 / hourDemonstrated experience with one or more of the following: World models (e.g., latent dynamics, diffusion-based models), Model-based RL or decision-making, 3D perception or sensor fusion, and Large-scale simulation for robotics or autonomous systems. Your work may focus on developing novel components of world models, improving decision-making through model-based RL, advancing 3D perception for dynamic scenes, or enhancing simulation-to-reality transfer.
Member of Technical Staff — Diffusion Model RadixArkMember of Technical Staff — Diffusion ModelPalo Alto, CaliforniaRadixArk is an infrastructure-first company built by engineers who've shipped production AI systems, created SGLang (30K+ GitHub stars, the fastest open LLM serving engine), and developed Miles (our large-scale RL framework). Our team has optimized kernels serving billions of tokens daily, designed distributed training systems coordinating 10,000+ GPUs, and contributed to infrastructure that powers leading AI companies and research labs.
NewProduct Manager, World Modeling and Interactive Media, DeepMind Google LLCProduct Manager, World Modeling and Interactive Media, DeepMindSan Francisco, CAAt DeepMind, we are a pioneering AI lab with exceptional interdisciplinary teams focused on advancing AI development to solve complex global challenges and accelerate high-quality product innovation for billions of users. You will develop a strong understanding of the multimodal AI landscape with a deep focus on world modeling and interactive media and translate product opportunities and complex technical concepts into clear, actionable product plans.
Research Engineer / Research Scientist - Personal AGI, Personality and Model Behavior OpenAI LLCResearch Engineer / Research Scientist - Personal AGI, Personality and Model BehaviorSan Francisco, CAFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. Were looking for individuals with strong ML engineering skills and research experience, especially with novel and highly capable models, and in areas like reinforcement learning and reward modeling.
Research Engineer/Research Scientist, Personal AGI-Model Experience OpenAI LLCResearch Engineer/Research Scientist, Personal AGI-Model ExperienceSan Francisco, CAFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity.
AI Systems, Model Optimization Unconventional AIAI Systems, Model OptimizationPalo Alto, CASystems Fluency: Demonstrated ability to map state-of-the-art AI model architectures (e.g., Transformers, Mixture of Experts, diffusion models) to system performance implications and apply advanced efficiency techniques such as sparsity, quantization, and distillation. Cross-Functional Collaboration: Act as a translator between AI model architects and hardware/infrastructure engineering teams, converting model requirements into concrete specifications and codifying learnings for tapeouts.
Member of Technical Staff - Language & Reasoning Models Unconventional AIMember of Technical Staff - Language & Reasoning ModelsPalo Alto, CAModel Development: Design, train, and scale next-generation language and reasoning architectures (such as transformers, state space models, diffusion/flow models, and deep equilibrium models) specifically tailored for unconventional compute. Evaluation & Scaling: Establish the training recipes, loss functions, and evaluation metrics needed to reach the frontier of language comprehension, logical reasoning, and generation speed while maintaining the massive energy efficiency of our platform.
Senior Machine Learning Scientist, Multimodal & Relational Foundation Models Altos Labs IncSenior Machine Learning Scientist, Multimodal & Relational Foundation ModelsSan Francisco, CA$270,600–$330,000 / yearArchitect and implement novel hybrid models that integrate Large Language Models (LLMs) with Graph Neural Networks (GNNs) for multi-hop reasoning over biological knowledge graphs. • Lead the design of efficient data loading strategies and distributed training recipes (e.g., FSDP, DeepSpeed) to train models across multiple GPU nodes.
Model Engineer - Member of Technical Staff MeterModel Engineer - Member of Technical StaffSan Francisco, CaliforniaEnd-to-end ownership : You'll ship models into production networks, collaborate with firmware and application teams sitting next to you, and rapidly iterate in the wild. Now, we’re assembling a founding core engineering team to build and train models that understand these systems, optimize operations, anticipate failures, and repair issues before humans even notice them.
Researcher, World Models MenloResearcher, World ModelsSan Francisco, CaliforniaConversant, ideally deep, in several of: SSL for visual and sensor representations; world models (JEPA, V-JEPA, I-JEPA, LeJEPA, MJEPA); generative and predictive architectures (diffusion, DiT, flow matching, VAEs); robotics ML (VLA, inverse dynamics, sim-to-real, optical flow); sensor fusion (vision, proprioception, force/torque, multi-modal encoders); PyTorch, JAX, and distributed training. Advance our self-supervised learning stack for visual and sensor representations, building on and extending the JEPA family (V-JEPA, I-JEPA, and related predictive-embedding approaches).
Senior Perception Engineer, Obstacle Foundation Models - Autonomous Vehicles NVIDIA CorpSenior Perception Engineer, Obstacle Foundation Models - Autonomous VehiclesSanta Clara, CAHands-on experience architecting and deploying DNN-based perception pipelines on embedded or real-time platforms, including optimization for latency, memory, and compute constraints, and experience with modern architectures such as CNNs and transformers, plus familiarity with techniques like large-scale pretraining, parameter-efficient fine-tuning (e.g., LoRA), or vision-language models (VLMs). Contribute to the data strategy for perception: specify data and labeling requirements, help prioritize data collection and annotation, and collaborate with data and ground-truth teams, including model-assisted workflows (e.g., active learning, auto-labeling, vision-language models (VLMs)) and model-in-the-loop tooling.
Senior Vision Language Model Engineer NVIDIA CorpSenior Vision Language Model EngineerSanta Clara, CAFluent with agentic AI workflows across the full applied research lifecycle, including prototyping novel algorithms and search pipelines, benchmarking, and integrating prototypes in production codebases. Excellent experience training and deploying deep learning models on real‑world datasets: data preprocessing, distributed training, evaluation, debugging, and iterative improvement.
Product Manager - Business Model Strategy AdobeProduct Manager - Business Model StrategySan Francisco, California$141,700–$205,150 / yearAdobe’s industry-leading offerings including Adobe Acrobat Studio, Adobe Express, Adobe Firefly, Creative Cloud, Adobe Experience Platform, Adobe Experience Manager, and GenStudio enable people and businesses to turn ideas into impact, powered by AI and driven by human ingenuity. This particular role will focus on opportunities within the Firefly Enterprise (FFE) portfolio of products, which focuses on providing enterprises creative workflow automation and brand management solutions which leverage innovative GenAi and agentic solutions and deploy novel revenue models.
Software Engineer - Hosted Model Infrastructure Palantir Technologies IncSoftware Engineer - Hosted Model InfrastructurePalo Alto, CA$145,000–$200,000 / yearWe deploy AI models to run in variety of environments: air-gapped government networks, forward-deployed defense environments, edge nodes, and enterprises with strict data sovereignty requirements. Debugging complex issues and performance problems throughout the stack, including open source inference engines, container runtimes, and GPU drivers, in environments you cannot always access directly.
Computational design of biological experiments for model development InceptiveComputational design of biological experiments for model developmentPalo Alto, CaliforniaOur team brings together vast expertise in molecular biology, machine learning, and software engineering, and we are all working towards becoming antedisciplinary, meaning we deepen the knowledge we have in our area of expertise while also expanding our knowledge of completely new fields. We approach our goals with a Beginner's mind, humbly and with fresh eyes, and aim to become the pioneers of a new discipline rooted in biology as much as in deep learning, whose impact will be realized together with out-of-the-box thinkers in business and entrepreneurship, defying established categorizations.
Senior Staff Machine Learning Engineer - Autonomous Driving Foundation Models XPeng Inc.Senior Staff Machine Learning Engineer - Autonomous Driving Foundation ModelsSanta Clara, CA$244,140–$413,160 / yearXPENG is a leading smart technology company at the forefront of innovation, integrating advanced AI and autonomous driving technologies into its vehicles, including electric vehicles (EVs), electric vertical take-off and landing (eVTOL) aircraft, and robotics. Key Responsibilities: Architectural Leadership: Lead the design of end-to-end VLA architectures, bridging multi-modal perception with high-level linguistic reasoning and precise action generation.
Senior AI Infrastructure Engineer - Model Training Kodiak Robotics, IncSenior AI Infrastructure Engineer - Model TrainingMountain View, CA$190,000–$260,000 / yearShould the position require, and Kodiak determines that a candidate's residence, U.S. person status, and/or citizenship status necessitate an export license, bar the candidate from the position, or otherwise fall under national security-related restrictions, Kodiak will consider the candidate for alternative positions unaffected by such restrictions, under terms and conditions set forth at Kodiak's sole discretion, or, as an alternative, opt not to proceed with the candidate's application. Experience building high-performance data pipelines for large-scale training, including streaming dataset formats (WebDataset, MosaicML Streaming/MDS, or similar), sharding, and storage/network-aware loading.
ML Engineer - Healthcare Data Curation & Model Workflows Stanford UniversityML Engineer - Healthcare Data Curation & Model WorkflowsStanford, CA$122,929–$145,389 / yearWORKING CONDITIONS: May be exposed to high voltage electricity, radiation or electromagnetic fields, lasers, noise > 80dB TWA, Allergens/Biohazards/Chemicals /Asbestos, confined spaces, working at heights 10 feet, temperature extremes, heavy metals, unusual work hours or routine overtime and/or inclement weather. Perform supervisory duties, including overseeing the work of technicians and other staff associated with the group/project, supervising the regular installation, maintenance, and operation of complex scientific or engineering projects, and training technicians, operators, and others working in particular scientific or engineering function area.