Senior Machine Learning Engineer - Model Inference AppleSenior Machine Learning Engineer - Model InferenceCupertino, CAYour responsibilities span the full server stack, including onboarding new use cases, optimizing inference across heterogeneous accelerated compute hardware, deploying services on Kubernetes, building and integrating inference engines and control-plane components, and ensuring seamless integration with Maps infrastructure. **Description** As a Software Engineer on the Apple Maps team, you will lead the design and implementation of large-scale, high-performance inference services that support a wide range of models used across Maps, including deep learning and large language models.
Senior Staff Machine Learning Engineer – Autonomous Driving Foundation Models XPeng MotorsSenior Staff Machine Learning Engineer – Autonomous Driving Foundation ModelsSanta Clara, CA$244,140–$413,160 / yearXPENG is a leading smart technology company at the forefront of innovation, integrating advanced AI and autonomous driving technologies into its vehicles, including electric vehicles (EVs), electric vertical take-off and landing (eVTOL) aircraft, and robotics. Key Responsibilities: Architectural Leadership: Lead the design of end-to-end VLA architectures, bridging multi-modal perception with high-level linguistic reasoning and precise action generation.
Staff Machine Learning Engineer - Foundation Model XPeng MotorsStaff Machine Learning Engineer - Foundation ModelSanta Clara, CA$215,280–$364,320 / yearDesign, train, and deploy large deep learning models that can leverage the vast amount of labeled and unlabeled data from a fleet of million vehicles. Within the range, individual pay is determined by work location and additional factors, including job-related skills, experience, and relevant education or training.
Software Engineer - Model Products BaseTenSoftware Engineer - Model ProductsSan Francisco, CARESPONSIBILITIES: Design, build, and operate the Model APIs surface with focus on advanced inference capabilities: structured outputs (JSON mode, grammar-constrained generation), tool/function calling and multi-modal serving. Profile and optimize TensorRT-LLM kernels, analyze CUDA kernel performance, implement custom CUDA operators, tune memory allocation patterns for maximum throughput and optimize communication patterns across multi-GPU setups.
Research Engineer, Production Model Post-Training AnthropicResearch Engineer, Production Model Post-TrainingSan Francisco, CAThis research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences. As a Research Engineer on our Post-Training team, you'll train our base models through the complete post-training stack to deliver the production Claude models that users interact with.
Research Product Manager, Model Behaviors AnthropicResearch Product Manager, Model BehaviorsSan Francisco, CAThis research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences. We offer competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and a lovely office space in which to collaborate with colleagues.
Product Manager, Claude Code Model Performance AnthropicProduct Manager, Claude Code Model PerformanceSan Francisco, CAAs a Product Manager on Claude Code's model performance team, you will drive model launches end-to-end, build evals that measure what matters, and partner directly with researchers and product engineers to translate model improvements into developer-facing outcomes. This research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences.
Threat Intel Manager, Model Exploitation & Fraud AnthropicThreat Intel Manager, Model Exploitation & FraudSan Francisco, CAThe team includes established senior investigators who own our deepest technical casework, tracing distillation networks, reseller and proxy ecosystems, and financially motivated actors across first-party surfaces and third-party platforms; your job is to direct, resource, and amplify that work, not duplicate it. Direct, prioritize, and resource complex investigations into model distillation, unauthorized AI R&D usage, unauthorized access, coordinated account abuse, and fraud/scam networks, partnering with the senior investigators who lead the deepest technical casework and clearing blockers from their path.
ML Engineer - Model Evaluation Expert MercorML Engineer - Model Evaluation ExpertSan Francisco, CaliforniaRemote$60–$90 / hourDesign tasks by transforming real ML research ideas into well-defined, multi-step tasks. Run experiments by implementing changes, executing training experiments, and analyzing results to define correct solutions.
Software Engineer - Model Performance Systems BaseTenSoftware Engineer - Model Performance SystemsSan Francisco, CABy uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. You will not just be building the automated "speedometer and diagnostic" suite for our next-generation AI infrastructure; you will be defining the roadmap, driving key technical decisions, and taking full ownership of the future of this work.
Sr. Manager, Transformation (Pricing & Business Model) AdobeSr. Manager, Transformation (Pricing & Business Model)San Jose, CaliforniaAdobe’s industry-leading offerings including Adobe Acrobat Studio, Adobe Express, Adobe Firefly, Creative Cloud, Adobe Experience Platform, Adobe Experience Manager, and GenStudio enable people and businesses to turn ideas into impact, powered by AI and driven by human ingenuity. Manager, Transformation (Pricing & Business Model) will drive end-to-end execution for prioritized transformation efforts – including translating strategic objectives into specific workplans and analyses, coaching cross-functional teams on transformation design and delivery, and driving problem-solving and alignment.
Robotic AI Engineer/Applied Scientist - Foundation Models Maven RoboticsRobotic AI Engineer/Applied Scientist - Foundation ModelsSan Francisco, CaliforniaMaster Data Efficiency: Develop novel co-training strategies and efficient learning algorithms that leverage diverse data sources—from Internet-scale video to sparse, high-fidelity human interventions. You are not expected to be a master of every domain listed below; however, you must be able to justify world-class excellence in at least one core factor (e.g., model architecture, RL formulations, or high-scale data systems).
Machine Learning Researcher / Engineer (Foundational Models) PathwayMachine Learning Researcher / Engineer (Foundational Models)Palo Alto, CARemotePathway is led by co-founder & CEO Zuzanna Stamirowska, a complexity scientist who created a team consisting of AI pioneers, including CTO Jan Chorowski who was the first person to apply Attention to speech and worked with Nobel laureate Geoff Hinton at Google Brain, as well as CSO Adrian Kosowski, a leading computer scientist and quantum physicist who obtained his PhD at the age of 20. The company is backed by leading investors and advisors, including Lukasz Kaiser, co-author of the Transformer (“the T” in ChatGPT) and a key researcher behind OpenAI’s reasoning models.
Manager, Multi-Modal Language Action Models ZooxManager, Multi-Modal Language Action ModelsFoster City, CAZoox is seeking an experienced Manager of Multi-Modal Language Action Models to lead a team focused on applying cutting-edge Large Multi-Modal Models (MLLMs) to solve concrete, offline autonomy problems. Proven track record of deploying ML/MLLM models to production or internal customers to solve complex, real-world problems, ideally within the autonomy or robotics domain.
Senior Applied Scientist, Efficient LLM Inference & Model Optimization NebiusSenior Applied Scientist, Efficient LLM Inference & Model OptimizationPalo Alto, California$195,200–$262,200 / yearInvent, evaluate, and productionize methods for quantization, QAT, distillation, speculative decoding, KV-cache reuse, KV-cache compression, long-context inference, MoE routing, and model/runtime co-optimization. Build high-quality prototypes in PyTorch, Triton, CUDA-adjacent tooling, or inference-serving frameworks, then work with MLEs and platform engineers to productionize them.
Machine Learning Researcher / Engineer (Foundation Models) PathwayMachine Learning Researcher / Engineer (Foundation Models)Palo Alto, CARemotePathway is led by co-founder & CEO Zuzanna Stamirowska, a complexity scientist who created a team consisting of AI pioneers, including CTO Jan Chorowski who was the first person to apply Attention to speech and worked with Nobel laureate Geoff Hinton at Google Brain, as well as CSO Adrian Kosowski, a leading computer scientist and quantum physicist who obtained his PhD at the age of 20. The company is backed by leading investors and advisors, including Lukasz Kaiser, co-author of the Transformer (“the T” in ChatGPT) and a key researcher behind OpenAI’s reasoning models.
Software Engineer - Voice Model SpaceXAISoftware Engineer - Voice ModelPalo Alto, CA$150,000–$450,000 / yearWork on pre-training and post-training of speech-language models, with targeted enhancements through supervised fine-tuning, reinforcement learning, and other techniques to ensure Grok Voice responses are accurate, factually grounded, natural and idiomatic in spoken style, conversational in tone, and fluent across multiple languages. Build and iterate a comprehensive evaluation framework covering objective metrics (accuracy, quality, latency, expressiveness), human preference studies, content factuality assessments, real-time interaction quality, and experimentation infrastructure to measure and improve performance.
Machine Learning Engineer - Multimodal Modeling Stand InsuranceMachine Learning Engineer - Multimodal ModelingSan Francisco, CaliforniaYou will own modeling work end-to-end, from architecture and training strategy through evaluation and production deployment, and partner closely with the Platform team to ensure the agentic harness and workflows your models plug into deliver strong results in production. Standing up the training-data pipelines, evaluation harnesses, and production monitoring that take these models from prototype to production, collaborating closely with Applied Science infrastructure engineers to build out proper tooling.
Research Program Manager – Adversarial Model Research OpenAIResearch Program Manager – Adversarial Model ResearchSan Francisco, CaliforniaFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. The Human Data team at OpenAI is responsible for identifying and mitigating risks in advanced AI systems by designing evaluations, surfacing vulnerabilities, and collaborating closely with researchers to strengthen model reliability and public trust.
Solutions Architect - AI Model Specialist FriendliAISolutions Architect - AI Model SpecialistSan Francisco, CaliforniaOur infrastructure powers high-throughput, low-latency workloads for global organizations and integrates directly with Hugging Face, providing instant access to over 600,000 open-source models. You will work closely with our customers to integrate FriendliAI’s inference and agent frameworks into real-world products, enabling them to build and scale AI applications effectively.