Director of Engineering for Open Data Analytics Engines, Data Processing & Experience Amazon.com IncDirector of Engineering for Open Data Analytics Engines, Data Processing & ExperienceRedmond, WAAWS Open Data Analytics Engines is a suite of fully managed, high-performance analytics services built on popular open-source frameworks, enabling customers to process, analyze, and derive insights from massive datasets with speed, flexibility, and cost efficiency. Draw from your deep and broad technical and management expertise to mentor senior engineers and managers, complete hands-on technical work, and provide leadership on complex technical issues, design tradeoffs, and feature and schedule prioritization.
Staff Software Engineer - Search / AI CVS Health CorpStaff Software Engineer - Search / AIWA$106,605–$284,280 / yearCurrently, we are seeking a Staff Software Engineer - Search / AI who as a Senior technical leader, be responsible for driving architecture, design, and delivery of scalable, cloud-native platforms built on microservices architecture and AI capabilities. CVS Health is looking for hands-on, passionate people who want to join a high energy and growing team to make a difference in customers' lives and who want to be on the forefront of digital innovation that aims to reinvent what a pharmacy and a health care company can be in the digital world.
Senior AI Search Product Manager Zoom Communications IncSenior AI Search Product ManagerSeattle, WAWe set out to build the best collaboration platform for the enterprise, and today help people communicate better with products like Zoom Contact Center, Zoom Phone, Zoom Events, Zoom Apps, Zoom Rooms, and Zoom Webinars. As part of our award-winning workplace culture and commitment to delivering happiness, our benefits program offers a variety of perks, benefits, and options to help employees maintain their physical, mental, emotional, and financial health; support work-life balance; and contribute to their community in meaningful ways.
Hardware Design Engineer, AI Inference Engine ElastixAIHardware Design Engineer, AI Inference EngineSeattle, WashingtonWe are developing a cutting-edge AI inference solution that dramatically improves efficiency through a holistic co-design approach, spanning from machine learning optimizations and a highly specialized software stack to the inference engine and underlying cloud hardware. Work hand-in-hand with software engineers to define a seamless hardware-software interface, ensuring the inference engine is highly programmable, efficient, and easy to integrate into our broader software stack and compiler.
NewSoftware Engineer Project Intern (Global E-Commerce Search Infrastructure) - 2026 Start TikTok IncSoftware Engineer Project Intern (Global E-Commerce Search Infrastructure) - 2026 StartSeattle, WASearch will propel TikTok Shop from tens of billions to hundreds of billions in annual GMV by activating real shopping intent, unlocking zero-sales products, and ensuring every item becomes instantly discoverable the moment a user types or taps. Data Pipelines: Build highly scalable and fault-tolerant data pipelines using Flink, Kafka, and Spark to ensure product changes (price, stock, and new listings) are reflected in search results in near real-time.
Senior/Staff Software Engineer - Machine Learning & System Optimization ZooxSenior/Staff Software Engineer - Machine Learning & System OptimizationSeattle, WA$226,000–$307,000 / yearOptimize large-scale models (Multi-Modal Sensor Fusion models, LLMs, VLMs) using advanced quantization (PTQ, QAT), pruning, mixed-precision inference frameworks, and parameter-efficient fine-tuning (LoRA, QLoRA). As a Machine Learning and System Optimization Engineer, you will orchestrate and allocate overall system capacity to various core perception models running on-bot, as well as drive large initiatives that allow for more efficient inference by sharing various parts of the perception stack with one another.
Principal Product Manager, Inference Engine DigitalOceanPrincipal Product Manager, Inference EngineSeattle, WA$218,400–$273,000 / yearMaximize GPU utilization and margin: Create a clear product and business framework for improving token revenue per GPU hour, reducing idle capacity, reclaiming underutilized infrastructure, and driving better gross margin as the business scales. Fluency in modern AI workloads: Familiarity with LLM inference, open-source models, model serving, prompt caching, batching, model routing, media models, latency tradeoffs, and production AI application patterns.
Tech Lead, Google Kubernetes Engine AI Platform Google LLCTech Lead, Google Kubernetes Engine AI PlatformSeattle, WA$207,000–$300,000 / yearCareers Skip navigation links home home Home work_outlinework_outline Jobs noogler_hat noogler_hat Students google google How we work handyman handyman How we hire person_outline person_outline Your career help_outline Help link feedback Send feedback more_vert Help Send Feedback Sign in. Were looking for engineers who bring fresh ideas from all areas including information retrieval, distributed computing, large-scale system design, networking and data storage, security, artificial intelligence, natural language processing, UI design and mobile-the list goes on and is growing every day.
Software Development Engineer, Aurora Open Source engines Amazon.com IncSoftware Development Engineer, Aurora Open Source enginesRedmond, WAWe pioneered cloud computing and never stopped innovating - that's why customers from the most successful startups to Global 500 companies trust our robust suite of products and services to power their businesses. We take databases to their limits - our customers rely on Aurora MySQL and Aurora PostgreSQL databases for their business and due to our scale, we solve challenges no other database environment sees.
Research Engineer - LLM/VLM Inference Optimization (Seed Infra) Beijing ByteDance Technology Co LtdResearch Engineer - LLM/VLM Inference Optimization (Seed Infra)Seattle, WABuild state-of-the-art model inference engines through advanced performance optimization techniques such as compiler-level optimizations, parallel computing, graph fusion, efficient CUDA kernel development, low-precision computation, streaming inference, speculative decoding, and high-concurrency request optimization. Design, develop, and optimize high-performance inference systems for large-scale LLMs and VLMs, covering inference engines, serving frameworks, and end-to-end deployment pipelines.
NewSoftware Engineer Intern (AML-Engine-Orchestration) - 2027 Start Beijing ByteDance Technology Co LtdSoftware Engineer Intern (AML-Engine-Orchestration) - 2027 StartSeattle, WABuild serving orchestration and traffic management capabilities for disaggregated serving clusters, including topology-aware scheduling, KV Cache affinity, intelligent request routing, and QoS/SLA management. The Data-AML-Engine Orchestration team builds large-scale machine learning infrastructure that powers online model serving across ByteDance products, including TikTok.
NewMachine Learning Engineer Intern (AML-Engine-Orchestration) - 2027 Start Beijing ByteDance Technology Co LtdMachine Learning Engineer Intern (AML-Engine-Orchestration) - 2027 StartSeattle, WABuild serving orchestration and traffic management capabilities for disaggregated serving clusters, including topology-aware scheduling, KV Cache affinity, intelligent request routing, and QoS/SLA management. The Data-AML-Engine Orchestration team builds large-scale machine learning infrastructure that powers online model serving across ByteDance products, including TikTok.
NewMachine Learning Engineer Graduate (AML-Engine-Orchestration) - 2027 Start Beijing ByteDance Technology Co LtdMachine Learning Engineer Graduate (AML-Engine-Orchestration) - 2027 StartSeattle, WABuild serving orchestration and traffic management capabilities for disaggregated serving clusters, including topology-aware scheduling, KV Cache affinity, intelligent request routing, and QoS/SLA management. The Data-AML-Engine Orchestration team builds large-scale machine learning infrastructure that powers online model serving across ByteDance products, including TikTok.
NewSoftware Engineer Graduate (AML-Engine-Orchestration) - 2027 Start Beijing ByteDance Technology Co LtdSoftware Engineer Graduate (AML-Engine-Orchestration) - 2027 StartSeattle, WABuild serving orchestration and traffic management capabilities for disaggregated serving clusters, including topology-aware scheduling, KV Cache affinity, intelligent request routing, and QoS/SLA management. The Data-AML-Engine Orchestration team builds large-scale machine learning infrastructure that powers online model serving across ByteDance products, including TikTok.
NewSoftware Engineer Graduate (AML-Engine-Orchestration) - 2027 Start (PhD) Beijing ByteDance Technology Co LtdSoftware Engineer Graduate (AML-Engine-Orchestration) - 2027 Start (PhD)Seattle, WABuild serving orchestration and traffic management capabilities for disaggregated serving clusters, including topology-aware scheduling, KV Cache affinity, intelligent request routing, and QoS/SLA management. The Data-AML-Engine Orchestration team builds large-scale machine learning infrastructure that powers online model serving across ByteDance products, including TikTok.
Staff Engineer, Inference Optimizations DigitalOceanStaff Engineer, Inference OptimizationsSeattle, WA$191,200–$239,000 / yearHardware & Ecosystem Mastery: Act as the subject matter expert on modern GPU families (NVIDIA/AMD) and their software stacks (CUDA, ROCm, TensorRT, OpenAI Triton), advising on hardware procurement and software integration. Deep-Dive Optimization: Engineer solutions for complex performance issues, including attention layer optimizations, memory and precision management, and advanced parallelization across multi-node GPU clusters.
Senior Engineer 2: Inference Optimizations DigitalOcean LLCSenior Engineer 2: Inference OptimizationsSeattle, WARemote$167,200–$209,000 / yearHardware & Ecosystem Mastery: Act as the subject matter expert on modern GPU families (NVIDIA/AMD) and their software stacks (CUDA, ROCm, TensorRT, OpenAI Triton), advising on hardware procurement and software integration. Deep-Dive Optimization: Engineer solutions for complex performance issues, including attention layer optimizations, memory and precision management, and advanced parallelization across multi-node GPU clusters.
Software Development Engineer, Amazon Search Autocomplete and Navigation AI Amazon.com IncSoftware Development Engineer, Amazon Search Autocomplete and Navigation AISeattle, WAThe team generates highly relevant, context-aware, and personalized search refinements, optimizing their presentation order in a variety of Search Navigation UIs and guiding users to the most relevant search results. Additionally, we are seeking candidates with determination and rigor in engineering, deep curiosity and interest for applied sciences, creativity, and sound logic and reasoning.
Senior AI Search Product Manager ZoomSenior AI Search Product ManagerSeattle, WAWe set out to build the best collaboration platform for the enterprise, and today help people communicate better with products like Zoom Contact Center, Zoom Phone, Zoom Events, Zoom Apps, Zoom Rooms, and Zoom Webinars. As part of our award-winning workplace culture and commitment to delivering happiness, our benefits program offers a variety of perks, benefits, and options to help employees maintain their physical, mental, emotional, and financial health; support work-life balance; and contribute to their community in meaningful ways.
Game Engine Engineer, Platform Worldscape TechnologyGame Engine Engineer, PlatformRedmond, Washington$100,000–$130,000 / yearWith a zero-trust security framework, seamless integration with legacy systems, and robust developer tools, Worldscape empowers rapid deployment of agentic applications that drive resilience and efficiency in sectors like logistics, telecom, defense, and infrastructure. Focus on core physics features: Collision detection (broad‑phase/narrow‑phase), entity controllers, vehicle dynamics with an emphasis on orbital mechanics, particles/fluid mechanics (with an emphasis on atmospheric dynamics).