Staff Software Engineer, ML Infrastructure VoxelStaff Software Engineer, ML InfrastructureSan Francisco, CaliforniaOwn the train-to-deploy handoff: export trained models to optimized inference formats (TensorRT, ONNX), quantify accuracy and latency impact, and partner with Platform on production deployment. Architect and build the training infrastructure that lets the applied ML team run multiple experiments concurrently and iterate quickly on new architectures (PyTorch, AWS).
Staff Software Engineer, Data Infrastructure PeregrineStaff Software Engineer, Data InfrastructureSan Francisco, CA$200,000–$275,000 / yearBacked by leading Silicon Valley investors, Peregrine helps public safety organizations, state and local and governments, federal agencies, and private-sector institutions address society's challenges with unprecedented speed and accuracy. Our AI-enabled platform turns siloed and disconnected data into operational intelligence - instantly surfacing mission-critical information to empower better, faster decisions that improve outcomes at every touchpoint.
Senior Software Engineer, Infrastructure Hayden AI Technologies, Inc.Senior Software Engineer, InfrastructureSan Francisco, CAThe Infrastructure Engineering team is crucial to the overall success of Hayden products: We own all the underlying fabric that connects thousands of devices deployed in the field with large-scale, multi-region cloud services and applications that interact with the data we collect from those devices. Responsibilities: Architect the Service Backbone: Lead the design and evolution of the core services architecture, providing a robust, high-availability backbone utilized by all cloud services and engineering teams.
ML Infrastructure Engineer zaimlerML Infrastructure EngineerSan Mateozaimler was founded by Biswajit Das (ex-VP Engineering, Truera), a Data Infra veteran and former Chief Architect at Visa, and Sofus Macskassy (ex-Director of Engineering, LinkedIn), who built one of the largest knowledge graphs in production in the industry at LinkedIn. zaimler is the context infrastructure for the agentic era: a platform that automatically discovers domain knowledge, maps relationships, and gives AI agents the semantic understanding to operate with precision at scale.
Cloud Infrastructure Engineer zaimlerCloud Infrastructure EngineerSan Mateozaimler was founded by Biswajit Das (ex-VP Engineering, Truera), a Data Infra veteran and former Chief Architect at Visa, and Sofus Macskassy (ex-Director of Engineering, LinkedIn), who built one of the largest knowledge graphs in production in the industry at LinkedIn. zaimler is the context infrastructure for the agentic era: a platform that automatically discovers domain knowledge, maps relationships, and gives AI agents the semantic understanding to operate with precision at scale.
NewSenior Cloud Engineer, Agentic Infrastructure HandshakeSenior Cloud Engineer, Agentic InfrastructureSan Francisco, CaliforniaYou'll work closely with engineering, platform, and security teams to define patterns that scale — and make it genuinely easy for developers to do the right thing. Deep hands-on experience with GCP (strongly preferred), AWS, or Azure — you've run production workloads at scale and know where the bodies are buried.
Member of Technical Staff — Cluster Infrastructure & Supercomputing RadixArkMember of Technical Staff — Cluster Infrastructure & SupercomputingPalo Alto, CaliforniaOur team has optimized kernels serving billions of tokens daily, designed distributed training systems coordinating 10,000+ GPUs, and contributed to infrastructure that powers leading AI companies and research labs. You will design and operate highly reliable, high-performance GPU/TPU clusters, build next-generation scheduling and resource management systems, and push the limits of large-scale distributed infrastructure for AI workloads.
Member of Technical Staff — Reliability-CI Infrastructure RadixArkMember of Technical Staff — Reliability-CI InfrastructurePalo Alto, California$200,000–$400,000 / yearRadixArk is an infrastructure-first company built by engineers who've shipped production AI systems, created SGLang (30K+ GitHub stars, the fastest open LLM serving engine), and developed Miles (our large-scale RL framework). Our team has optimized kernels serving billions of tokens daily, designed distributed training systems coordinating 10,000+ GPUs, and contributed to infrastructure that powers leading AI companies and research labs.
Senior Software Engineer - Backend/Infrastructure SitelineSenior Software Engineer - Backend/InfrastructureSan Francisco, CaliforniaAs a senior member of our small but mighty engineering team, you'll work closely with our cross-functional product team to conceptualize, implement, and ship features from start to finish, as well as own the maintenance and evolution of our backend tech stack and infrastructure. Work tenaciously to build smarter systems that solve real problems for our customers, positively impact their businesses and lives, and make them loyal fans.
Engineering Manager, Platform And AI Infrastructure Anrok, Inc.Engineering Manager, Platform And AI InfrastructureSan Francisco, CAAs we grow, well-designed platform spanning authn and authz, RBAC, FGAC, audit trails and more, along with AI infrastructure is the backbone that lets every other team move fast without compromising the compliance workflows our customers depend on. As the digital economy has grown 6x over the last decade, software businesses have gone from not worrying about sales tax to needing to monitor exposure, calculate rates, and file returns across 50 US jurisdictions and 100+ countries.
Technical Program Manager, Platform Infrastructure (Foundations) AnyscaleTechnical Program Manager, Platform Infrastructure (Foundations)San Francisco, CaliforniaThe Anyscale Technical Program Management (TPM) team is expected to play a critical role executing high impact programs while continuously improving processes to sustainably grow and increase the effectiveness of the Tech organization spanning Design, Engineering and Product teams. Collaborate with the product teams and align all the stakeholders to assemble project teams, assign responsibilities, identify appropriate resources needed, and develop schedules to ensure timely completion of projects by meeting project milestones.
NewIT Infrastructure Engineer SierraIT Infrastructure EngineerSan Francisco, CaliforniaOnboarding & Offboarding Workflows: Develop automated user-lifecycle pipelines (provisioning through deprovisioning) to reduce manual provisioning time. Identity Analytics & Access Governance: Implement an identity-analytics capability to automate orphaned-account discovery, anomalous-login detection, and access-review workflows.
Technical Program Manager, Infrastructure BasetenTechnical Program Manager, InfrastructureSan Francisco, CaliforniaMaintain a clear-eyed view of cross-team dependencies and risks; surface them early, drive mitigation, and keep leadership informed with honest, signal-rich updates. The work is less about owning a single system and more about imposing order on ambiguity: standing up the right structures, driving decisions to closure, and making sure nothing falls through the cracks across dozens of stakeholders.
Staff Engineer, Distributed Storage and HPC & AI Infrastructure Together AIStaff Engineer, Distributed Storage and HPC & AI InfrastructureSan Francisco, CA$250,000–$300,000 / yearYou'll manage and scale high-performance parallel filesystems and object stores, evaluate and integrate cutting-edge technologies such as Vast, Weka, Ceph, and Lustre, and solve the complex engineering challenges of operating at extreme throughput, low-latency data paths, and massive cluster-scale storage operations. Engineer end-to-end data paths to achieve 10+ GB/s per GPU node; architect multi-tier caching for model weights and datasets; tune parallel filesystems using advanced profiling; and scale storage infrastructure across thousands of nodes.
Staff Engineer, Distributed Storage And HPC & AI Infrastructure Together AIStaff Engineer, Distributed Storage And HPC & AI InfrastructureSan Francisco, CA$250,000–$300,000 / yearYou'll manage and scale high-performance parallel filesystems and object stores, evaluate and integrate cutting-edge technologies such as Vast, Weka, Ceph, and Lustre, and solve the complex engineering challenges of operating at extreme throughput, low-latency data paths, and massive cluster-scale storage operations. Engineer end-to-end data paths to achieve 10+ GB/s per GPU node; architect multi-tier caching for model weights and datasets; tune parallel filesystems using advanced profiling; and scale storage infrastructure across thousands of nodes.
Senior Software Engineer, AI Infrastructure - LVM Inference & Evaluation Ambient.aiSenior Software Engineer, AI Infrastructure - LVM Inference & EvaluationRedwood City, CaliforniaPowered by Ambient Pulsar, the first reasoning Vision-Language Model purpose-built for physical security, our platform seamlessly integrates with existing security cameras and physical access control systems to unify monitoring, access control, threat assessment, response, and investigations through an always-on reasoning layer that augments security operators with superhuman capabilities. Hands-on experience running deep learning models in production, ideally including LLMs, LVMs, vision-language models, or multimodal models.
Staff Product Manager, Platform & Infrastructure SkydioStaff Product Manager, Platform & InfrastructureSan Mateo, California$230,000–$280,000 / yearProductize and deliver Skydio On-Prem: Define the roadmap for air-gapped and self-hosted offerings; lead cross-functional delivery (engineering, SRE, security, legal, GTM); own releases, onboarding/runbooks and launch artifacts; run pilots/rollouts to cut time-to-deploy and manual effort while preserving a secure, frictionless experience; and support exponential growth in the national security business. Customer & sales engagement: Run discovery with government and commercial customers: Surface the “why”, build business case for investment, validate proposed solutions in the field, and help grow the business by providing platform capabilities with clearly communicated delivery expectations.
Staff Product Manager, Platform & Infrastructure Skydio, Inc.Staff Product Manager, Platform & InfrastructureSan Mateo, CA$230,000–$280,000 / yearProductize and deliver Skydio On-Prem: Define the roadmap for air-gapped and self-hosted offerings; lead cross-functional delivery (engineering, SRE, security, legal, GTM); own releases, onboarding/runbooks and launch artifacts; run pilots/rollouts to cut time-to-deploy and manual effort while preserving a secure, frictionless experience; and support exponential growth in the national security business. How you'll make an impact: Customer & sales engagement: Run discovery with government and commercial customers: Surface the "why", build business case for investment, validate proposed solutions in the field, and help grow the business by providing platform capabilities with clearly communicated delivery expectations.
Infrastructure Engineer LatchBioInfrastructure EngineerSan Francisco, California$180,000–$250,000 / yearWhile our team is operationally proficient, we do not have the bandwidth to tackle exciting areas of improvement such as speeding up container starts, improving filesystem throughput, scaling up RL operations, and reducing costs. Our charter and continued mission is to accelerate biological science so that humanity can benefit from better bioengineering, whether that be medical technology, farming, or environmental conservation.
Engineering Manager, Machine Learning Infrastructure, Ads RobloxEngineering Manager, Machine Learning Infrastructure, AdsSan Mateo, CA$295,250–$345,040 / yearEvery day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. With Roblox Ads business growing at a rapid rate, we are building large scale ads machine learning infrastructure to deliver effective performance ads to our users, and more business values to our advertisers.