Lead Member Of Technical Staff, Inference Infrastructure CohereLead Member Of Technical Staff, Inference InfrastructureSan Francisco, CAIn this role, you will provide technical leadership across multiple teams, driving the architecture and strategy for deploying optimized NLP models to production in low latency, high throughput, and high availability environments. You will serve as a key point of contact for customers, leading the design of customized deployments to meet their specific needs, and mentoring engineers to raise the technical bar across the team.
AI Platform Engineer, Infrastructure Brain Co.AI Platform Engineer, InfrastructureSan Francisco, CaliforniaBrain Co. is entering its next phase of production deployments on a national scale with an elite team built from Palantir, Google, Meta, and Nvidia, and a growing footprint across government, insurance, health, and financial services. Underneath it all is Atlas, our proprietary platform that keeps customers in control, secure by design, and never locked into one model.
NewGPU Infrastructure SRE 39 AI IncGPU Infrastructure SRESan Francisco, CAYou will support GPU clusters at 100+ card scale, respond to production incidents, and work closely with training and inference teams on stability, performance, and resource efficiency. Provide infrastructure support for large-model training and online inference workloads, responding quickly to production incidents and operational issues.
Software Engineer - Test Infrastructure Development Katalyst HealthCares & Life Sciences IncSoftware Engineer - Test Infrastructure DevelopmentSanta Clara, CAImplement end-to-end features based on requirements and designs from senior engineers, including backend services, REST APIs, database integrations, and frontend UI. Working knowledge of Git, code review, testing, documentation, CI/CD, Linux deployment, logs, processes, basic networking, and troubleshooting.
Senior Backend Engineer - Infrastructure (ClickPipes) ClickHouseSenior Backend Engineer - Infrastructure (ClickPipes)San Francisco, CaliforniaWorking at a petabyte scale and in high-velocity environments, this role offers an exceptional opportunity to solve challenging technical problems, operate with significant autonomy, and make a measurable impact. We are looking for exceptional backend engineers who can work across a variety of technologies, with deep expertise in Kubernetes, distributed systems, cloud services, and platform reliability.
Member of Technical Staff, Infrastructure SieveMember of Technical Staff, InfrastructureSan Francisco, CaliforniaYou’re likely a good fit if you love optimizing for system uptime, have worked with cloud technologies, optimizing hyper-fast distributed systems at the scale of thousands of GPUs, and building great internal tooling and CI/CD for rapid iteration. As an infrastructure engineer at Sieve , you’ll design and engineer systems that handle the compute, scheduling, and orchestration of complex ML + ETL pipelines that need to run quickly, reliably, and cost-effectively on large sums of video.
Staff Infrastructure Engineer, Cluster Infrastructure Anthropic PBCStaff Infrastructure Engineer, Cluster InfrastructureSan Francisco, CAExperience with cloud networking: VPC design and peering, Shared VPC/Transit Gateway, Cloud Interconnect/Direct Connect, Cloud NAT, cross-cloud private connectivity, BGP and route control, edge load balancing and DDoS mitigation (Cloud Armor / AWS Shield). This research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences.
Software Engineer - Tools & Infrastructure / DevOps Cerebras SystemsSoftware Engineer - Tools & Infrastructure / DevOpsSunnyvale, CAOpenAI recently announced a multi-year partnership with Cerebras, to deploy 750 megawatts of scale, transforming key workloads with ultra high-speed inference. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services.
Senior ML Infrastructure Engineer (Compute) General MotorsSenior ML Infrastructure Engineer (Compute)Mountain View, CA$155,420–$205,900 / yearBenefit options include medical, dental, vision, Health Savings Account, Flexible Spending Accounts, retirement savings plan, sickness and accident benefits, life insurance, paid vacation & holidays, tuition assistance programs, employee assistance program, GM vehicle discounts and more. Collaborate with Simulation engineers, ML engineers and researchers to understand critical workflows, parse them to platform requirements, and deliver incremental value.
Capital Markets - Infrastructure Financing Anthropic PBCCapital Markets - Infrastructure FinancingSan Francisco, CA$325,000–$425,000 / yearIn this role, you will lead financing workstreams end-to-end, engage external financing counterparties, and partner closely with our compute, treasury, legal, finance, and accounting teams to bring transactions from structuring through close. This research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences.
Software Engineer, Research Infrastructure Anthropic PBCSoftware Engineer, Research InfrastructureSan Francisco, CA$405,000–$625,000 / yearThis research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences. You''ll independently scope complex, multi-month projects, drive cross-org alignment through ambiguous problem spaces, and make the architectural decisions that shape how the infrastructure behind our research tooling gets built.
Software Engineer, AI/ML Infrastructure (US-Based) ThumbtackSoftware Engineer, AI/ML Infrastructure (US-Based)San Francisco, California$145,400–$188,100 / yearContribute to the design, development, and deployment of scalable tools and infrastructure to support the efforts of our applied scientists, including traditional ML model training and serving systems, feature and data workflows, CI/CD, orchestration, deployment, and evaluation tooling. For candidates living in San Francisco / Bay Area, San Jose, New York City, or Seattle metros, the expected salary range for the role is currently $145,400 - $188,100.
Software Engineer - Infrastructure Tooling Applied Intuition IncSoftware Engineer - Infrastructure ToolingSunnyvale, CA$149,365–$163,974 / yearApplied Intuition is headquartered in Sunnyvale, California, with offices in Washington, D.C. San Diego; Ft. Walton Beach, Florida; Ann Arbor, Michigan; London; Stuttgart; Munich; Stockholm; Bangalore; Seoul; and Tokyo. Applied Intuition services the automotive, defense, trucking, construction, mining and agriculture industries in three core areas: tools and infrastructure, operating systems, and autonomy.
Founding Product Infrastructure Engineer EmanateFounding Product Infrastructure EngineerSan Francisco, CaliforniaAs a Founding Product Infrastructure Engineer at Emanate, you’ll build the core systems that power our AI revenue engine for companies that are building the backbone of the physical economy.
Principal Software Engineer - Performance Engineering (Cloud Infrastructure) SnowflakePrincipal Software Engineer - Performance Engineering (Cloud Infrastructure)Menlo Park, CaliforniaYou will drive the evaluation of new hardware generations across all major cloud providers, build the benchmarking and modeling infrastructure that turns raw performance data into pricing and rollout decisions, and represent Snowflake in technical discussions with CSPs and silicon partners. We drive Snowflake’s cloud infrastructure adoption and evolution strategy, price/performance decisions, platform footprint expansion amidst capacity constraints, and strategic alignment with cloud service provider (CSP) technical roadmaps.
Sr. Cloud AI Infrastructure Engineer Tencent LTDSr. Cloud AI Infrastructure EngineerPalo Alto, CA$151,300–$283,800 / year1.Architecture Research: Conduct in-depth research into the underlying hardware logic of various AI accelerators; evaluate the power-efficiency ratio and suitability of different heterogeneous architectures in the context of Large Language Model (LLM) inference and training. 2.Operator & Performance Optimization: Design and optimize high-performance operator libraries for large-scale cloud computing environments; resolve long-tail latency issues in hardware scheduling, memory management, and distributed communication.
Staff/Senior Software Engineer, Onboard Infrastructure NuroStaff/Senior Software Engineer, Onboard InfrastructureMountain View, CA$193,930–$352,290 / yearYou have experience in one or more of the following areas: large-scale distributed systems; computer architecture and operating systems; advanced algorithms using C++ and Python; highly-concurrent, multi-processor, and multi-threaded environments; software performance tuning and optimization;profiling and tracing tools and infrastructure (perf, eBPF, Perfetto, pprof, NVIDIA Nsight Systems/Compute); robotics software frameworks; robotics hardware components (including sensors, embedded platforms, etc); and different compute modalities (x86, ARM, GPU, FPGA, etc). The team builds systems and tools for continuous performance analysis, and drives latency reduction and resource efficiency efforts to ensure the autonomy teams can implement an autonomy stack that is efficient and performant for current and future generations of the Nuro Driver.
Software Engineer, ML Infrastructure, Optimization NuroSoftware Engineer, ML Infrastructure, OptimizationMountain View, CA$160,360–$240,540 / yearWith years of real-world deployment experience and a flexible, partner-led business model, Nuro is working toward a future where millions of autonomous vehicles powered by our technology help make everyday life safer, easier, and more connected. You will have an opportunity to work across the full stack of machine learning solutions - from designing robust and scalable model pipelines to building to deploying the optimized models on Nuro's fleet of self-driving robots!
Staff Site Reliability Engineer, Core AI Infrastructure Coinbase Global IncStaff Site Reliability Engineer, Core AI InfrastructureCA$218,025–$256,500 / yearThis team builds custom products through full-stack engineering and scales the infrastructure powering Coinbase''s AI products, with direct exposure to senior leadership in a fast-paced, incubator-style environment. What you''ll do: Own end-to-end delivery of AI products by building production-grade distributed systems, including serving infrastructure, data pipelines, and deployment orchestration across the full stack throughout the SDLC.
Senior Site Reliability Engineer, Core AI Infrastructure Coinbase Global IncSenior Site Reliability Engineer, Core AI InfrastructureCA$186,065–$218,900 / yearThis team builds custom products through full-stack engineering and scales the infrastructure powering Coinbase''s AI products, with direct exposure to senior leadership in a fast-paced, incubator-style environment. What you''ll do: Own end-to-end delivery of AI products by building production-grade distributed systems, including serving infrastructure, data pipelines, and deployment orchestration across the full stack throughout the SDLC.