Lead Member Of Technical Staff, Inference Infrastructure CohereLead Member Of Technical Staff, Inference InfrastructureSan Francisco, CAIn this role, you will provide technical leadership across multiple teams, driving the architecture and strategy for deploying optimized NLP models to production in low latency, high throughput, and high availability environments. You will serve as a key point of contact for customers, leading the design of customized deployments to meet their specific needs, and mentoring engineers to raise the technical bar across the team.
Sr Cloud Infrastructure Engineer LendingClub CorpSr Cloud Infrastructure EngineerSan Francisco, CA$175,000–$207,000 / yearHands-on experience using AI tools to accelerate infrastructure work, improve output quality, and reduce operational toil, and you help colleagues apply them responsibly while maintaining appropriate security, validation, and human oversight. 4+ years of experience designing, building, and operating production cloud infrastructure, with deep hands-on expertise in AWS; bachelor's degree or higher, or equivalent combination of education and work experience.
Software Engineer, Cloud Infrastructure DatologyAISoftware Engineer, Cloud InfrastructureRedwood City, CaliforniaTraining on curated data can dramatically reduce training time and cost ( 7-40x faster training depending on the use case), dramatically increase model performance as if you had trained on >10x more raw data without increasing the cost of training, and allow smaller models with fewer than half the parameters to outperform larger models despite using far less compute at inference time, substantially reducing the cost of deployment. We raised a total of $57.5M in two rounds, a Seed and Series A. Our investors include Felicis Ventures, Radical Ventures, Amplify Partners, Microsoft, Amazon, and AI visionaries like Geoff Hinton, Yann LeCun, Jeff Dean, and many others who deeply understand the importance and difficulty of identifying and optimizing the best possible training data for models.
Infrastructure Engineer RelaceInfrastructure EngineerSan Francisco, CaliforniaWe power the fastest model on OpenRouter (10,000 tok/s) and deliver optimized small language models designed for retrieval, application, and core code generation functions. Our technology supports some of the world’s fastest-moving companies — including Lovable, Figma, and Vercel — as they deploy and scale code generation to hundreds of millions of users.
ML Infrastructure Engineer Echo NeurotechnologiesML Infrastructure EngineerSan Francisco, California$180,000–$230,000 / yearThe person who fills this role will design, build, and scale infrastructure to power massive-scale data, modeling, and analysis platforms, playing a critical role in shaping a high-performance, production-grade ML ecosystem to support rapid experimentation with diverse datasets spanning neural signals, behavior, and more. The work will ultimately enable the development of cutting-edge models for neuroscientific discovery and neural decoding, empowering brain-computer interface technology to improve the lives of patients living with severe neurological conditions.
Forward Deployed Infrastructure Engineer SiriusForward Deployed Infrastructure EngineerSan Francisco, CaliforniaWork directly with customers' platform engineering, security, and DevOps teams to navigate their infrastructure constraints — particularly in regulated industries where subscription data is sensitive. Own deployment architecture conversations and execution — designing and implementing deployment strategies across customer environments, including VPC configuration, permissioning, networking, and infrastructure provisioning.
Senior AI Platform Engineer, Infrastructure Services SentinelOneSenior AI Platform Engineer, Infrastructure ServicesSan Francisco, CA$132,000–$182,000 / yearDesign across the platform, not just the gateway: work fluently with our CI/CD (Jenkins, JPAAS), GitOps and Kubernetes deployment tooling (ArgoCD across dev/gov/prod), artifact management (Artifactory/Xray), GitHub Enterprise administration, and GitHub Actions runner fleet, so that AI infrastructure decisions account for how the rest of the platform actually works. As a Senior AI Platform Engineer, Infrastructure Services, you will be tasked with taking ownership of our AI Gateway infrastructure (built on Kong AI Gateway), the system that authenticates, routes, rate-limits, and monitors AI coding assistant traffic org-wide, while also being fluent enough across our broader platform stack to design solutions that span the two.
Software Engineer - Infrastructure Emergent LabsSoftware Engineer - InfrastructureSan Francisco, CaliforniaTech Stack: We don't require previous experience with our entire stack, but enthusiasm for learning is key: Go, Python, Kubernetes, ArgoCD, Helm, GCP, AWS, Cloudflare, Grafana, Prometheus, Loki, New Relic, PagerDuty, PostgreSQL, MongoDB, Redis, Kafka, and GitHub. Emergent builds autonomous coding agents that replace traditional software development by generating, testing, and deploying production applications directly from plain-language intent.
Director Of Corporate Infrastructure SofiDirector Of Corporate InfrastructureSan Francisco, CA$192,000–$330,000 / yearThis includes managing all aspects of our network - engineering and operations to improve user experience and performance, and also supporting the multi-terabit backbone network that interconnects edge PoPs, corporate offices, data centers, and cloud gateways. Develop the multi-year networking architecture, strategy, and roadmap aimed at implementing an advanced network architecture to support SoFi's evolving business needs, inclusive of SoFi and its subsidiaries.
Senior Software Engineer - Data Infrastructure, Safety RobloxSenior Software Engineer - Data Infrastructure, SafetySan Mateo, CA$196,750–$243,290 / yearAligned and partnering with product teams, we use this tool-belt to discover new opportunities, influence and shape the product roadmap and prioritization, build safety products, and measure the impact on our community of users and developers. To pro-actively find bad actors and protect good users, Roblox needs to ingest enormous amounts of data and make it usable both to our automated detection systems and for human moderators and customer support agents doing investigations.
Senior Software Engineer, Test Infrastructure ZooxSenior Software Engineer, Test InfrastructureFoster City, CA$215,000–$237,000 / yearDemonstrated ability to design and maintain test frameworks for large codebases (deep pytest experience: fixtures, parametrization, scalable suites), and to design REST and library/framework APIs that other engineers build on. Qualifications: Strong production Python, including hands-on experience building distributed task/queue systems (e.g., Celery or equivalent) with a message broker and result backend (Redis / RabbitMQ) - task lifecycles, retries, cancellation, and worker reliability.
Software Engineer, Data Infrastructure DatologyAISoftware Engineer, Data InfrastructureRedwood City, California$180,000–$300,000 / yearTraining on curated data can dramatically reduce training time and cost ( 7-40x faster training depending on the use case), dramatically increase model performance as if you had trained on >10x more raw data without increasing the cost of training, and allow smaller models with fewer than half the parameters to outperform larger models despite using far less compute at inference time, substantially reducing the cost of deployment. We raised a total of $57.5M in two rounds, a Seed and Series A. Our investors include Felicis Ventures, Radical Ventures, Amplify Partners, Microsoft, Amazon, and AI visionaries like Geoff Hinton, Yann LeCun, Jeff Dean, and many others who deeply understand the importance and difficulty of identifying and optimizing the best possible training data for models.
Infrastructure Consultant 3 Infosys LTDInfrastructure Consultant 3Sunnyvale, CAIn the assigned Job Role of Infrastructure Consultant 3, your Area Of Responsibility will be as below: Develop diagnostic tools, establish monitoring protocols, and resolve complex incidents through root cause analysis while ensuring SLA adherence and creating action plans to prevent recurrence. Conduct detailed IT infrastructure evaluations and due diligence across multiple tracks, identify gaps and risks, and collaborate with stakeholders to recommend tailored enhancements and validate infrastructure details.
Senior Software Engineer, Build Infrastructure ZooxSenior Software Engineer, Build InfrastructureFoster City, CA$208,000–$271,000 / yearDirectly accelerate engineering velocity by minimizing Bazel overhead, streamlining incremental builds to reduce iteration times, and working with our vendor Engflow to optimize our RBE platform. You'll join the team that owns the build infrastructure making that possible: maintaining and evolving our Bazel-based build system in a large, multi-language monorepo used by hundreds of engineers.
Senior GPU Infrastructure Engineer Hyperbolic LabsSenior GPU Infrastructure EngineerSan Francisco, CaliforniaYou'll work at the cutting edge of cloud infrastructure, building the core orchestration layer that enables our platform to deliver up to 75% cost savings compared to traditional cloud providers. Experience with storage and data infrastructure for AI/ML workloads, including object storage, high-IOPS block storage, and distributed file systems for training data and checkpoints.
Machine Learning Infrastructure Engineer, Safeguards Research AnthropicMachine Learning Infrastructure Engineer, Safeguards ResearchSan Francisco, CAThis research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences. Research shows that people who identify as being from underrepresented groups are more prone to experiencing imposter syndrome and doubting the strength of their candidacy, so we urge you not to exclude yourself prematurely and to submit an application if you're interested in this work.
Infrastructure Engineer MercorInfrastructure EngineerSan Francisco, CaliforniaMercor Enterprise brings this same infrastructure to Fortune 500 companies: helping companies capture how their best people actually work, translating that expertise directly back into agents. You’ll work closely with engineers across product, research, and operations to design scalable architectures, streamline deployments, and improve observability.
Corporate Development Associate - Digital Infrastructure M&A LambdaCorporate Development Associate - Digital Infrastructure M&ASan Francisco, CaliforniaOur investors notably include TWG Global, US Innovative Technology Fund (USIT), Andra Capital, SGW, Andrej Karpathy, ARK Invest, Fincadia Advisors, G Squared, In-Q-Tel (IQT), KHK & Partners, NVIDIA, Pegatron, Supermicro, Wistron, Wiwynn, Gradient Ventures, Mercato Partners, SVB, 1517, and Crescent Cove. We are seeking a Corporate Development Associate to play a key role in managing and executing the company’s end-to-end acquisition and strategic investment activities on the real asset / digital infrastructure side of our business.
Member of Technical Staff - Infrastructure PlixMember of Technical Staff - InfrastructureSan MateoExperience working with cloud computing offerings like AWS and GCP, specifically with services like GCS/S3, GCE/EC2, GKE, Container Registries, etc. Experience with SQL and no-SQL databases (e.g., MySQL, PostgreSQL, MongoDB, Cassandra) and web servers (e.g., Apache, Nginx).
AI Infrastructure Operations, Demand Planning AnthropicAI Infrastructure Operations, Demand PlanningSan Francisco, CARun a portfolio of bring-ups in parallel - new cloud regions, on-prem sites, neocloud blocks - with one integrated schedule spanning provider milestones, cluster creation, network turn-up, storage readiness, health burn-in, and first-workload landing. Capacity Engineering owns the data, tooling, and systems that let Anthropic plan, measure, and maximize utilization of that fleet: we partner on supply deals, wire telemetry from day zero, own the canonical capacity data layer, and build the planning and enforcement tools every research and product team relies on.