Senior DevOps Engineer

Avantos AI
      30 days ago

      Job Description

      Senior DevOps Engineer

      NYC (Hybrid) or US Remote | Competitive + Equity


      About Avantos

      Avantos is building the AI-native operating system for financial services — transforming fragmented data into a single, intelligent system that powers workflows, automation, and decision-making.

      We’re a product-led, fast-moving team at the intersection of AI, fintech, and modern infrastructure.

       

      The Role

      We're seeking a Senior DevOps Engineer / Site Reliability Engineer to own and evolve our infrastructure, reliability, and deployment practices. You'll be responsible for building the foundational platform that enables our engineering teams to ship quickly and reliably while maintaining the security and compliance standards required in financial services.

       

      What You’ll Do

      • Design, implement, and maintain our AWS cloud infrastructure using infrastructure-as-code principles with Terraform

      • Build and optimize CI/CD pipelines to enable rapid, safe deployments across multiple environments

      • Own observability strategy—implement comprehensive monitoring, logging, and alerting systems using Datadog and other tooling

      • Architect and manage containerized workloads on ECS Fargate and evaluate migration paths to Kubernetes

      • Establish and enforce security best practices, working closely with compliance teams on financial services requirements

      • Design and implement disaster recovery, backup, and business continuity strategies

      • Optimize system performance, cost efficiency, and resource utilization across AWS services

      • Collaborate with engineering teams to improve service reliability, reduce toil, and establish SLOs/SLIs

      • Participate in incident response and conduct thorough post-mortems to drive continuous improvement

      • Mentor engineers on DevOps practices, cloud architecture patterns, and operational excellence

      What We’re Looking For

      • 8+ years of experience in DevOps, SRE, or infrastructure engineering roles

      • Expert-level proficiency with AWS services including ECS Fargate, ALB, Cognito, S3, SQS, and related services

      • Deep hands-on experience with Terraform for managing complex, multi-account AWS environments

      • Strong scripting and automation skills in Python and/or Bash

      • Proven experience designing and implementing CI/CD pipelines (GitHub Actions, ArgoCD, or similar)

      • Solid understanding of containerization technologies (Docker) and orchestration platforms (Kubernetes/ECS)

      • Experience with observability and monitoring tools (Datadog, CloudWatch, or equivalent)

      • Deep knowledge of networking, security, and AWS best practices

      • Strong problem-solving abilities and experience troubleshooting complex distributed systems

      • Excellent communication skills and ability to work cross-functionally with engineering teams

      Bonus

      • Have startup or early-stage experience

      • Experience with PostgreSQL performance tuning and RDS management

      Why Join

      • Build the foundation layer of an AI-native product

      • High ownership + direct impact

      The Fit

      You’re a systems thinker with a passion for reliability and automation. You love tackling complex infrastructure challenges, thrive in fast-paced environments, and are motivated by the responsibility of building secure, scalable platforms from the ground up. You collaborate deeply, care about operational excellence, and are energized by turning ambiguity into robust solutions that help teams move faster and safer.

      Numbers & Facts

      Location
      Websitehttps://avantos.ai/

      Skills

      • Amazon Web Services (AWS)unmatched
      • Artificial Intelligence (AI)unmatched
      • Automationunmatched
      • Best Practicesunmatched
      • Business Strategyunmatched
      • Cloud Architectureunmatched
      • Cloud Computingunmatched
      • Communication Skillsunmatched
      • Continuous Deployment/Deliveryunmatched
      • Continuous Improvementunmatched
      • Continuous Integrationunmatched
      • Cross-Functionalunmatched
      • Data Recoveryunmatched
      • DevOpsunmatched
      • Disaster Recoveryunmatched
      • Distributed Computingunmatched
      • Dockerunmatched
      • Energy Efficiencyunmatched
      • Financial Complianceunmatched
      • Financial Servicesunmatched
      • Financial Systemsunmatched
      • GitHubunmatched
      • Identify Issuesunmatched
      • Incident Responseunmatched
      • Machine Toolunmatched
      • Mentoringunmatched
      • Network Securityunmatched
      • Operating Systemsunmatched
      • Performance Tuning/Optimizationunmatched
      • PostgreSQLunmatched
      • Problem Solving Skillsunmatched
      • Process Improvementunmatched
      • Regulatory Complianceunmatched
      • Reliability Engineeringunmatched
      • Resource Utilizationunmatched
      • Scalable System Developmentunmatched
      • Software Engineeringunmatched
      • Team Playerunmatched

      Be found by employers

      5,500+ employers search our resume database daily. Add yours to get found by recruiters looking for candidates like you.

      Level up your application

      Professional resume templates

      Browse dozens of recruiter approved resume templates, layouts and formats. Choose your favorite and make it your own in minutes.

      Free resume templates

      Free resume builder

      Improve your existing resume or start from scratch and create a standout, ATS-friendly resume. Add job-specific content, download and apply.

      Free resume builder