Principal Site Reliability Engineer
An established fintech institution is seeking a Principal Site Reliability Engineer (SRE) to apply software engineering and systems engineering best practices to ensure the stability, scalability, and overall health of critical production services. This full-time position is hybrid, with 3 days on-site in Downtown Chicago and 2 days remote.
This is a highly influential Principal-level role with significant impact on engineering and reliability strategy. This role collaborates with engineering and technology teams to build reliable, scalable, and observable systems throughout their lifecycle. The Principal SRE leads technical initiatives across the enterprise by leveraging data-driven engineering, automation, and strong architectural expertise to solve complex challenges and align teams around scalable solutions. Required Skills & Experience
15+ years of professional experience in Software Engineering, Site Reliability Engineering, Systems Engineering, Cloud/Platform Engineering, DevOps, or Infrastructure Engineering.
Strong leadership capabilities including the ability to mentor junior engineers and the capacity to drive architecture decisions and influence technical direction.
The ability to operates with significant autonomy across the organization's most consequential reliability challenges.
Proven track record of using objective analysis and technical expertise to validate assumptions, assess risks, and guide engineering decisions
Desired Skills & Experience
AWS/Cloud experience
C++ experience
Bachelor's degree in Computer Science, Software Engineering, Computer Engineering, or Information Systems
Experience in the Security, Payments, or Financial Services industries
What You Will Be Doing Tech Breakdown
Languages/scripting: Java, Python
Containers/clusters: Kubernetes
IaC: Terraform
Observability/monitoring: Datadog, Grafana
OS: Linux
Daily Responsibilities
Apply automation, DevOps, and software engineering best practices to improve how services are built, deployed, monitored, and maintained.
Use data and analysis to identify reliability risks, validate solutions, and guide technical decisions.
Define and enhance service reliability metrics, objectives, and performance standards.
Improve system observability through monitoring, logging, tracing, alerting, and dashboards.
Drive continuous improvements in CI/CD, Infrastructure as Code, automation, testing, incident response, and system reliability.
Identify recurring production issues and implement long-term improvements in code, architecture, tooling, and operational processes.
The Offer
This role is bonus eligible.
You will receive the following benefits:
Competitive medical (PPO and HDHP), dental, and vision insurance, plus employer contributions to Health Savings Accounts (HSA) and Flexible Spending Accounts (FSA) for healthcare, commuting, and dependent care expenses.
Dollar-for-dollar 401(k) match on the first 6% of employee contributions, available upon eligibility.
Flexible Time Off (FTO) for salaried employees, generous PTO for hourly employees, 11 paid company holidays, and a paid volunteer day.
Up to 12 weeks of paid parental leave to support growing families.
Access to Maven's comprehensive family planning benefits, including fertility treatments, egg freezing, adoption, surrogacy, pregnancy and postpartum care, pediatric support, and return-to-work resources.
Applicants must be currently authorized to work in the US on a full-time basis now and in the future.
Numbers & Facts
Location
Chicago, IL
Skills
Amazon Web Services (AWS)unmatched
Analysis Skillsunmatched
Architectural Servicesunmatched
Automationunmatched
Best Practicesunmatched
C++ Programming Languageunmatched
Cloud Computingunmatched
Computer Engineeringunmatched
Computer Scienceunmatched
Continuous Deployment/Deliveryunmatched
Continuous Improvementunmatched
Continuous Integrationunmatched
Data Analysisunmatched
DevOpsunmatched
Financial Servicesunmatched
Flexible Spending Accountsunmatched
Healthcareunmatched
Identify Issuesunmatched
Incident Responseunmatched
Information Technology & Information Systemsunmatched
Insuranceunmatched
Javaunmatched
Leadershipunmatched
Linux Operating Systemunmatched
Machine Toolunmatched
Mentoringunmatched
Metricsunmatched
Operating Systemsunmatched
Operations Processesunmatched
Preferred Provider Organization (PPO)unmatched
Process Improvementunmatched
Python Programming/Scripting Languageunmatched
Reliability Engineeringunmatched
Reporting Dashboardsunmatched
Risk Analysisunmatched
Scalable System Developmentunmatched
Scripting (Scripting Languages)unmatched
Software Engineeringunmatched
System Lifecycleunmatched
Systems Engineeringunmatched
Systems Reliabilityunmatched
Systems Scalabilityunmatched
Technical Leadershipunmatched
Test Automationunmatched
🎯
Be found by employers
5,500+ employers search our resume database daily. Add yours to get found by recruiters looking for candidates like you.
Level up your application
Professional resume templates
Browse dozens of recruiter approved resume templates, layouts and formats. Choose your favorite and make it your own in minutes.