Genesis10 is currently seeking a Site Reliability / Platform Engineer - Remote position with a Leading Asset Management Firm. This role is open to Remote US based resources, however candidates that are able to work hybrid in either New York City or Austin, TX are preferred. This is a 12+ month contract opportunity.
Compensation Ranges:
W2: $90–$105/hr
C2C: $110–$125/hr
We are seeking a highly hands-on SRE / Platform Engineer for a hybrid software development, Site Reliability Engineering, and systems engineering role. This position is ideal for a strong developer who has expanded into infrastructure, cloud, automation, and production engineering and wants to take a more holistic view of how applications and systems operate together.
The team follows an engineering-focused SRE model centered on using software and automation to solve infrastructure and operational problems. Engineers on the team write code every day and work across application and infrastructure layers to improve reliability, performance, scalability, observability, and system integration.
A major initiative for the team is establishing a centralized observability capability across an environment where monitoring and operational data have historically been siloed. The organization is bringing telemetry together using Datadog and enterprise data lake capabilities, creating a common observability foundation that can ultimately support AIOps, agentic AI, automated remediation, and self-healing systems.
This is not a traditional operations or Solutions Architecture position. The successful candidate will be expected to build, deploy, automate, troubleshoot, and improve the environment using technologies such as Kubernetes, Terraform, Ansible, and Datadog.
Responsibilities:
- Develop software and automation to improve the reliability, scalability, performance, and operational efficiency of production systems.
- Write and maintain infrastructure and configuration code using Terraform and Ansible as part of day-to-day engineering activities.
- Deploy, manage, automate, and troubleshoot applications and services running on Kubernetes.
- Automate the deployment and configuration of Datadog and other observability capabilities, including Ansible-based deployments.
- Help establish centralized observability across previously siloed applications, infrastructure, cloud, and data environments.
- Bring together metrics, logs, traces, events, and other telemetry to provide a comprehensive view of system health, dependencies, and potential business impact.
- Build monitoring, dashboards, alerting, and observability capabilities that enable teams to identify and resolve production issues more effectively.
- Help create the observability and operational-data foundation required to support future AIOps and agentic AI capabilities.
- Support longer-term initiatives around intelligent incident detection, root-cause analysis, impact analysis, automated remediation, and self-healing.
- Partner with application developers to improve application performance, resiliency, deployment patterns, and integration with other systems.
- Evaluate how applications interact with infrastructure, networks, APIs, databases, and data platforms and identify opportunities for improvement.
- Troubleshoot complex production issues spanning applications, Kubernetes, cloud infrastructure, networking, and data environments.
- Provide advanced production engineering support across traditional L2/L3 boundaries, focusing on root cause and permanent solutions rather than temporary fixes.
- Support large-scale data environments that may include Databricks, Snowflake, and enterprise data lake technologies.
- Identify repetitive or manual operational processes and replace them with scalable, code-based solutions.
Required Skills:
- Strong hands-on experience in software development, with some additional experience in SRE, DevOps, Platform Engineering, or Infrastructure Engineering.
- Strong Terraform experience, including the ability to write, deploy, maintain, and troubleshoot Infrastructure-as-Code.
- Strong Ansible experience, including developing playbooks/roles and automating software and infrastructure deployments.
- Hands-on Kubernetes experience, including deployment, configuration, troubleshooting, and production support.
- Strong understanding of observability, including metrics, logs, traces, monitoring, alerting, dashboards, and production troubleshooting.
- Experience with Datadog or a comparable enterprise observability platform; direct Datadog experience is strongly preferred.
- Experience automating the deployment and configuration of infrastructure and/or observability tooling.
- Strong coding or scripting capabilities; Python and/or Java experience is highly desirable.
- Experience working with cloud infrastructure and an understanding of how applications, networks, infrastructure, and data platforms interact.
- Strong troubleshooting and root-cause analysis skills across both application and infrastructure layers.
- Experience supporting highly available, production-scale systems.
- Demonstrated ability to solve infrastructure and operational problems through code and automation rather than manual processes.
- Ability to work independently, learn unfamiliar technologies quickly, and take ownership of problems from investigation through implementation.
- Only candidates available and ready to work directly as Genesis10 employees will be considered for this position.
Preferred Qualifications:
- Experience building, consolidating, or modernizing enterprise observability platforms.
- Hands-on experience deploying or managing Datadog through Ansible, Terraform, or other automation.
- Experience with data lakes, Databricks, Snowflake, or other large-scale data platforms.
- Exposure to AIOps, AI/ML, agentic AI, automated remediation, or self-healing systems.
- Experience using Python or Java for infrastructure, platform, or operational automation.
- Experience working in a startup or similarly broad engineering environment where engineers have responsibility across multiple layers of the technology stack.
- Financial services experience is a plus but not required.
Pay rate range: $ - $ hourly.
If you have the described qualifications and are interested in this exciting opportunity, please apply!
Ranked a Top Staffing Firm in the U.S. by Staffing Industry Analysts for six consecutive years, Genesis10 puts thousands of consultants and employees to work across the United States every year in contract, contract-for-hire, and permanent placement roles. With more than 300 active clients, Genesis10 provides access to many of the Fortune 100 firms and a variety of mid-market organizations across the full spectrum of industry verticals.
For contract roles, Genesis10 offers the benefits listed below. If this is a perm-placement opportunity, our recruiter can talk you through the unique benefits offered for that particular client. Benefits of Working with Genesis10:
- Access to hundreds of clients, most who have been working with Genesis10 for 5-20+ years.
- The opportunity to have a career-home in Genesis10; many of our consultants have been working exclusively with Genesis10 for years.
- Access to an experienced, caring recruiting team (more than 7 years of experience, on average.)
- Behavioral Health Platform
- Medical, Dental, Vision
- Health Savings Account
- Voluntary Hospital Indemnity (Critical Illness & Accident)
- Voluntary Term Life Insurance
- 401K
- Sick Pay (for applicable states/municipalities)
- Commuter Benefits (Dallas, NYC, SF)
For multiple years running, Genesis10 has been recognized as a Top Staffing Firm in the U.S., as a Best Company for Work-Life Balance, as a Best Company for Career Growth, for Diversity, and for Leadership, amongst others. To learn more and to view all our available career opportunities, please visit us at our website.
Genesis10 is an Equal Opportunity Employer. Candidates will receive consideration without regard to their race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or status as a protected veteran.
#DIG10-MN