Overview One of Insight Global's customers is looking to onboard a Sr. Site Reliability Engineer with strong expertise in modern DevOps practices, cloud infrastructure, observability, and platform security. This role partners directly with product teams to support deployments, build reliable systems, and strengthen platform capabilities across Kubernetes, AWS, and CI/CD pipelines. They will own the observability strategy across a growing application ecosystem. This person will partner closely with product and engineering teams, helping drive reliability, monitoring, logging, and platform scalability initiatives. This is a rolling contract and will be 5 days a week onsite in Irvine, CA.We are a company committed to creating diverse and inclusive environments where people can bring their full, authentic selves to work every day. We are an equal opportunity/affirmative action employer that believes everyone matters. Qualified candidates will receive consideration for employment regardless of their race, color, ethnicity, religion, sex (including pregnancy), sexual orientation, gender identity and expression, marital status, national origin, ancestry, genetic factors, age, disability, protected veteran status, military or uniformed service member status, or any other status or characteristic protected by applicable laws, regulations, and ordinances. If you need assistance and/or a reasonable accommodation due to a disability during the application or recruiting process, please send a request to HR@insightglobal.com.To learn more about how we collect, keep, and process your private information, please review Insight Global's Workforce Privacy Policy: Provide reliability engineering and platform engineering support for deployments and system design across Kubernetes, AWS, and CI/CD pipelines.Own the observability strategy and drive monitoring, logging, and metrics across a growing application ecosystem.Collaborate with product and engineering teams to improve platform reliability, performance, and scalability.Partner with cross-functional teams to implement best practices for security, resilience, and incident response.Qualifications 3–6+ years of Site Reliability Engineering, DevOps, or Platform Engineering experienceStrong AWS experience (GCP and Azure are a plus)Hands-on Kubernetes platform engineering (multi cluster, deployment, troubleshooting); experience building and maintaining CI/CD pipelinesObservability and monitoring experience within modern application environmentsExperience with Prometheus and Grafana for metrics, monitoring, and loggingExcellent communication and stakeholder management skillsPython and/or React development experienceExperience with Kafka or event streaming platformsNeo4j experienceAI Infrastructure experience#J-18808-Ljbffr
| Location | Irvine, CA |
| Industry | Healthcare Services |
| Company Size | 2,500 to 4,999 employees |
| Year Founded | 2001 |
| Website | http://www.insightglobal.net |
We are a staffing agency helping individuals find jobs and employers fill open positions.
Based in Atlanta, Insight Global is a premier provider of employment and staffing solutions to Fortune 1000 customers across the United States and Canada. We provide long-term contract, short-term contract, temporary-to-permanent, direct placement, and enhanced staffing services. Insight Global specializes in placing contract job seekers into Information Technology, Accounting and Finance, Engineering (non-IT), and Government jobs.
Since our inception in 2001, we have experienced unprecedented growth within our industry, rapidly expanding from an Atlanta-based start-up to one of the most successful staffing firms in America.
Our core staffing services are the backbone upon which Insight Global was founded and have driven our success. We cater our delivery approach and recruiting efforts to meet each client’s unique demands, ensuring that we deliver both maximum client value and the differentiated Insight Global experience.