Skills Required: This position requires five (5) years of experience with the following: Utilizing Sev1/Sev2 on-call including triaging, mitigating, coordinating restoration, RCA, and blameless postmortems; MTTR recurrence reduction; Observability including metrics, logs, traces; low-noise alerts, fast detection; maintaining runbooks and escalations; Automation including Python and Bash; utilizing CI/CD with blue and green or canary, feature flags, pre-deploy validation, and reliable rollback; using IaC including Terraform for public cloud; using Modules, remote state, drift detection, LUT/policy-as-code, and targeted applies; using Kubernetes for deployments, autoscaling, health probes, progressive rollouts and rollbacks, quotas, and cluster troubleshooting; utilizing AWS production operations including VPC networking, IAM policy design, encryption/ KMS, load balancing, multi-AZ resilience, backup and restore, and regional failover; Database reliability including PostgreSQL, MySQL, Oracle backups/PITR, replication and failover, online schema changes, SQL and index tuning under load; Linux and networking including processing memory/IO diagnostics, kernel and sysctl tuning, TLS, TCP/IP, DNS, HTTP, load balancing, and service discovery; Security including least privilege, secrets management, patch, vulnerability management, immutable audit logging and framework alignment. QUALIFICATIONS: Minimum education and experience required: Bachelor''s degree in Information Systems Engineering, Computer Engineering, or related field of study plus 5 years of experience in the job offered or as Site Reliability Engineer, Data Engineer, Data Analyst, MSSQL Server Developer, Support Engineer, Software Developer, or related occupation.