Must Have Technical/Functional Skills
24x7 data center operations monitoring, including system, server, application, and network availability. Hands-on experience with incident monitoring, triage, escalation, and follow-up with internal support teams, vendors, and service providers. Strong understanding of facilities alarm monitoring, environmental alert handling, and escalation procedures. Experience managing after-hours operational support, including evenings, nights, weekends, and holiday coverage. Proficiency in monitoring and responding to ServiceNow incidents, requests, and tickets within defined SLA timelines. Knowledge of network availability monitoring and the ability to identify, escalate, and track outages or performance degradation. Experience opening, updating, and tracking incident tickets through resolution. Familiarity with data center environmental monitoring, health checks, and periodic operational testing. Understanding of ITIL-based incident management, escalation workflows, and operational support processes. Strong documentation, communication, and coordination skills for reporting issues and ensuring timely resolution across support teams.
Roles & Responsibilities
Monitor system, server, application, and network availability across the data center environment. Identify failures, outages, and performance issues, and promptly escalate them to the appropriate support teams. Perform continuous facilities alarm and environmental monitoring to ensure data center stability and compliance. Respond to operational alerts and incidents during after-hours shifts, including nights, weekends, and holidays. Monitor, update, and manage ServiceNow tickets within defined SLA requirements. Open incident tickets with vendors and service providers, and track them through closure. Coordinate with internal infrastructure, network, application, and facilities teams during incidents and service disruptions. Perform periodic checks and testing of data center systems, environmental controls, and monitoring tools. Maintain accurate incident logs, escalation records, and operational documentation. Ensure timely communication of outages, critical events, and status updates to stakeholders and support teams. Support incident response, troubleshooting, and service restoration activities in line with operational procedures. Contribute to continuous improvement of monitoring, escalation, and support processes.
TCS Employee Benefits Summary:
Discretionary Annual Incentive. Comprehensive Medical Coverage: Medical & Health, Dental & Vision, Disability Planning & Insurance, Pet Insurance Plans. Family Support: Maternal & Parental Leaves. Insurance Options: Aut& Home Insurance, Identity Theft Protection. Convenience & Professional Growth: Commuter Benefits & Certification & Training Reimbursement. Time Off: Vacation, Time Off, Sick Leave & Holidays. Legal & Financial Assistance: Legal Assistance, 401K Plan, Performance Bonus, College Fund, Student Loan Refinancing.
#LI-KR3
Salary Range-$60,000-$70,000 a year
| Location | Bloomington, IL |
| Salary | $60,000–$70,000 Per Year |