VLink, founded in 2006, is a leading global provider of software engineering services with next-gen technologies and best-in-class talent. Our Headquarters are in the U.S, and we have offices in 7+ countries from North America-Europe to APAC, with expansion plans in the Middle East. With over 1,000 employees working globally, VLink has helped SMBs, and large enterprises achieve their business goals, and gained the trust of Fortune-250 companies. VLink is 'Great Place to Work CertifiedTM' and has been a consistent winner as- Best Places to Work in CT. Trust, collaboration, and accountability are the three elements that are at the core of VLink's work culture. We value our professionals, providing comprehensive benefits and the opportunity for growth.
Provide L2/L3 production support for enterprise applications built with React, Python, APIs, and microservices.
Troubleshoot UI, backend service, integration, database, cache, and application performance issues.
Support MongoDB operations, including data validation, connectivity, query performance, and indexing issues.
Monitor and troubleshoot Redis Cache availability, memory utilization, expiration, and synchronization problems.
Support AWS-hosted applications, including React UI deployments, Amazon S3, CloudFront, and related cloud services.
Use Dynatrace and Splunk for application monitoring, log analysis, distributed tracing, alert investigation, and dashboards.
Manage production incidents, participate in incident bridges, perform root-cause analysis, and implement preventive actions.
Develop Python scripts to automate health checks, log analysis, operational reporting, and repetitive support activities.
Use Generative AI, prompt engineering, RAG, and AI-assisted tools for incident summarization, troubleshooting, and knowledge management.
Support releases, maintain runbooks, communicate with stakeholders, and participate in rotational on-call support.
Provide L2/L3 production support for enterprise applications developed using React, Python, APIs, microservices, MongoDB, Redis, and AWS services.
Monitor application availability, performance, and reliability, and proactively identify production risks and service degradation.
Troubleshoot end-to-end issues across the UI, backend services, APIs, integrations, databases, caching layers, and cloud infrastructure.
Support MongoDB operations, including connectivity, data validation, query optimization, indexing, and performance troubleshooting.
Monitor and resolve Redis Cache issues related to availability, memory utilization, key expiration, data synchronization, and application connectivity.
Support AWS-hosted applications and deployments involving Amazon S3, CloudFront, React UI components, and related cloud services.
Use Dynatrace and Splunk for log analysis, distributed tracing, alert investigation, performance monitoring, dashboarding, and production diagnostics.
Manage production incidents by performing impact assessment, participating in incident bridges, coordinating resolution, completing root-cause analysis, and implementing preventive actions.
Develop Python-based automation for application health checks, log analysis, alert enrichment, operational reporting, and repetitive support activities.
Apply Generative AI, prompt engineering, RAG, and AI-assisted tools to accelerate incident analysis, generate incident summaries, support troubleshooting, and improve knowledge management.
Support application releases by completing readiness checks, validating deployments, monitoring post-release performance, and coordinating rollback or remediation activities when required.
Create and maintain runbooks, standard operating procedures, troubleshooting guides, knowledge articles, incident records, and operational documentation.
Communicate incident status, risks, technical findings, and recovery progress clearly to business stakeholders, development teams, infrastructure teams, and leadership.
Participate in rotational on-call support and continuously improve application stability, support efficiency, monitoring coverage, and incident prevention.
6 12 years of overall IT experience, including at least 4 years of L2/L3 production support for enterprise and business-critical applications.
Strong hands-on experience supporting applications developed using React, Python, REST APIs, microservices, MongoDB, Redis Cache, and AWS services.
Experience troubleshooting end-to-end production issues across the UI, backend services, APIs, integrations, databases, caching layers, and cloud infrastructure.
Experience designing and building an enterprise alerting framework for proactive monitoring, alert correlation, notification, escalation, and incident prevention.
Experience supporting Doctor Tools and clinician-facing applications, including application availability, integrations, workflow issues, and production performance.
Strong experience managing incident and change queues, including ticket prioritization, assignment, SLA tracking, technical analysis, change validation, stakeholder communication, and timely closure.
Experience supporting MongoDB, including connectivity, query performance, indexing, data validation, and production troubleshooting.
Working knowledge of Redis Cache, including availability, memory utilization, key expiration, synchronization, connectivity, and performance issues.
Practical experience using AI tools, Generative AI, prompt engineering, and RAG for incident analysis, troubleshooting, operational automation, incident summarization, and knowledge management.
Experience supporting, migrating, and operating AI-enabled IVR and contact-center solutions, including Kore.ai, Sierra AI for Health Services, Amazon Connect Outbound, Doctor Tools, and related AI tools, covering integrations, monitoring, incident resolution, production stability, and continuous improvement.
Experience supporting AWS-hosted applications involving Amazon S3, CloudFront, application deployments, monitoring, and related AWS services.
Strong experience using Dynatrace and Splunk for log analysis, distributed tracing, dashboarding, alert investigation, performance monitoring, and root-cause analysis.
Proven experience managing major incidents, incident bridges, problem management, root-cause analysis, corrective actions, and preventive measures.
Experience developing Python automation scripts for health checks, log analysis, alert enrichment, operational reporting, and repetitive support activities.
Experience supporting production releases, deployment validation, post-release monitoring, rollback coordination, runbooks, SOPs, and knowledge documentation.
Ability to participate in rotational on-call support and communicate effectively with business stakeholders, client teams, development teams, infrastructure teams, and leadership.
VLink is an equal opportunity employer committed to fostering an inclusive environment where diversity is celebrated. All qualified applicants will be considered for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or veteran status. Employment is contingent upon successful completion of a background check.
This job application process may use AI-powered tools to assist in screening and evaluating applications based on objective, job-related qualifications. AI is used solely to support the recruitment process and all final hiring decisions are made exclusively by our human recruitment team.
Applicant information will be handled in accordance with VLink's privacy policy.
| Location | Lafayette, LA |
| Industry | Computer/IT Services |
| Salary | $65–$70 Per Hour |
| Company Size | 100 to 499 employees |
Started in 2006, VLink has built a solid foundation of providing end-to-end project delivery services, IT services, and talent acquisition solutions to various clients of all sizes -from small, medium to large Fortune 500 companies. We have a stellar history of providing continuous and superior quality services to our customers. This is a testament to the quality of our employees, who are our greatest asset. We believe providing our employees with the tools, training, and processes will not only improve their skills but provide a high caliber of service to our clients. This employee-centric approach, coupled with our financial stability and retention policies and procedures, has resulted in an employee turnover rate well below industry averages. In addition, over the past twelve years, VLink has created a robust database of thousands of pre-screened candidates who may be actively seeking new opportunities. Presently, VLink has over 400+ employees working on-site at client sites in the United States, and in our offshore delivery centers in India and Indonesia.
Browse dozens of recruiter approved resume templates, layouts and formats. Choose your favorite and make it your own in minutes.
Free resume templatesImprove your existing resume or start from scratch and create a standout, ATS-friendly resume. Add job-specific content, download and apply.
Free resume builder