Position: AWS Site Reliability Engineering With Dynatrace
Location: Allentown, PA (100% Remote)
Roles and Responsibility :
Supporting Incident Escalation and Trouble shooting
Documenting process and related Knowledge articles (Documenting Tribal Knowledge)
Optimizing on-call rotations & processes: Need to have skills to priorities task
Conducting Post-incident reviews (PIRs): Evaluating Incidents after resolution to build work book and automation to avoid/solve them when happen again
Understanding of all the stages of SDLC
Precise Communication:
Ability to communicate clearly and concisely.
Often the project will need them to relay important information about system alerts or outages to other members of the team.
Java Developer: who is experience in building Microservices bases architecture on AWS tech.
Problem solving: The resource should be able to debug code to identify the issue/defect and provide fix details. The engineer should be able to read logs, stack trace and other telemetry to identify the issues. Should have a good understanding of infra and network layers as well.
Monitoring tools: Should have working experience on Dynatrace, AppDynamics, DataDog, Pager Duty and Cloud Watch
Detail Dev/ DevOps skills needed
Node. JS
Memento Design Pattern
Rabbit MQ
Redis ElasticCache
Mongo DB (Clusters)
Ansible
Secrets Management
Key Management
CI/CD pipeline development
Using version control tools such as GitHub
Experience in working on AWS Tech stack
Dynamo DB
ECS Clusters
Lambda
Simple Notification Services
Simple Email Services
Simple Queue Services
Simple Storage Services
Mastered distributed computing using AWS as a tech stack