5+ years' extensive experience of Site Reliability Engineer across all Phases of the software lifecycle
Experienced in Azure or AWS cloud technologies
Experience in building CI/CD pipeline automation, tooling (Github, Jenkins, Artifactory, and Docker) and Compliance as code;
Have an in-depth understanding of miniapps and microservice architecture, API management, and distributed systems concepts;
Experience with cloud services is essential, in particular, our core Azure and AWS Technologies (EC2, ECS, Lambda, S3, SQS, SNS, & Cloud Watch);
Experience hands-on Infrastructure as Code abilities e.g. Terraform, Ansible, Packer, Python.
Ability to develop code and work with
testers and
developers
Ability to lead and coach on code reviews between Developers and Quality Engineering.
Experienced and highly capable in continually developing and balancing technical and soft skills with an understanding that making great software requires both problem identification and prevention
Strong communication, documentation, organization, and time management skills.
Be the gate keeper for all things quality in Production environment and strive to ensure 99.99% service availability.
Ability to define and articulate what good looks like in terms of key metrices for Incidences, Problem Records and Service Availability.