Site Reliability Engineer L3
Job Summary
Role Purpose
Required Skills:
ÃÂÂ 5Years of experience in system administration application development infrastructure development or related areas
ÃÂÂ 5 years of experience with programming in languages like Javascript Python PHP Go Java or Ruby
ÃÂÂ 3 years of in reading understanding and writing code in the same
ÃÂÂ 3years Mastery of infrastructure automation technologies (like Terraform Code Deploy Puppet Ansible Chef)
ÃÂÂ 3years expertise in container/container-fleet-orchestration technologies (like Kubernetes Openshift AKS EKS Docker Vagrant etcd zookeeper)
ÃÂÂ 5 years Cloud and container native Linux administration /build/ management skills
Key Responsibilities:
ÃÂÂ Hands-on design analysis development and troubleshooting of highly-distributed large-scale production systems and event-driven cloud-based services
ÃÂÂ Primarily Linux Administration managing a fleet of Linux and Windows VMs as part of the application solutions
ÃÂÂ Involved in Pull Requests for site reliability goals
ÃÂÂ Advocate IaC (Infrastructure as Code) and CaC (Configuration as Code) practices within Honeywell HCE
ÃÂÂ Ownership of reliability up time system security cost operations capacity and performance-analysis
Monitor and report on service level objectives for a given applications services. Work with the business Technology teams and product owners to establish key service level indicators.
ÃÂÂ Ensuring the repeatability traceability and transparency of our infrastructure automation
ÃÂÂ Support on-call rotations for operational duties that have not been addressed with automation
ÃÂÂ Support healthy software development practices including complying with the chosen software development methodology (Agile or alternatives) building standards for code reviews work packaging etc.
ÃÂÂ Create and maintain monitoring technologies and processes that improve the visibility to our applications performance and business metrics and keep operational workload in-check.
ÃÂÂ Partnering with security engineers and developing plans and automation to aggressively and safely respond to new risks and vulnerabilities.
ÃÂÂ Develop communicate collaborate and monitor standard processes to promote the long-term health and sustainability of operational development tasks.
Required Experience:
IC
About Company
As a global leader, Wipro blends consulting and AI expertise across design, engineering and operations to accelerate business transformation and deliver future-ready technology.