Linux Admin
Redmond, WA - USA
Job Summary
Role Purpose
The purpose of this role is to assist in the triage and documentation of priority incidents support resolution of software and hardware service requests and ensure SLA adherence by escalating delays when required.
Â
Areas of responsibility
Triaging and resolution of tickets-Perform initial triage and assist in response for priority incident tickets (P3 tickets) along with its documentation and following the defined SLA and escalation procedures.
Resolution of service Requests-Track and assist in resolution of software/hardware/network service requests ensuring they are accurately logged processed and completed within defined timelines as per established procedures.
SLA Monitoring-Track and monitor SLA timelines for entire lifecycle of priority incident tickets (P3 tickets). Escalate to higher levels for delays or potential SLA breaches.
Change request execution-Assist in creation of change requests basis types of incident tickets which are raised and resolved.
JD-
Understanding of Linux operating systems and Linux system administration (SysAdmin roles).
* Good understanding of Linux/Unix commands (Strong User).
* Experience automating tasks with scripting languages such as Python Bash and JavaScript.
* Systematic problem-solving approach strong communication skills a sense of ownership and drive solutions.
* Deep understand of service metrics and alarms through the development of dashboards service KPIs alarming systems.
* Experience working in an operational environment with mission critical tier one services with associated pager duty.
* Software development experience focused on Services and Operational tools.
* CI/CD process and software knowledge.
* Operational toil mitigation and reduction.
Responsibilities
* Work closely with development team on maintaining operational health of core compute services for API availability and low latency
* Managing and triaging tickets. Driving prioritization and execution of work based on impact.
* Scale systems sustainably through mechanisms such as easy-to-use tooling and automation.
* Work in concert with service developers to evolve systems/products for better scalability reliability and development velocity.
* Drives new runbooks to help reduce mean triage time of incidents.
* Prioritize and automate high hit count runbooks.
* Practice sustainable incident response and drive root case analysis
About Company
As a global leader, Wipro blends consulting and AI expertise across design, engineering and operations to accelerate business transformation and deliver future-ready technology.