Job Title:  Associate | System Admin | Bengaluru | Engineering as a Service/ Operate

Job Summary:

As a technical front-line resource, the Systems Administrator II is a mentor and escalation point for junior members of the Systems Operations team. This role’s primary responsibility is to monitor, triage, and resolve or escalate issues as necessary in a world class 24/7/365 production environment. The Systems Administrator II will also act as a resource to collaborate with teams to prevent or minimize impact of potential issues. They will achieve this by blending a strong set of troubleshooting skills with a strong understanding of our infrastructure and production pipelines.

Eyes of Glass monitoring for and react to alerts within the environment providing level 1 support

Provide clear communication/escalation/follow-up/closure of alerts and maintenances tomultiple teams within the organization

Support various engineering and other internal teams involved with maintaining our serverinfrastructure; including but not limited to firmware upgrades, system provisioning anddecommissions.

Complete general day to day tickets within jira/SNOW (user-add/removal, monitoringadjustments, general permission related issues, etc.).

Interact with external parties for escalation or hardware /software support tickets.

Responsibilities and Duties of the Role:

•Participate on small to medium application and infrastructure implementation projects andensure the result conforms to System Operations monitoring standards

•Support in carrying out Operational tasks as assigned.

•Interacts and communicates with supervisor and peers. Receives and transmits routineinformation requiring some explanation or interpretation, to members of the team.

•Ability to resolve infrastructure operational issues with a moderate to high degree ofcomplexity while following established procedures or analyzing data to diagnose and identifyroot causes when necessary, to resolve issues or complete assignments.

•Communicate with internal contacts at a variety of organizational levels to diagnose, explainand resolve problems for both technical and non-technical audiences.

•Participate in supporting high profile events.

•Demonstrates a strong work ethic, initiative and the ability to aggressively and effectivelytroubleshoot problems and perform root cause analysis.

•Comfort working in a fast-paced, growth environment with minimal direction.

•Able to operate with limited resources, exhibit dispassionate critical-thinking, and creativelysolve problems. Motivated and inventive, always finds a way forward—eternal optimist witha committed drive for contributing and making a difference.

•Any other duties needed to help achieve business objectives, including participating in an on-call rotation.

•Escalate and triage monitoring alerts

Required Education, Experience/Skills/Training:

•Three or more years of progressively complex related experience.

•BA/BS degree in Computer Science or related software engineering field or equivalent practical experience

•Proactive thinker with strong problem solving skills.

•Experience with Linux (Installation, Configuration, Tuning).

•Strong written and verbal skills

•Strong troubleshooting skills.

•Excellent communication and interpersonal skills.

•Knowledge with scripting languages (BASH, Python,).

•Experience with Ansible and Ansible Tower.

•Good understanding of security and networking concepts.

•Experience with IPMI tool

•Experience utilizing infrastructure monitoring solutions. (Icinga, Zabbix)

•Able to work non-traditional hours, in non-traditional settings. This includes weekends,evenings, and holidays

Preferred :

•Three or more years of infrastructure monitoring experience

•Experience with Varnish caching

• Experience in installing and configuring Icinga2

• Experience with Icingaweb2 Interface.

• Experience Creating NAGIOS check plugins.

• Experience in installing and configuring ZABBIX

• Experience with Zabbix web interface