Job Title:  Assistant Manager | Observability Engineer | Hyderabad | Engineering as a Service/ Operate

Sr. Systems Engineer - Observability engineer

Key Responsibilities:

System Installation & Configuration :

  Install, configure, and manage Linux operating systems (SUSE, Ubuntu, Debian, Rocky Linux) and associated software/tools.

   Manage system lifecycle including OS upgrades, patching, and package management.

Virtualization:

   Provided day-to-day operational support for VMware vSphere environments including ESXi, vCenter Server, and virtual machine lifecycle management.

   Supported Proxmox environments for virtual machine and container-based workloads.

Storage: Strong with Storage File Shares: NFS/CIFS

Network and DC: Strong HW knowledge especially in the Server, Networking, Storage, BIOS & OOB arenas.

Performance Monitoring & Optimization:

   Monitor system health and performance using industry-standard tools.

   Diagnose and resolve bottlenecks to ensure high availability, scalability, and reliability.

Security Management:

   Implement and maintain security policies, access controls, and hardening standards.

   Apply patches, updates, and vulnerability remediations in a timely manner.

Scripting & Automation:

   Develop scripts using Bash, Python, or Perl to automate tasks and improve operational efficiency.

   Utilize configuration management and automation tools such as Ansible, Salt, etc.

Troubleshooting & Support:

   Provide advanced technical support for Linux-related issues in production environments.

   Troubleshoot complex hardware, software, and system-level issues.

Collaboration:

   Partner with development, networking, cloud, and DevOps teams to support application deployments and infrastructure projects.

   Assist with integrating new technologies and improving existing systems.

 

Documentation:

   Maintain accurate and detailed documentation for configurations, procedures, standards, and troubleshooting guides.

 

Qualifications & Skills:

   Experience: Proven experience as a Linux Administrator or similar role in production environments.

   Operating Systems: Strong knowledge of Linux distributions, kernel concepts, file systems, system initialization, and networking.

   Automation: Hands-on experience with Ansible, Salt

   Scripting: Proficiency in Bash is a strong plus.

   Cloud & Virtualization: Experience with VMware, Proxmox, and cloud platforms such as AWS, Azure, or Google Cloud.

   Networking: Working knowledge of TCP/IP, DNS, DHCP, VPN, SSH, and firewall concepts.

   Problem-Solving: Strong analytical abilities with high attention to detail.

   Communication: Excellent communication and interpersonal skills; ability to clearly articulate technical concepts.