Job Title: Assistant Manager | Observability Engineer | Hyderabad | Engineering as a Service/ Operate
Sr. Systems Engineer - Observability engineer
Key Responsibilities:
System Installation & Configuration :
Install, configure, and manage Linux operating systems (SUSE, Ubuntu, Debian, Rocky Linux) and associated software/tools.
Manage system lifecycle including OS upgrades, patching, and package management.
Virtualization:
Provided day-to-day operational support for VMware vSphere environments including ESXi, vCenter Server, and virtual machine lifecycle management.
Supported Proxmox environments for virtual machine and container-based workloads.
Storage: Strong with Storage File Shares: NFS/CIFS
Network and DC: Strong HW knowledge especially in the Server, Networking, Storage, BIOS & OOB arenas.
Performance Monitoring & Optimization:
Monitor system health and performance using industry-standard tools.
Diagnose and resolve bottlenecks to ensure high availability, scalability, and reliability.
Security Management:
Implement and maintain security policies, access controls, and hardening standards.
Apply patches, updates, and vulnerability remediations in a timely manner.
Scripting & Automation:
Develop scripts using Bash, Python, or Perl to automate tasks and improve operational efficiency.
Utilize configuration management and automation tools such as Ansible, Salt, etc.
Troubleshooting & Support:
Provide advanced technical support for Linux-related issues in production environments.
Troubleshoot complex hardware, software, and system-level issues.
Collaboration:
Partner with development, networking, cloud, and DevOps teams to support application deployments and infrastructure projects.
Assist with integrating new technologies and improving existing systems.
Documentation:
Maintain accurate and detailed documentation for configurations, procedures, standards, and troubleshooting guides.
Qualifications & Skills:
Experience: Proven experience as a Linux Administrator or similar role in production environments.
Operating Systems: Strong knowledge of Linux distributions, kernel concepts, file systems, system initialization, and networking.
Automation: Hands-on experience with Ansible, Salt
Scripting: Proficiency in Bash is a strong plus.
Cloud & Virtualization: Experience with VMware, Proxmox, and cloud platforms such as AWS, Azure, or Google Cloud.
Networking: Working knowledge of TCP/IP, DNS, DHCP, VPN, SSH, and firewall concepts.
Problem-Solving: Strong analytical abilities with high attention to detail.
Communication: Excellent communication and interpersonal skills; ability to clearly articulate technical concepts.