Job Title: Senior Associate | Site Reliability Engineer (SRE) | Bengaluru | Engineering as a Service/ Operate
April 2025
JOB DESCRIPTION
Job Title:
Senior Site Reliability Engineer
Location:
Bangalore
Date:
Reason:
Job Family:
SRE
Full-Time/Part-Time:
FTE
Job Level:
P3
FLSA (Exempt/Non-Exempt):
Department/Group Overview:
Describe the department or group. This information will be used to help attract and engage candidates.
Disney Entertainment and ESPN Product & Technology
Technology is at the heart of Disney’s past, present, and future. Disney Entertainment and ESPN Product & Technology is a global organization of engineers, product developers, designers, technologists, data scientists, and more – all working to build and advance the technological backbone for Disney’s media business globally.
The team marries technology with creativity to build world-class products, enhance storytelling, and drive velocity, innovation, and scalability for our businesses. We are Storytellers and Innovators. Creators and Builders. Entertainers and Engineers. We work with every part of The Walt Disney Company’s media portfolio to advance the technological foundation and consumer media touch points serving millions of people around the world.
Here are a few reasons why we think you’d love working here:
1.Building the future of Disney’s media: Our Technologists are designing and building the products and platforms that will power our media, advertising, and distribution businesses for years to come.
2.Reach, Scale & Impact: More than ever, Disney’s technology and products serve as a signature doorway for fans' connections with the company’s brands and stories. Disney+. Hulu. ESPN. ABC. ABC News…and many more. These products and brands – and the unmatched stories, storytellers, and events they carry – matter to millions of people globally.
3.Innovation: We develop and implement groundbreaking products and techniques that shape industry norms, and solve complex and distinctive technical problems.
April 2025
JOB DESCRIPTION
Product Engineering is a unified team responsible for the engineering of Disney Entertainment & ESPN digital and streaming products and platforms. This includes product engineering, media engineering, quality assurance, engineering behind personalization, commerce, lifecycle, and identity.
The Infrastructure Reliability Engineering (IRE) team is a new group focused on ensuring the stability, performance, and scalability of our company's infrastructure. We are a team of engineers who build tools to test and verify our infrastructure, automate issue remediation, and create self-healing systems. Our work is critical to providing a seamless and reliable experience for our customers.
Job Summary:
Describe what the person will do in the role - how he/she will impact the organization.
As a Senior Site Reliability Engineer on the Infrastructure Reliability team, you will design, build, and maintain the software and tools that test and verify the performance and reliabilityof our infrastructure. You will work closely with other engineers to identify potential issues, develop automated tests, and create solutions that ensure our systems are robust and resilient.
In this role, you will own the solutions you build, collaborating with cross-functional teams in a fast-paced environment to ensure successful implementation. You will be responsible for continuously improving service resiliency through automation, identifying and resolving performance issues, and conducting capacity planning.
Additionally, you will participate in incident reviews, assist in root cause analysis, and deliver SRE solutions in a globally distributed, multi-cloud hybrid environment (AWS, GCP, and On-prem) to ensure the highest level of uptime and Quality of Service (QoS) for internal customers.
Responsibilities and Duties of the Role:
Summarize job responsibilities, core deliverables and major duties. What is required for the position to exist?
-Focus on major areas of work, typically 20% or more of role
% of Time
Design and develop software tools for infrastructure testing and verification.
40%
Create and maintain automated testing frameworks for our infrastructure.
30%
Collaborate with infrastructure and development teams to identify and resolve reliability issues.
20%
Participate in code reviews and contribute to the team's high standards for software quality.
10%
April 2025
JOB DESCRIPTION
Required Education, Experience/Skills/Training:
Minimum and Preferred. Inclusive of Licenses/Certs (include functional experience as well as behavioral attributesand/or leadership capabilities)Basic Qualifications
Required Skills and Experience:
•Bachelor’s degree in computer science or a related field, or equivalent experience.
•8+ years of software development experience.
•Proficiency in at least one programming language (e.g., Python, Go, Java).
•Proficiency in Kubernetes administration, modern CI/CD techniques and Infrastructure as Code (IaC).
•Deep understanding of Linux operating systems and TCP/IP fundamentals.
•Experience with software testing methodologies and tools.
•Proficient in monitoring, metrics gathering, APM, container management, and log collection tools.
•Creative problem solver with excellent debugging skills and great documentation abilities.Preferred Qualifications
•Experience with performance and chaos engineering.
•Experience with CI/CD pipelines.•Understands complex system architectures and infrastructures.•Passion for automation, scalability, and building reliable systems from the ground up.
•Familiarity with cloud infrastructure (AWS, Azure, GCP).