Source description
About the role
Job_Description":" **Job details Job Title: Automated Reliability Test Engineer (Chaos Engineering) Division: Technology / Platform Engineering / SRE Years of Experience: 3+ Years Education: Bachelor/Master in Computer Science, Electronics/Electrical Engineering, Data Science, or related technical field Employment Type: Full-time Location: Hyderabad **Role Type: Permanent **Key responsibilities Design and implement automated reliability and resilience testing strategies. Develop test plans to simulate real-world system failures and identify weaknesses. Execute chaos engineering experiments covering network, hardware, and software faults. Utilize Linux as the core environment for debugging, testing, and infrastructure validation. Work with Docker and Kubernetes for containerized and orchestrated test environments. Build automation workflows using shell scripting and integrate them with testing frameworks. Implement fault injection scripts and techniques for controlled system disruption. Document failure modes and propose mitigation strategies to improve reliability. Collaborate with development, operations, and QA teams to align test execution with business needs. Communicate experiment results, impacts, and improvement recommendations to stakeholders. **Required Technical Skills Proficiency in Linux operating systems. Hands\u2011on experience with Docker and Kubernetes. Strong shell scripting capability (Bash, Shell). Experience in automated testing or automation workflows. Practical exposure to chaos tools or fault injection techniques. Familiarity with CI/CD tools (Jenkins, GitLab CI/CD). Experience with infrastructure automation tools such as Terraform or Ansible. **Key skills and experience required Strong understanding of reliability, resilience, and non\u2011functional testing concepts. Practical exposure to chaos engineering methodologies and failure simulation. Hands\u2011on experience in Linux\u2011based environments. Strong analytical and problem\u2011solving abilities for system behavior diagnosis. Ability to work cross\u2011functionally with DevOps, SRE, and development teams. Clear communication and documentation skills for experiment results and system insights. **Good to have skills and experience required Exposure to AI\u2011powered testing or automation tools. Experience with ML\u2011based system insights or anomaly detection. Background in Python scripting for automation. Understanding of performance testing or site reliability engineering (SRE) practices. Knowledge of distributed systems behavior under failure conditions. Requirements **Required Technical Skills Proficiency in Linux operating systems. Hands\u2011on experience with Docker and Kubernetes. Strong shell scripting capability (Bash, Shell). Experience in automated testing or automation workflows. Practical exposure to chaos tools or fault injection techniques. Familiarity with CI/CD tools (Jenkins, GitLab CI/CD). Experience with infrastructure automation tools such as Terraform or Ansible. **Key skills and experience required Strong understanding of reliability, resilience, and non\u2011functional testing concepts. Practical exposure to chaos engineering methodologies and failure simulation. Hands\u2011on experience in Linux\u2011based environments. Strong analytical and problem\u2011solving abilities for system behavior diagnosis. Ability to work cross\u2011functionally with DevOps, SRE, and development teams. Clear communication and documentation skills for experiment results and system insights. Benefits **Why Join Us At NPCI, youll be part of a purpose-driven organization shaping the future of digital payments in India and beyond. NPCI offers a unique opportunity to work on cutting-edge projects that directly impact millions. We foster a culture of innovation, inclusion, and high performance, where every individual is empowered to lead with purpose and deliver with passion. With a strong focus on employee wellbeing, continuous learning, and collaborative success, NPCI is more than just a workplaceit a platform to grow, contribute, and make a meaningful difference. **Life at NPCI Job_Description":" **Job details Job Title: Automated Reliability Test Engineer (Chaos Engineering) Division: Technology / Platform Engineering / SRE Years of Experience: 3+ Years Education: Bachelor/Master in Computer Science, Electronics/Electrical Engineering, Data Science, or related technical field Employment Type: Full-time Location: Hyderabad **Role Type: Permanent **Key responsibilities Design and implement automated reliability and resilience testing strategies. Develop test plans to simulate real-world system failures and identify weaknesses. Execute chaos engineering experiments covering network, hardware, and software faults. Utilize Linux as the core environment for debugging, testing, and infrastructure validation. Work with Docker and Kubernetes for containerized and orch
More at National Payments Corporation Of India (NPCI)
Related open roles
Associate/Sr. Associate SRE (Telangana)
Hyderabad
Senior Associate Software Development Engineer
Mumbai
SRE -Leadership role (Telangana)
Hyderabad
Senior Associate Infra & Networking (Chennai)
Chennai
Senior Associate - Automated Reliability Test u2013 (Development) (Hyderabad)
Hyderabad
Data / Big Data Testing Engineer
India