Padmi

Technical Duty Officer (SRE)

BangalorePosted 2 months ago
Infrastructure And DatabasesMid-level
Apply at The Networker

Opens the source posting on naukri.com

Source description

About the role

View original

Join eBay as a Site Reliability Engineer in our Site Engineering Center (SEC), where you'll play a crucial role in ensuring the availability and reliability of eBay's platform. You'll lead major incidents, manage critical service health, and collaborate with cross-functional teams to develop innovative solutions. With opportunities to enhance automation and improve monitoring tools, this role is perfect for someone with a passion for technology and a drive to make a significant impact on the customer experience of the eBay community. What you will accomplish: Lead Incident Management: Act as the Incident Commander to drive resolution of major incidents, manage alarms, and ensure effective communication with leadership and partner teams. Proactive Monitoring: Continuously monitor the health of eBay's critical services to identify and address potential issues before they escalate. Collaborative Problem Solving: Work closely with partner teams to resolve recurring technical issues, onboard new alerts, and develop high-quality Standard Operating Procedures (SOPs). Automation and Process Enhancement: Identify and implement opportunities to enhance automation and reduce manual workload, improving overall efficiency. Solution Development: Collaborate with Architecture, Engineering, and Operations teams to develop solutions that ensure high site availability and reliability. Enhance Monitoring Tools: Improve tools for monitoring and mitigating site incidents, and conduct reliability audits and tests to strengthen eBay's reliability and incident management capabilities. What you will bring: Experience in highly-scaled internet/server environments. Strong technical triage and troubleshooting skills, particularly in crisis-level incident management. Excellent incident management skills with proven leadership abilities. Experience with large enterprise server environments, including cloud computing, web services, and multi-tier architectures, with knowledge of UNIX, Linux, networking, and database technologies. Proficient in creating software solutions for task automation using technologies like GO, Python, Java, NodeJS, Docker, and Kubernetes. Familiarity with observability tools such as Grafana and Prometheus, and experience in documenting procedures and knowledge management. NOTE: As part of the operation staff members of the SEC work a fixed shift. This position is for a day shift in BanTeam members will work four days in a row in 10 hour shifts, with no on call responsibilities. Additional Details Disclaimer: This job posting has been aggregated from external source. Role details, content, and availability are subject to change. Applicants are advised to confirm the latest information directly on the company website before applying.

One address, no account. We’ll tell you when matching roles go live.

More at The Networker

Related open roles

View all roles
Technical Duty Officer (SRE) at The Networker · Padmi