Source description
About the role
Team Introduction Global Infra Reliability and Excellence (Global RE) is responsible for IT infrastructure reliability across AMS, EMEA, and APAC, covering both network and system domains. The team drives architecture evolution, regional operations, service resilience, delivery standards, automation, and roadmap execution for global offices, POPs, backbone connectivity, and core infrastructure services. Our direction is to enhance a globally integrated network and system team, strengthen regional ownership, shift operational capability left, and coordinate cross-region execution through virtual teams. This Singapore-based role will focus on APAC network architecture, office and workplace network delivery, POP and service operations, project acceptance, user experience, and local risk closure while aligning with global architecture principles, service baselines, and delivery and operations playbooks.
Responsibilities
- Own architecture design, buildout, optimization, and operations governance for global office networks, workplace networks, data center access, and cross-border backbone connectivity, ensuring high availability, performance, security, and scalability. - Lead capacity expansion for regional office network, research and development, and operations scenarios, ensuring delivery meets quality, cost, and timeline expectations. - Design enterprise network architecture aligned with IT strategy and business growth, including campus networks, wireless networks, WAN, SD-WAN, data center interconnect, cloud connectivity, IPv4 and IPv6 dual-stack coverage, network security boundaries, and disaster recovery solutions. - Act as the technical owner for key projects, driving cross-region network build, transformation, migration, and incident remediation initiatives; coordinate network, system, security, facilities, procurement, legal and compliance teams, carriers, and external vendors to identify and resolve project risks. - Build and maintain global network delivery standards, design templates, implementation guidelines, acceptance criteria, operations SOPs, MOPs, and emergency response playbooks to create repeatable network delivery methodology across countries and regions. - Lead troubleshooting and root cause analysis for complex network issues, own major incident response, RCA reviews, and long-term improvements, and reduce the impact of network risk on business continuity. - Track and evaluate emerging technologies such as SD-WAN, SDN, SASE, Zero Trust, SRv6, Wi-Fi 6, Wi-Fi 7, network observability, and automation; apply them to architecture evolution, operational efficiency, and cost optimization. - Regularly assess office network infrastructure capacity, performance, stability, and security risk, and drive improvements across link quality, wireless experience, network device lifecycle, configuration compliance, and asset efficiency.
More at ByteDance
Related open roles
Site Reliability Engineer - Video Infrastructure
San Francisco Bay Area
Site Reliability Engineer (Cloud) - Infrastructure Engineering
Singapore
Datacenter Operations Engineer
Saudi Arabia
Site Reliability Engineer - Traffic Infrastructure
Singapore
Traffic Access Architectural SRE Graduate (Traffic Infrastructure) - 2026 Start (BS/MS)
Singapore
Site Reliability Engineer, Traffic Solution - System Service Global
Singapore