Beranda Loker Detail
B
Information & Communication Technology 🏢 Full Time ⭐️ Terverifikasi

Site Reliability Engineer, Traffic Solution - System Service Global

ByteDance
Singapore
Estimasi Gaji
SGD 150.000 – SGD 250.000
Live Update
15 Mei 2026
Batas Akhir
15 Mei 2027

Deskripsi Pekerjaan

ByteDance is seeking a passionate and experienced Site Reliability Engineer (Traffic Solution) to join our Global System Service team in Singapore. This team is the backbone of our global infrastructure, owning the critical services and management solutions that power ByteDance's rapidly expanding data centers around the world.

As an SRE focused on Traffic Solutions, you will be at the forefront of ensuring the reliability, scalability, and performance of the network traffic that drives iconic products like TikTok, Douyin, and Lark. You will tackle complex engineering challenges involving global load balancing, intelligent DNS routing, CDN optimization, and advanced traffic management systems at an unprecedented scale.

We are looking for engineers who are deeply passionate about automation, observability, and building resilient, self-healing systems. You will have the unique opportunity to design systems that serve billions of users and collaborate with some of the brightest minds in the industry. This role requires a proactive engineer who can drive incident response, perform root cause analysis, and implement long-term solutions to prevent recurrence.

If you thrive on solving complex infrastructure puzzles and want to make a direct, tangible impact on a global scale, this is the perfect role for you. Join us in shaping the future of real-time, large-scale traffic engineering.

Why Join ByteDance?

  • Work on world-class infrastructure at a massive scale.
  • Collaborate with top engineering talent across the globe.
  • Competitive compensation, equity, and comprehensive benefits.
  • Opportunity to innovate and drive cutting-edge technology in traffic engineering.

Tanggung Jawab

  • Design, build, and maintain highly available and scalable traffic systems for ByteDance's global data centers.
  • Automate operational tasks to improve efficiency and reliability of the traffic infrastructure.
  • Participate in an on-call rotation to respond to incidents, perform root cause analysis, and implement long-term fixes.
  • Collaborate with development teams on capacity planning and performance optimization of traffic-facing services.
  • Develop monitoring, alerting, and dashboard solutions to ensure end-to-end service health.
  • Drive architectural improvements and standardize best practices for traffic engineering and incident response.
  • Manage cross-region disaster recovery and ensure business continuity for critical traffic services.

Kualifikasi

  • Bachelor's degree or higher in Computer Science, Engineering, or a related technical field.
  • Strong experience in Site Reliability Engineering, DevOps, or Systems Engineering focusing on large-scale traffic or network systems.
  • Proficiency in at least one programming language (Go, Python, or C++).
  • Deep understanding of Linux/Unix systems internals and networking protocols (TCP/IP, HTTP/HTTPS, DNS, BGP).
  • Experience with containerization and orchestration technologies (Kubernetes, Docker) and service meshes (Istio, Envoy) is highly desirable.
  • Solid experience with incident management, root cause analysis, and blameless postmortems.
  • Excellent problem-solving skills and the ability to thrive in a fast-paced, global environment.

Keahlian yang Dibutuhkan

Site Reliability Engineering SRE Traffic Engineering Linux Administration Networking TCP/IP DNS Global Load Balancing CDN Kubernetes Docker Go Python Incident Management Automation Terraform Istio Envoy System Design

Siap Mengambil Tantangan Ini?

Pastikan resume Anda sudah siap. Kirimkan lamaran Anda sekarang sebelum tanggal deadline.

Lamar Sekarang

Lowongan Terkait

Rekomendasi pekerjaan serupa untuk Anda

Lihat Semua