Beranda Loker Detail
B
Information & Communication Technology 🏢 Full Time ⭐️ Terverifikasi

Site Reliability Engineer, System - System Service Global

ByteDance
Singapore
Estimasi Gaji
SGD 100.000 – SGD 160.000
Live Update
15 Mei 2026
Batas Akhir
15 Mei 2027

Deskripsi Pekerjaan

Are you ready to build the infrastructure that powers the world's most popular content platforms? ByteDance is seeking a skilled Site Reliability Engineer (SRE) to join our Global System Service team. In this pivotal role, you will own the infrastructure services and management solutions that power ByteDance's extensive data centers outside of China. We are looking for individuals who are passionate about scalability, reliability, and automation to help us maintain the high standards of our global infrastructure.

The Global System Service team bridges the gap between software development and IT operations, ensuring that our services are resilient, secure, and performant at scale. You will work in a dynamic environment where your code directly impacts the user experience of millions of users worldwide. By leveraging modern cloud technologies and robust engineering practices, you will help us drive efficiency and innovation across our data center operations.

If you thrive in a fast-paced, collaborative environment and want to solve complex technical challenges, we want to hear from you.

Tanggung Jawab

  • Design, build, and maintain scalable infrastructure services for global data centers.
  • Implement and manage automation tools to streamline deployment and operations workflows.
  • Monitor system health and performance to ensure high availability and low latency.
  • Collaborate with software engineering teams to integrate reliability practices into the development lifecycle.
  • Conduct incident response and root cause analysis to improve system resilience.
  • Optimize resource utilization and reduce operational costs through technical improvements.
  • Develop and maintain runbooks and documentation for complex systems.

Kualifikasi

  • Bachelor’s degree in Computer Science, Engineering, or a related technical field.
  • 3+ years of experience as a Site Reliability Engineer, DevOps Engineer, or in a similar technical role.
  • Proficiency in programming languages such as Python, Go, or Java.
  • Strong experience with containerization and orchestration technologies, specifically Kubernetes.
  • Deep understanding of Linux system administration and networking protocols.
  • Experience with cloud platforms (AWS, Azure, or GCP) and infrastructure as code (Terraform, Ansible).
  • Familiarity with monitoring and logging tools (Prometheus, Grafana, ELK Stack).

Keahlian yang Dibutuhkan

Kubernetes Python Go Java AWS Azure Terraform Ansible Docker Linux SRE CI/CD Prometheus Grafana Data Center Operations

Siap Mengambil Tantangan Ini?

Pastikan resume Anda sudah siap. Kirimkan lamaran Anda sekarang sebelum tanggal deadline.

Lamar Sekarang

Lowongan Terkait

Rekomendasi pekerjaan serupa untuk Anda

Lihat Semua