Deskripsi Pekerjaan
Are you an experienced Cloud Professional looking for a long-term, stable career with the comfort of a Permanent Work-From-Home setup? Scalable OS is seeking high-caliber L2 and L3 Cloud Operations Engineers to join our dynamic team. This role is specifically designed for technical experts who thrive in high-availability environments and are comfortable working on a night shift schedule to support our global clientele.
As a Cloud Operations Engineer, you will be at the forefront of managing and optimizing complex cloud infrastructures. You will play a critical role in ensuring system reliability, scalability, and security across multi-cloud environments (AWS/Azure/GCP). This is an excellent opportunity for an individual who enjoys deep-dive troubleshooting and implementing automation to drive operational excellence. At Scalable OS, we value innovation and provide a collaborative platform where your technical contributions directly impact our partners' success.
We offer a competitive compensation package, comprehensive benefits, and a culture that champions work-life balance through remote work. If you are a proactive problem-solver with a passion for cloud technology and a commitment to maintaining 99.9% uptime, we want to hear from you.
Tanggung Jawab
- Monitor, maintain, and optimize multi-cloud infrastructure to ensure maximum uptime and performance.
- Serve as the escalation point for complex L2 and L3 technical incidents, providing rapid resolution and root cause analysis.
- Automate routine operational tasks using Infrastructure as Code (IaC) tools like Terraform, Ansible, or CloudFormation.
- Manage cloud security configurations, including IAM roles, security groups, and compliance monitoring.
- Collaborate with DevOps and Development teams to streamline CI/CD pipelines and deployment processes.
- Perform regular system health checks, backup management, and disaster recovery testing.
- Implement and manage monitoring and alerting solutions such as CloudWatch, Datadog, or New Relic.
- Document technical procedures, system architectures, and incident reports to maintain a robust knowledge base.
Kualifikasi
- Bachelor’s degree in Computer Science, Information Technology, or a related field.
- Minimum of 3-5 years of professional experience in Cloud Operations, Systems Administration, or Site Reliability Engineering.
- Strong expertise in at least one major cloud provider (AWS, Azure, or GCP); multi-cloud experience is a significant plus.
- Hands-on experience with Linux/Unix administration and shell scripting (Python, Bash, or PowerShell).
- Solid understanding of networking concepts (DNS, VPN, VPC, Load Balancing).
- Experience with containerization and orchestration tools like Docker and Kubernetes.
- Professional certifications such as AWS Certified SysOps Administrator or Azure Administrator Associate are highly preferred.
- Excellent communication skills and the ability to work effectively in a night shift, remote team environment.