Deskripsi Pekerjaan
Join MaiStorage as an AI Infrastructure Engineer and shape the future of on-premise AI deployment. We're seeking a tech visionary to architect, implement, and optimize cutting-edge AI infrastructure solutions that power our clients' transformative AI initiatives. In this pivotal role, you'll bridge the gap between AI innovation and scalable infrastructure, ensuring seamless end-to-end deployment of machine learning models and frameworks.
Your expertise will directly impact the performance and reliability of AI systems across diverse industries. You'll collaborate with data scientists and ML engineers to translate complex requirements into robust infrastructure designs, while maintaining security, scalability, and cost-efficiency. This position offers the unique opportunity to work at the intersection of artificial intelligence and enterprise infrastructure, driving technological excellence in Malaysia's growing tech ecosystem.
MaiStorage provides a dynamic environment where your skills will accelerate real-world AI adoption. If you're passionate about building the backbone of tomorrow's AI applications and thrive in hands-on technical challenges, this role is your chance to make a tangible impact on the future of intelligent systems.
Tanggung Jawab
- Design and implement end-to-end on-premise AI infrastructure solutions including GPU clusters and high-performance computing systems
- Deploy and maintain AI frameworks (TensorFlow, PyTorch) and containerized environments using Kubernetes and Docker
- Optimize infrastructure for AI workloads, ensuring scalability, reliability, and cost-efficiency
- Collaborate with data scientists to translate ML requirements into robust infrastructure specifications
- Monitor system performance, troubleshoot bottlenecks, and implement proactive maintenance protocols
- Document infrastructure configurations, deployment procedures, and technical specifications
- Stay current with emerging AI infrastructure technologies and best practices
Kualifikasi
- Bachelor's degree in Computer Science, Engineering, or related technical field
- 3+ years experience in IT infrastructure with focus on AI/ML environments
- Expertise in Linux system administration, networking, and cloud technologies (AWS/GCP/Azure)
- Strong knowledge of GPU optimization, CUDA, and distributed computing frameworks
- Experience with container orchestration (Kubernetes, Docker) and CI/CD pipelines
- Proficiency in infrastructure-as-code tools (Terraform, Ansible)
- Problem-solving skills with ability to debug complex distributed systems