Deskripsi Pekerjaan
Join Capgemini's dynamic data team as a Data Engineer and become the architect of our data ecosystem. You'll design, implement, and maintain robust ETL pipelines that transform raw data from diverse source systems into actionable insights for our central data warehouse. This role is critical in enabling data-driven decision-making across our organization while ensuring data integrity, scalability, and security. You'll collaborate with data scientists, analysts, and business stakeholders to understand requirements and deliver high-quality data solutions. At Capgemini, you'll work with cutting-edge technologies including cloud platforms (AWS/Azure), big data tools (Spark, Hadoop), and modern data warehouse solutions. This position offers continuous learning opportunities in a global technology leader committed to innovation and professional growth.
Tanggung Jawab
- Design, develop, and maintain scalable ETL/ELT pipelines for data ingestion and transformation
- Optimize data workflows for performance, reliability, and cost-efficiency
- Implement data quality checks and validation rules to ensure data accuracy
- Collaborate with cross-functional teams to understand data requirements and business needs
- Monitor and troubleshoot data pipeline issues using monitoring tools and logs
- Document data architectures, processes, and technical specifications
- Stay current with emerging data engineering technologies and best practices
Kualifikasi
- Bachelor's degree in Computer Science, Engineering, or related field
- 3+ years of experience in data engineering or ETL development
- Proficiency in SQL and database management systems (PostgreSQL, MySQL)
- Experience with cloud data platforms (AWS Glue, Azure Data Factory, GCP Dataflow)
- Strong knowledge of big data technologies (Spark, Hadoop, Kafka)
- Familiarity with data warehouse concepts and dimensional modeling
- Experience with scripting languages (Python, Shell scripting)
- Excellent problem-solving and communication skills