Job Overview
We are looking for an experienced and talented Data Engineer to join our team. In this role, you will be responsible for designing, building, and maintaining robust, scalable data pipelines (ETL/ELT processes) to support our analytics and business intelligence systems. You will work extensively with PySpark, Alteryx, and SQL to transform large-scale raw data into high-quality, actionable data assets for decision-making.
---
Key Responsibilities
1. Data Pipeline Development & Maintenance:
- Design, develop, and optimize high-performance distributed data pipelines using PySpark for large-scale structured and semi-structured datasets.
- Build, automate, and maintain efficient data workflows and data integration solutions using Alteryx Designer.
- Write, optimize, and maintain complex SQL queries, stored procedures, and data transformations for maximum performance and scalability.
2. Data Architecture & Performance Tuning:
- Contribute to the design, modeling, and management of Data Warehouses and Data Lakes.
- Monitor, troubleshoot, and optimize existing ETL/ELT workflows to ensure high data availability, accuracy, and operational efficiency.
3. Cross-Functional Collaboration:
- Work closely with data analysts, data scientists, and business stakeholders to understand data requirements and deliver clean, production-ready datasets.
Qualifications & Requirements
- Education & Experience:
- Bachelor’s or Master’s degree in Computer Science, Software Engineering, Information Systems, or a related field.
- 3+ years of hands-on experience in data engineering or big data development.
Benefit
- Annaul leave
- Sick leave
- MPF
- Medical insurance
- Five days work