Senior Data Engineer

🏢 Global Corporation
📍 Dubai, United Arab EmiratesFull-timeOn-site
📅 Posted: 1mo ago🔄 Updated: 1mo ago
CV%
✨ AI Summary
We are seeking an experienced Senior Data Engineer with a strong background in PySpark, Python, and Cloudera Data Platform (CDP) to join our data engineering team in Dubai. This role involves designing, building, and optimizing large-scale data pipelines, developing ETL/ELT solutions, and ensuring data quality and governance for enterprise-scale digital products and transaction banking initiatives. Key responsibilities include developing robust batch and distributed data processing solutions, designing data ingestion frameworks, optimizing Spark jobs, and collaborating with various stakeholders. The ideal candidate possesses 8-10+ years of data engineering experience, strong SQL skills, and a solid understanding of big data technologies and best practices. Experience in banking, financial services, and digital products is preferred.
Required Skills
Information Technology
○PySpark○Python○Data Platforms○Data Quality○Data Governance○SQL○Performance Optimization○Data Modeling○Data Transformation○Git○Version Control○CI/CD○REST API○Data Integration○Debugging
Engineering, Construction & Trades
○Gas Processing
Other
○ETL/ELT pipeline development○data ingestion○ownership mindset
Science & Research
○Optimization
Soft Skills & Professional Competencies
○Analytical Skills○Problem Solving○Communication○Stakeholder Management○Collaboration○Prioritization
Business, Sales & Management
○Agile
Nice to have:
Information Technology
○Cloud Native○Machine Learning○Feature Engineering○Orchestration○Docker○Kubernetes○DevOps
Requirements
The ideal candidate will have 8-10+ years of experience in data engineering with strong expertise in PySpark, Python, and Cloudera Data Platform (CDP). Key responsibilities include designing, building, and optimizing large-scale data pipelines, developing ETL/ELT processes, ensuring data quality and governance, and optimizing Spark jobs. Strong SQL programming, data modelling, and version control skills are required, along with analytical and problem-solving competencies.
Description
Job description / Role Job Type Full Time Job Location Dubai, UAE Nationality Any Nationality Salary Not Specified Gender Not Specified Arabic Fluency Not Specified Job Function IT - Software & Web Development Company Industry Finance, Investment & Asset Management We are looking for an experienced senior data engineer with strong expertise in PySpark, Python, and Cloudera Data Platform (CDP) to join a high-performing data engineering team supporting enterprise-scale digital products and transaction banking initiatives. The ideal candidate will have extensive experience designing, building, and optimizing large-scale data pipelines within modern big data ecosystems. This role requires strong technical expertise in distributed data processing, cloud-native data platforms, data quality, and enterprise data engineering best practices. The successful candidate will work closely with data architects, data scientists, analytics teams, product owners, and business stakeholders to deliver scalable, secure, and high-performance data solutions that power business intelligence, analytics, and machine learning initiatives. Requirements Key responsibilities Design, develop, and maintain scalable, high-performance data pipelines using PySpark and Python. Build robust batch and distributed data processing solutions using Cloudera Data Platform (CDP). Develop and optimize ETL/ELT pipelines for structured and unstructured enterprise datasets. Design scalable data ingestion frameworks from multiple enterprise data sources. Ensure high data quality, integrity, governance, and availability across enterprise platforms. Perform data profiling, cleansing, transformation, and validation activities. Optimize Spark jobs for performance, scalability, and resource utilization. Work closely with data scientists to prepare datasets for analytics and machine learning use cases. Collaborate with product owners, business analysts, architects, and cross-functional engineering teams. Monitor, troubleshoot, and resolve production data pipeline issues. Participate in code reviews and implement engineering best practices. Create technical documentation and maintain data engineering standards. Support continuous improvement of enterprise data platforms and engineering processes. Participate in Agile/Scrum ceremonies including sprint planning, backlog grooming, stand-ups, and retrospectives. Required technical skills 8–10+ years of experience in data engineering. Strong hands-on expertise in Python. Strong hands-on expertise in PySpark. Extensive experience with Cloudera Data Platform (CDP). Strong understanding of distributed data processing frameworks. Experience building enterprise-scale ETL/ELT pipelines. Strong knowledge of big data technologies. Experience with Hadoop ecosystem technologies. Strong SQL programming and query optimization skills. Experience with data modelling and data transformation techniques. Knowledge of data quality, validation, and governance principles. Experience working with Git and version control systems. Understanding of CI/CD practices for data engineering. Strong understanding of REST APIs and data integration patterns. Experience working with structured, semi-structured, and unstructured datasets. Nice to have Experience working with cloud-native data platforms. Exposure to machine learning data pipelines. Knowledge of feature engineering and data preparation for AI/ML workloads. Experience with workflow orchestration tools. Exposure to containerization technologies such as Docker and Kubernetes. Experience with DevOps practices for data engineering. Required competencies Strong analytical and problem-solving skills. Excellent communication and stakeholder management skills. Ability to work in fast-paced Agile delivery environments. Strong ownership mindset with focus on quality and delivery. Ability to collaborate effectively with business and technical stakeholders. Strong debugging and performance optimization capabilities. Ability to manage multiple priorities and deliver within tight timelines. Preferred domain experience Banking Financial services Digital products Transaction banking Enterprise data platforms Apply Now
✨ Premium Match Details
Deep-dive CV analysis, customized Cover Letters, and Interview prep!
📊 Match Analysis
Insights against your active CV
📊
Personalized Match Analysis
Upload your CV to see exact matching percentages, detailed skills mapping, and gap analysis for this role.
🎯 Overalli74%
⚡ Skillsi85%
View Breakdown
Ontology Match: 85.0
Matched:✓ Requirements Matching✓ Ontology Skills Mapping
📜 Eligibilityi49%
View Breakdown
Local: 19600%
🏗️ Career Fiti91%
View Breakdown
Seniority: 91.0
📋 Requirementsi67%
View Breakdown
Domain: 67.0
🔥 Motivationi78%
View Breakdown
Title Fit: 78.00