AWS Data Lakehouse Engineer- (Permanent/Contract)
Overview
We are seeking a talented and experienced to join the Data Solutions domain in EMEA In this role you will design build Data Lakehouse Engineer and optimize scalable Lakehouse and data platform solutions that enable data driven decision-making across the organization You will collaborate closely with cross functional teams to create resilient ingestion and transformation pipelines while ensuring reliability performance and high quality data across the ecosystem
Key Responsibilities
- Contribute to the design and implementation of Lakehouse integration architectures including data flows process flows and ELTETL patterns
- Build optimize and maintain scalable data pipelines for ingestion processing and storage across data lake and Lakehouse layers
- Implement and uphold data governance security lineage and compliance best practices throughout the data lifecycle
- Monitors troubleshoot and finetune data workflows compute jobs and storage layers to ensure high performance and reliability
- Provide production support ensuring robust system observability with proactive detection and resolution of issues
- Collaborate with platform analytics and engineering teams to ensure efficient data modeling table design and storage optimization eg
Iceberg Delta patterns
Technical Qualifications
- 7 years of experience in a data engineering or similar role within a fast-paced enterprise scale environment
- Bachelor’s or master’s degree in computer science data engineering or a related technical field
- Strong Proficiency in SQL and Data Integration Platforms such as IBM Stream sets or Informatica with thorough understanding of data transformation and orchestration processes
- Extensive Experience with AWS services including IAM Lambda EKS S3 Data Lake EMR PySpark MWAA and Lakehouse technologies Apache Iceberg
- Proven expertise with Snowflake including performance tuning and efficient query design Experience with batch pipelines eg Stream Sets and streaming technologies eg Kafka
Nice to have
- Data Transformation Frameworks such as DBT
- Experience with both batch Stream Sets and streaming Kafka data solutions
- Familiarity with data modelling concepts Inmon Kimball Data Vault etc
- Exposure to CICD or Infrastructure as code Terraform or any other is a plus
- Application Performance Monitoring Framework experience is a plus
- Knowledge of data quality and security best practices is advantageous
- Ability to understand complex architecture solutions and make sound estimates for the implementation
Nontechnical Skills
- Strong attention to detail with a proactive approach to identifying and resolving problems
- A genuine interest in emerging data technologies and a curiosity about continuous improvements
- Excellent verbal and written communication skills
- Attention to detail mindset proactively identify problems evaluate solutions
- Comfort working in an AI empowered automation driven environment
Skills
Mandatory Skills : AWS Lambda, Data Lakehouse Architecture, ETL Concepts
Good to Have Skills : DBT, Snowflake