We are looking for a hands-on, all-round Data Lake Developer to design and develop an enterprise Data Lake & Analytics platform integrating data from multiple systems.
Key Responsibilities & Skills:
-
Strong Python & SQL development
-
Data ingestion through APIs, databases, files, batch & real-time pipelines
-
Hands-on with AWS S3 Data Lake, Iceberg
-
Experience with Apache Airflow for pipeline orchestration
-
Knowledge of Kafka for real-time/event-driven ingestion
-
ETL/ELT, data cleansing, transformation, reconciliation & data quality
-
Experience with analytical databases such as ClickHouse
-
Understanding of Data Catalog, Metadata, Data Lineage & Governance
-
Ability to design Bronze → Silver → Gold data architecture
-
Integration with Tableau/Power BI or similar BI tools
-
Good understanding of REST APIs, Git, Docker and AWS
-
Exposure to AI/LLM-based analytics is an added advantage
Experience: 3–7 years, preferably with end-to-end ownership of Data Lake/Data Engineering projects.
Ideal Candidate: A strong problem solver who can independently handle source understanding → ingestion → transformation → data modelling → analytics-ready datasets → BI integration, rather than specializing in only one part of the data pipeline.