Description
Job Summary:
We are seeking a Data Lake specialist to lead data structuring, organization, and governance, ensuring a reliable, scalable, and secure data ecosystem.
Key Highlights:
1. Expertise in data structuring, organization, and governance
2. Development of data ingestion and transformation pipelines
3. Experience with cloud storage (AWS S3, Azure Data Lake, GCP)
We are looking for a Data Lake specialist to lead the structuring, organization, and governance of the company's data.
This person will be responsible for ensuring the data ecosystem is reliable, scalable, and secure — supporting BI, Analytics, and Product teams in making data-driven decisions.
**Key Responsibilities:**
. Design, implement, and maintain data ingestion and transformation pipelines for the Data Lake;
. Integrate internal and external data sources in an automated and secure manner;
. Work with cloud storage (AWS S3, Azure Data Lake, GCP Storage);
. Implement best practices for data governance, versioning, and data quality;
. Support BI and Analytics teams in delivering trustworthy data;
. Monitor data environment performance, costs, and security.
**Requirements:**
. Experience with cloud-based Data Lakes (AWS, Azure, or GCP);
. Proficiency in ETL/ELT and tools such as Apache Airflow, Glue, Databricks, or similar;
. Advanced knowledge of Python and SQL;
. Experience handling unstructured data storage (JSON, Parquet, Avro, etc.);
. Experience in data architecture, modeling, and integration;
. Understanding of cloud security, governance, and cost management.
**Nice-to-Have:**
. Experience with Data Mesh or Data Governance frameworks;
. Knowledge of Kafka (real-time data streaming);
. AWS / Azure / GCP Data certifications;
. Experience in complex or highly scalable enterprise environments.