We are looking for a seasoned Data Engineer with over five years of experience in designing and implementing cloud-based data solutions, primarily focusing on Google Cloud Platform (GCP) and Microsoft Azure, with additional exposure to AWS. The ideal candidate will have a strong background in building scalable data platforms, orchestrating efficient data pipelines, and enabling analytics using Microsoft Fabric, BigQuery, and related technologies.
Key Responsibilities
- Design, build, and maintain end-to-end data ingestion, transformation, and analytics solutions across multi-cloud environments including GCP, Azure, and AWS.
- Develop and manage ETL/ELT pipelines, data models, and data warehouses to support business intelligence and analytics needs.
- Implement data lakehouse architectures and Fabric-based workflows to unify data analytics processes.
- Integrate and streamline data pipelines, dataflows, lakehouses, and notebooks within Microsoft Fabric for optimized ingestion, transformation, and reporting.
- Collaborate with analytics teams to deliver dynamic data models, semantic layers, and interactive dashboards using Power BI and Looker Studio.
- Develop reusable and scalable scripts in SQL and Python for data transformation, validation, and orchestration.
- Utilize orchestration tools such as Airflow (Composer), Azure Data Factory (ADF), and Databricks Workflows to automate data jobs.
- Ensure data accuracy, security, and performance optimization across various database systems including Azure SQL, PostgreSQL, BigQuery, MySQL, and SQL Server.
- Apply version control best practices using Git and participate in CI/CD pipelines for deploying and maintaining data pipelines and cloud artifacts in multi-environment setups.
- Communicate effectively with both technical and business stakeholders to translate complex requirements into actionable data solutions.
Required Qualifications
- Minimum of 5 years of hands-on experience in data engineering with a strong focus on cloud platforms, especially GCP and Azure.
- Proficiency with GCP services such as BigQuery, Dataflow, Pub/Sub, Composer (Airflow), Cloud Storage, and Cloud Functions.
- Expertise in Azure services including Azure Data Factory, Azure Databricks, Azure Synapse Analytics, Azure Data Lake Storage, and Azure SQL Database.
- Familiarity with Microsoft Fabric components like Data Pipelines, Dataflows Gen2, Lakehouses, and Notebooks.
- Working knowledge of AWS data services such as AWS Glue, S3, Lambda, and Redshift.
- Strong programming skills in SQL and Python for data processing and automation.
- Experience with orchestration and workflow management tools like Airflow, ADF, and Databricks Workflows.
- Solid understanding of relational and cloud-native databases, ensuring data integrity and optimized query performance.
- Proficient in Git and familiar with CI/CD processes for cloud-based data solutions.
Preferred Qualifications and Benefits
- Experience integrating reporting and analytics across Microsoft Fabric, BigQuery, and Azure Synapse ecosystems.
- Ability to work effectively in fast-paced, multi-cloud environments with a collaborative and detail-oriented approach.
- Strong analytical skills with the ability to translate complex business needs into data-driven solutions.
- Excellent communication and stakeholder management capabilities.
This role offers the opportunity to work with cutting-edge cloud technologies and contribute to the development of scalable, modern data platforms that drive business insights and innovation.