Responsibilities
- Build and support ETL workflows in Python (PySpark) using Azure Synapse Analytics Notebooks or Pipelines for reliable data processing.
- Design and implement data warehouse models using star schemas, fact tables, and dimension tables within a Massively Parallel Processing SQL Pool.
- Pull data from diverse sources such as REST APIs, SQL database tables, and CSV files.
- Leverage in-depth knowledge of Azure Synapse Analytics to build high-performance, scalable data pipelines and notebooks.
- Support the adoption of Data Fabric components including data lakes, lakehouses, delta lakes, and data cataloging to improve data handling.
- Partner with data architects to develop data models and schemas that meet business needs.
- Establish data validation rules and quality checks to ensure accuracy and consistency across systems.
- Detect and fix performance issues in ETL processes to meet service level agreements.
- Monitor execution of ETL jobs, troubleshoot failures, and apply fixes to maintain pipeline stability.
- Keep detailed records of ETL workflows, data movement, and transformation logic.
- Engage with cross-functional teams to understand data needs and support data-driven projects.
- Enforce data security practices and comply with governance and privacy regulations.
Work Arrangement
Remote (Worldwide) — Latin America
Other
- Serving as a consultant on this team offers an engaging, demanding, and fulfilling professional path.
- Your work is critically valued by clients and frequently drives meaningful business outcomes.
- You’ll engage in diverse projects for exceptional clients, accelerating your professional development.
- You’ll use cutting-edge tools and collaborate with top-tier industry experts.