05 ago
|
Bluelightconsulting
|
Hermosillo
05 ago
Bluelightconsulting
Hermosillo
About the role
Bluelight is a leading software consultancy dedicated to designing and developing innovative technology that enhances users' lives. With a steadfast commitment to delivering exceptional service to our clients, Bluelight excels in its focus on quality and customer satisfaction. Our mission is not only to create cutting‑edge applications but also to foster a collaborative and enriching work environment where each team member can grow and thrive. With a presence across the United States, Central, and South America, Bluelight is in an exciting phase of expansion, continually seeking exceptional talent to join its dynamic and diverse community.
Responsibilities Develop and maintain ETL data engineering processes using Python (Py Spark) within Azure Synapse Analytics Notebooks and/or Azure Synapse Analytics Pipelines to ensure efficient data extraction, transformation, and loading. Apply data warehousing expertise, understanding star schemas, facts, and dimensions, to design and build effective data storage structures in a Massively Parallel Processing (MPP) SWL Pool. Extract data from various sources, including REST APIs, SWL database tables, and CSV files. Utilize deep knowledge of Azure Synapse Analytics to design and optimize data notebooks/pipelines for scalability and performance. Contribute to the implementation and understanding of other Data Fabric concepts, such as data lakes, lakehouses, delta lakes, and data cataloging, to enhance data management capabilities. Collaborate with data architects to create data models and schemas that align with business requirements. Implement data quality checks and validation processes to maintain data accuracy and consistency.
Identify and resolve performance bottlenecks and optimize ETL data notebooks/pipelines to meet SLAs. Monitor ETL jobs, diagnose issues, and implement solutions to ensure data pipeline reliability. Maintain comprehensive documentation of ETL data engineering processes, data flows, and data transformations. Work closely with cross‑functional teams to understand data requirements and provide support for data‑related initiatives. Ensure data security and compliance with data governance and privacy standards. Qualifications Bachelor’s degree in Computer Science, Information Technology, or a related field; or equivalent work experience, with certifications related to data engineering or data science (e.g., Azure Data Engineer) being a plus. Proven experience in ETL data engineering with significant expertise in using Python (Py Spark) to perform data extraction, transformation, and loading from REST APIs, SQL database tables, and CSV files. Proficiency in using Azure Synapse Analytics resources including Notebooks, Pipelines, Linked Services, and Azure Key Vault. Demonstrated ability to write complex SQL queries, optimize query performance, and work with both Spark SQL and MS SQL to effectively extract, transform, and load data. Knowledge of data integration best practices and tools. Experience with version control systems, such as Git (Azure Dev Ops). Strong problem‑solving and analytical skills, with a keen attention to detail. Excellent communication skills, both verbal and written, with the ability to work collaboratively in a team environment with shifting priorities. Familiarity with big data technologies, machine learning, and data analysis preferred. Experience with data visualization tools (e.g., Power BI, Tableau) and Agile Methodologies a plus. #J-18808-Ljbffr
📌 Azure data engineer - remote, latin america (Hermosillo)
🏢 Bluelightconsulting
📍 Hermosillo