31 ago
|
Capgemini Engineering
|
Aguascalientes
31 ago
Capgemini Engineering
Aguascalientes
MID -Data Engineer- | REMOTEAt Capgemini Engineering, the world leader in engineering services, we bring together a general team of engineers, scientists, and architects to help the world's most innovative companies unleash their potential.
From autonomous cars to life‐saving robots, our digital and software technology experts think outside the box as they provide unique R&D; and engineering services across all industries.
Join us for a career full of opportunities.
Where you can make a difference.
Where no two days are the same.YOUR ROLEDesign, develop, and support data replication and integration solutions using HVRDesign, build, and maintain scalable data pipelines using Databricks (Spark, Delta Lake)Develop and optimize ETL/ELT processes for structured and unstructured dataWork with large datasets to ensure data quality, integrity, and performance optimizationImplement data models and transformations for analytics and reportingCollaborate with data scientists and analysts to enable advanced analytics and ML workloadsIntegrate data from multiple sources including databases, APIs, and streaming systemsOptimize Spark jobs for performance tuning and cost efficiencyImplement data governance, security, and access controlsMonitor and troubleshoot data pipelines and production issuesSupport CI/CD pipelines and DevOps best practices for data engineering workflowsSupport data migration and modernization initiativesEnsure data quality, governance, security, and compliance standardsCreate operational documentation, runbooks, and support proceduresParticipate in production support, issue resolution, and performance tuning activitiesYOUR PROFILEBachelor's degree in computer science,
Engineering, or related fieldHands‐on experience with data engineering or big data developmentStrong knowledge of SQL / Oracle / SQL ServerStrong knowledge of Java / JavaScript (preferred)Strong knowledge of Web‐based applications and APIs (REST/SOAP)Strong experience with Databricks PlatformStrong experience with Apache Spark (PySpark/Scala)Strong experience with SQL & PythonExperience with Delta Lake and data lake architectureHands‐on experience with cloud platforms (Azure, AWS, or GCP)Familiarity with data orchestration tools (Azure Data Factory, Airflow, etc.)Knowledge of data warehousing concepts (Star schema, Snowflake schema)Experience with version control (Git) and CI/CD pipelinesStrong understanding of data pipeline optimization and performance tuningTech StackDatabricks (Spark, Delta Lake)Python / PySpark / SQLAzure (ADLS, Synapse, ADF) / AWS (S3, Redshift, Glue)Git, CI/CD toolsData visualization tools (Power BI, Tableau)WHAT YOU'LL LOVE ABOUT WORKING HERE?
At Capgemini Engineering, we encourage flexibility in how, when, and where people get their work done, allowing a better work‐life balance, and greater empowerment.
They partner with their managers to find an arrangement that works best for their role and their circumstances.At Capgemini Engineering, we're always looking ahead.
We're part of a team that creates opportunities to achieve valuable change.
Change that makes a difference.
New connections, new technologies, new ways to work.
It's so energizing.At Capgemini Engineering, we make it easy for you to deepen knowledge and learn new skills while you're still doing the day job.
#J-*****-Ljbffr
📌 Data Engineer (Aguascalientes)
🏢 Capgemini Engineering
📍 Aguascalientes