31 ago
|
The Cervantes Group
|
Ciudad de México
31 ago
The Cervantes Group
Ciudad de México
The Data Engineer will focus on the design and build out of data models, codification of business rules, mapping of data sources to the data models (structured and unstructured), engineering of scalable ETL pipelines, development of data quality solutions, and continuous evaluation of technologies to continue to enhance the capabilities of the Data Engineer team and broader Innovation group.
Job Duties
- Build data pipelines and batch data pipeline with relational and columnar database engines as and evaluate their respective strengths and weaknesses.
- Build scalable and performant data models utilizing data structures, algorithms, programming languages, distributed systems, and information retrieval
- Lead data pipeline execution, testing, cutover, and data modernization solution while identifying opportunities for automating tasks.
- Work with large data sets and deriving insights from data using various BI and data analytics tools.
- Collaborate with Security teams and security requirements for handling data both in motion and at rest such as communication protocols, encryption, authentication, and authorization.
- Assemble large, complex data sets that meet non-functional and functional business requirements in addition to helping to establish Data Engineering architecture strategy, best practices, and standards for data lake support engineering.
- Design and build reliable, scalable data infrastructure with leading privacy and security techniques to safeguard data using AWS technologies.
- Assemble large,
complex data sets that meet non-functional and functional business requirements and prepare technical design specifications.
- Build Data Flows mapping Source systems and Process flows and work with stakeholders including data, design, product and executive teams and assisting them with data-related technical issues.
- Collaborate with cross-functional teams to understand end user needs to be translated into technical solutions while monitoring scheduled production related tasks.
Education & Requirements:
- Bachelor's Degree in Computer Science or related field
- Bilingual in English/Spanish
- 5+ years’ of overall industry experience with a minimum of 2 years’ experience working preferably as a data engineer
- Object-oriented/object function scripting languages: Python, R, C/C++, Java, Scala, etc.
- Relational SQL, distributed SQL and NoSQL databases including MS SQL, PostgreSQL, MySQL, etc. MemSQL, CrateDB, MongoDB, Cassandra, Neo4j, AllegroGraph, ArangoDB, etc.
- Big data tools(Hadoop, Spark, and/or Kafka)
- Data modeling tools (ERWin, Enterprise Architect, and/or Visio)
- Data integration tools (SSIS, Informatica, and/or SnapLogic)
- Data pipeline and workflow management (Azkaban, Luigi, and/or Airflow)
- Cloud technologies such as SaaS, IaaS and PaaS within Azure, AWS or Google Linux and comfortable with bash scripting Docker and Puppet
Plus
- Certifications in AWS are highly desired, such as AWS Certified Data Engineer Associate
📌 Data Engineer (Ciudad de México)
🏢 The Cervantes Group
📍 Ciudad de México