LEAD, Site Reliability Engineering (Ecatepec de Morelos)

LEAD, Site Reliability Engineering (Ecatepec de Morelos)

15 sep
|
Capital one
|
Ecatepec de Morelos

15 sep

Capital one

Ecatepec de Morelos

WeWork Reforma Latino (97001), Mexico, Ciudad de Mexico, Ciudad de Mexico Lead Site Reliability Engineer We're building a Site Reliability Engineering center in Mexico City, and we're hiring a Manager-level Backend Engineer to own the reliability and operational maturity of our settlement platforms. These are batch-critical systems that process every credit and debit transaction across the network. You'll work across hybrid infrastructure (on-prem data centers and AWS), partner closely with UK-based engineers, and build the automation and observability that allows Mexico City to operate settlement. Own reliability for batch settlement systems - ensure cycle completion windows are met, data integrity is maintained, and failures are detected before they reach downstream consumers Partner with UK-based settlement engineers - acquire domain expertise on Durbin compliance windows, cross-border DCI routing, and acquirer/issuer SLA adherence Participate in incident management - respond to settlement failures, drive root cause analysis, and implement durable fixes that prevent recurrence You'll work with batch processing systems that handle financial transactions across multiple on-prem data centers with active/active and active/passive configurations. The stack includes Java, Python, shell scripting, SQL, AWS, Kubernetes, OpenShift containers, Datadog, Observe, and legacy payment platforms.



CI/CD pipelines, API automation, and secret management via HashiCorp Vault are part of daily operations. You'll need strong troubleshooting and debugging skills and be comfortable with both modern cloud-native tooling and traditional enterprise batch systems. Professional English fluency ~ Bachelor's degree ~ At least 6 years of experience in SRE, production operations, or reliability engineering ~ Experience in DevOps Engineering (internship experience does not apply) ~Java, Python, Go ~ At least 4 years of experience with Cloud Native technologies (Amazon Web Services, Microsoft Azure, Google Cloud Platform) ~3+ years of experience with container orchestration services including Docker or Kubernetes ~ Experience with Shell or Bash scripting ~ At least 3 years of Unix or Linux system administration experience Knowledge or experience of Networking concepts (TCP/DNS/TLS) In the hiring process, we seek to provide equal employment opportunities to candidates, regardless of race, color, religion, gender, sexual orientation, marital or civil status, national origin, disability, or any other situation protected by federal, state, or local laws. For technical support or questions about Capital One's recruiting process, please send an email to Careers@capitalone.

📌 LEAD, Site Reliability Engineering (Ecatepec de Morelos)
🏢 Capital one
📍 Ecatepec de Morelos

Postulate a este anuncio

Muestra tus habilidades a la empresa, rellenar el formulario y deja un toque personal en la carta, ayudará el reclutador en la elección del candidato.

Suscribete a esta alerta:

Recibe por email las nuevas ofertas de trabajo para: lead, site reliability engineering (ecatepec de morelos) / ecatepec de morelos

Suscribete a esta alerta:

Recibe por email las nuevas ofertas de trabajo para: lead, site reliability engineering (ecatepec de morelos) / ecatepec de morelos