Site Reliability Engineer (Ciudad de México)

Site Reliability Engineer (Ciudad de México)

09 oct
|
Cognizant
|
Ciudad de México

09 oct

Cognizant

Ciudad de México

We’re hiring

At Cognizant we have an adecuado opportunity for you to be part of one of the largest companies in the digital sector worldwide. A Great Place To Work where we look for people who contribute new ideas, experiencing a dynamic and growing environment. At Cognizant we promote an inclusive culture, where we value different perspectives providing career growth and development opportunities. #WelcomeToCognizant

Role Overview

The Application Reliability Engineer will lead a structured reliability‑improvement initiative across critical applications. The consultant will analyze recurring production issues, identify systemic weaknesses, and drive measurable improvements in stability, resiliency, performance, observability, and operational effectiveness. This role requires a strong application‑centric mindset, deep diagnostic skills, and the ability to convert operational noise into actionable reliability outcomes

Key Responsibilities

Reliability & Production Stability

• Lead reliability uplift initiatives across assigned applications and services.




• Assess current reliability posture: recurring incidents, degradation patterns, operational pain points, manual toil, and risk areas.
• Define and prioritize improvements that reduce incidents, customer impact, alert noise, and MTTR. Build a practical roadmap for resilience, availability, recoverability, and operational readiness. Evaluate operational procedures and identify gaps affecting responsiveness and production support. Partner with engineering, operations, platform, infrastructure, and product teams to execute the reliability plan.

Incident Management, Analysis & Root Cause Resolution

• Lead production incident reviews for high‑impact and complex issues.
• Perform deep‑dive analysis to identify root causes, not just symptoms.
• Review logs, metrics, traces, dependency health, and post‑incident findings to identify systemic failure patterns.
• Differentiate actionable incidents from duplicates or noisy ale

📌 Site Reliability Engineer (Ciudad de México)
🏢 Cognizant
📍 Ciudad de México

Postulate a este anuncio

Muestra tus habilidades a la empresa, rellenar el formulario y deja un toque personal en la carta, ayudará el reclutador en la elección del candidato.

Suscribete a esta alerta:

Recibe por email las nuevas ofertas de trabajo para: site reliability engineer (ciudad de méxico) / ciudad de méxico

Suscribete a esta alerta:

Recibe por email las nuevas ofertas de trabajo para: site reliability engineer (ciudad de méxico) / ciudad de méxico