TCS is looking for a Site Reliability Engineer
Work modality: Hybrid
(Candidate needs to be located or relocate to Querétaro, CDMX, Guadalajara or Monterrey, it will be requested to attend office at least 3 day per week)
Required Technical Skill Set (Desired Experience Range: min. 4 years):
- JAVA Application Trouble shooting - On Premise Dev Ops tools knowledge - Application Performance monitoring tools Knowledge - Server Load Balancing - Strong SQL and Oracle/Post GRE/Cassandra - Proficiency with Shell/Bash Scripting knowledge
Job summary:
Must-Have: JAVA Application Troubleshooting, Scripting (Python/Shell/Bash), Unix/Linux, Strong SQL writing experience, concepts with any one or more (Oracle/Postgre SQL/ My SQL/ Cassandra), Shell Scripting, Apache Tomcat Server, Splunk, Dynatrace, Grafana or Data Dog, Basics of Dev Ops pipeline creation and maintenance
Expectations from the Role:
- Responsible for maintaining Production environment and provide SRE Services - Strong knowledge of ITSM/ITIL Process, incident and problem management experience. - Triaging and troubleshooting incidents perform root cause analysis post incident remediation, resolving tickets and restoring services as per the SLA. - Collaborate and coordinate with internal and external stakeholders during incident or outages. - Keep all stakeholders updated during the incident lifecycle and escalate as required as per the predefined call structures. - Have a strong knowledge of Payment system and/or payment application from financial/banking domain. - Ready to work in rotational shift model and on-call support during the weekend
What we offer to you:
- 100% payroll under law - Direct contract - Benefits above law