Nuevo

IC3 - Infra Engineer - SRE

México, México

Objective of the Role
The IC3 SRE Engineer is responsible for supporting and enhancing the reliability, availability, and performance of the company's IT infrastructure and applications. This semi-senior role focuses on improving system stability and efficiency through advanced monitoring, automation, and incident response, contributing to the overall success of IT operations and strategic initiatives.

Main Responsibilities
1. Advanced System Monitoring: Implement and maintain advanced monitoring solutions to ensure the health and performance of infrastructure and applications.
2. Incident Response: Lead incident response activities, diagnosing and resolving system reliability issues, and conducting post-incident reviews.
3. Automation and Scripting: Develop and implement automation scripts and tools to improve system reliability and operational efficiency.
4. Performance Analysis: Collect, analyze, and interpret performance data to identify trends, anomalies, and potential issues, providing actionable insights.
5. Documentation: Maintain accurate and up-to-date documentation of system configurations, processes, and procedures.
6. Collaboration: Work closely with other IT team members and departments to support reliability engineering projects and initiatives.
7. Mentorship: Provide guidance and support to junior engineers, helping to enhance their technical skills and knowledge.
8. Security Compliance: Implement and enforce security measures to protect systems and ensure compliance with security policies.
9. Continuous Improvement: Drive continuous improvement initiatives, exploring new technologies and methodologies to enhance system reliability.
10. Autonomous Work Culture: Actively contribute to creating an autonomous work culture by taking initiative, being self-motivated, and collaborating effectively in an agile and lean environment.
11. Spin Culture Ambassador: Embody and promote Spin's values in every action, fostering a positive and inclusive work environment.
12. Disaster Recovery: Develop and maintain disaster recovery plans to ensure business continuity in case of system failures.

Required Knowledge and Experience
1. Bachelor's degree in computer science, Information Technology, or a related field, or equivalent work experience.
2. Minimum of 5+ years of experience in site reliability engineering or related fields.
3. Strong understanding of system reliability concepts, including monitoring, automation, and incident response.
4. Proficiency with scripting languages and automation tools.
5. Strong problem-solving and troubleshooting skills.
6. Excellent communication and teamwork skills.
7. Willingness to learn and adapt to new technologies and processes.
8. Data-driven mindset
9. Strong communication skills
10. English Level: Intermediate profiency.

En Spin estamos comprometidos con construir un lugar de trabajo diverso e inclusivo.

Creemos en la igualdad de oportunidades y promovemos un entorno libre de discriminación por motivos de raza, origen nacional, género, identidad de género, orientación sexual, discapacidad, edad o cualquier otra condición legalmente protegida.

Crear una alerta de empleo

¿Le interesa desarrollar su carrera en Spin - Job Board External ? Reciba futuras oportunidades directamente en su correo electrónico.

Solicitar este puesto

*

indica un campo obligatorio

Teléfono
Currículum*

Tipos de archivos aceptados: pdf, doc, docx, txt, rtf


Please share your LinkedIn profile link