This job was published 3 months ago

Ingeniero/a Sr. de Infraestructuras y Observabilidad CLOUD

Mapfre
Mapfre
Madrid, SpainOn-siteCompetitiveAdded 3 months agoInternship
🇪🇸 Translated from Spanish

This job was originally posted in Spanish and automatically translated to English. You'll most likely need Spanish to apply.

Job Description

LOCATION.- MAJADAHONDA - CR. POZUELO DE ALARCON, 50-1

We are looking for an experienced Infrastructure and Operations Engineer (I&O IBERIA) to join our cross-platform team, responsible for leading the evolution of the operating model, driving and evolving observability, ensuring the cloud operation of applications, platforms and products, implementing Service Delivery Manager best practices and guaranteeing "Production Readiness".

You will join a dynamic and highly goal-oriented team, which must guarantee the availability, performance, service levels and compliance of MAPFRE Spain's Cloud applications.

You must have "everything as code" in your DNA: from Infrastructure as Code (IaC) to self-remediation, Pipelines, CD, etc., with strong knowledge of Cloud related to computing, serverless, storage and networking, all geared towards a Platform Engineering-based model.

HIERARCHICAL AND/OR FUNCTIONAL DEPENDENCE

Infrastructure and Operations Department, Iberia

JOB FUNCTIONS (KEY FUNCTIONS)

Service Delivery Manager:

• Responsible for ensuring that technology services are delivered effectively, aligned with business objectives and meeting quality standards, ensuring the Operation model. To do this, you will have to supervise the execution of the contracted services, ensuring compliance with the service level agreements (SLA).

• Act as a liaison between the technical teams for products, platforms, applications, SRE, Architecture and Security.

• Organize and lead teams to ensure the timely and efficient delivery of services.

• Identify opportunities for optimization in delivery processes, tools, and methodologies.

• Monitor the performance of external providers and ensure that outsourced services meet established standards.

• Analyze key performance indicators and prepare reports for management.

Strategic management:

• You will be involved in the planning, design and execution of the Observability 360 roadmap at MAPFRE Spain, ensuring its alignment with the strategic objectives of Management, guaranteeing key decision-making based on data and operational needs of availability, performance and safety.

• You will design and evolve the Production Readiness of applications, with an Agile / Lean perspective, to facilitate the development and maintenance teams, SRE and Architecture, ensuring that applications reach Production with the necessary guarantees of Availability, Performance, Security and operational resilience.

• You will drive automation and the "everything as code" goal as part of the new application operating model.

Technical supervision:

• You will ensure the implementation and operation of observability tools such as Dynatrace, Grafana, standards such as Open Telemetry, alerting, auto-scaling, auto-remediation, etc... to guarantee compliance with the operating model and Production Readiness.

• You will need to collaborate with Architecture, the SRE team, Release Management and Product teams, so that the design and subsequent implementation of the applications meet the requirements and objectives of Management, adapting and aligning Production Readiness according to the needs that are identified.

• You will drive the use of reference architectures and automation with the use of Terraform, Ansible, and CAP in Infrastructure pipelines and operations.

Team management:

• You will lead and coordinate multidisciplinary teams that will help you achieve the objectives of the I&O Iberia Management and Business.

• You will act as a link between different product teams, platforms and suppliers, and you will have to establish mechanisms that optimize this collaboration in order to achieve common goals, promoting data-driven decision-making and prioritization aligned with the Organization's strategy.

Supervision and Reporting:

• You will ensure that the area's Management has the necessary reports, correctly updated, that must be presented and defended in the different Monitoring, Management and Strategic Committees: SLA, SLO, APPDEX, COMPLIANCE, etc...

• You will be responsible for ensuring the existence of dashboards that represent the status of the activities and responsibilities of the I&O Iberia Management, adapting, evolving or creating those that are necessary.

Communication:

• You will disseminate and communicate the pillars of the MAPFRE Spain Cross-Platform Service operating model based on CloudOps (NoOps), production readiness, end-to-end observability and extreme automation with the "everything as code" mentality

• You will be responsible for clearly and promptly communicating any changes, improvements, or benefits introduced in existing or new processes that impact the Governance model and the relationship with the Product, Platforms, Architecture, and SRE teams.

REQUIRED TRAINING AND KNOWLEDGE

Advanced level, with demonstrable experience working with AWS and Azure, with in-depth knowledge of HA and DR models, focused on the resilience of Cloud Applications

Advanced level, with demonstrable experience working with Kafka, MongoDB, Kubernetes (openShift, AKS, EKS, ...)

Expert level in ITIL processes

Solid knowledge in automation, preferably focused on infrastructures (Terraform, Ansible) but experience working with SDKs or APIs of Hyper-scalars and OpenShift is highly valued.

Advanced level in the administration, implementation, and operation of Dynatrace and the LGTM stack (Loki, Grafana, Tempo, and Mimir). Proficient in the OpenTelemetry standard.

Access and Policy Management, Dashboard Design and Alert Configuration, Performance Monitoring and Optimization, Log Analysis...)

PROFESSIONAL EXPERIENCE REQUIRED FOR THE POSITION

5 years in positions related to ITIL processes, of which 3 were as Service Delivery Manager

5 years in coordinating multidisciplinary teams and managing projects with multiple dependencies in cloud environments

3 years in the design of application architectures, platforms and Cloud services, etc., preferably in Serverless or PaaS models and their automated deployments and operation

3 years of experience in the implementation, operation, and use of observability tools such as Dynatrace and the LGTM stack (Loki, Grafana, Tempo, and Mimir). Proficient in the OpenTelemetry standard.

KEY SKILLS

• Strategic leadership and ability to influence across departments

• Technological vision and results orientation.

• Excellent communication and collaboration skills.

• Analytical and decision-making skills in complex environments.

SUCCESS INDICATORS

• Technological evolution of the platform aligned with strategic objectives.

• Level of adoption and reuse of the platform by business and technology teams.

• Improvement in operational efficiency and technical quality of deployed solutions.

Skills and soft skills

• Ability to work in multidisciplinary teams, in a dynamic environment.

• Communication skills.

• Analytical, pragmatic thinking and problem-solving orientation.

• Proactivity, autonomy and a desire to teach, learn and think as a team.

• Good vibes, good humor, and a desire to have fun

View original advert (Spanish)

UBICACIÓN.- MAJADAHONDA - CR. POZUELO DE ALARCON, 50-1

Buscamos a un/a Ingeniero/a de Infraestructuras y Operaciones IBERIA (I&O IBERIA) experto/a, para unirse a nuestro equipo de plataforma trasversal, responsable de liderar la evolución del modelo de operación, impulsar y evolucionar la observabilidad, asegurar la operación cloud de las aplicaciones, plataformas y productos, las mejores prácticas de Service Delivery Manager y garantizar el "Production Readiness".

Te integrarás en un equipo dinámico y altamente orientado a resolución y objetivos, que debe garantizar la disponibilidad, rendimiento, niveles de servicio y compliance de las aplicaciones Cloud de MAPFRE España.

Deberás tener en tu ADN el "everything as code": desde la Infraestructura como Código (IaC) hasta las auto-remediaciones, Pipelines, CD, etc. con fuertes conocimientos en Cloud relativos a cómputo, serverless, storage y networking, orientado todo a un modelo basado en Platform Engineering.

DEPENDENCIA JERARQUICA Y/0 FUNCIONAL

Dirección Infraestructura y Operación iberia

FUNCIONES DEL PUESTO (FUNCIONES CLAVE)

Service Delivery Manager:

• Responsable de asegurar que los servicios tecnológicos se entreguen de forma eficaz, alineados con los objetivos del negocio y cumpliendo los estándares de calidad asegurando el modelo de Operación para ello tendrás que supervisar la ejecución de los servicios contratados, asegurando el cumplimiento de los acuerdos de nivel de servicio (SLA).

• Actuar como enlace entre los equipos técnicos de productos, plataformas, aplicaciones, SRE, Arquitectura y Seguridad.

• Organizar y liderar equipos para garantizar la entrega puntual y eficiente de los servicios.

• Identificar oportunidades de optimización en procesos, herramientas y metodologías de entrega.

• Supervisar el rendimiento de proveedores externos y asegurar que los servicios subcontratados cumplen con los estándares establecidos.

• Analizar indicadores clave de rendimiento y elaborar informes para la dirección.

Gestión estratégica:

• Tendrás que participar en la planificación, diseño y ejecución del roadmap de Observabilidad 360 en MAPFRE España, asegurando su alineamiento con los objetivos estratégicos de la Dirección, garantizando la toma de decisiones clave basándote en datos y necesidades operativas de disponibilidad, rendimiento y seguridad.

• Diseñarás y evolucionarás el Production Readiness de las aplicaciones, con una perspectiva Agile / Lean, para facilitar a los equipos de desarrollo y mantenimiento, SRE y Arquitectura, que las aplicaciones llegan a Producción con las garantías necesarias de Disponibilidad, Rendimiento, Seguridad y resiliencia operativa.

• Impulsarás la automatización y el objetivo "everything as code" como parte del nuevo modelo de operación de aplicaciones.

Supervisión técnica:

• Asegurarás la implantación y explotación de herramientas de observabilidad como Dynatrace, Grafana, estándares como Open Telemetry, alertado, auto-escalados, auto-remediaciones, etc... para garantizar el cumplimiento del modelo operación y Production Readiness.

• Deberás colaborar con Arquitectura, el equipo de SRE, Release Management y equipos de Producto, para que el diseño y posterior implantación de las aplicaciones cumplan con las exigencias y objetivos de la Dirección, adaptando y alineando el Production Readiness según las necesidades que se vayan identificando.

• Impulsarás el uso de las arquitecturas de referencia y la automatización con el uso de Terraform, Ansible y CAP en los pipelines y operaciones de Infraestructura.

Gestión de equipos:

• Liderarás y coordinarás equipos multidisciplinares que te ayudarán a cumplir los objetivos de la Dirección de I&O Iberia y de Negocio.

• Actuarás como punto de unión entre diferentes equipos de productos, plataformas y proveedores, y tendrás que establecer mecanismos que optimicen esta colaboración en pro de la consecución de los objetivos comunes, fomentando la toma de decisiones y la priorización basada en datos y alineadas a la estrategia de la Organización.

Supervisión y Reporting:

• Garantizarás que la Dirección del área cuenta con los informes necesarios, correctamente actualizados, que se deban presentar y defender en los diferentes Comités de Seguimiento, Dirección y Estratégicos: SLA, SLO, APPDEX, COMPLIANCE, etc...

• Serás responsable de garantizar la existencia de los dashboards que representen el estado de las actividades y responsabilidades de la Dirección de I&O Iberia, adaptando, evolucionando o creando aquellos que sean necesarios.

Comunicación:

• Divulgarás y comunicarás los pilares del modelo de operación del Servicio de Plataformas Transversales de MAPFRE España basado en CloudOps (NoOps), production readiness, observabilidad end to end y automatización extrema con la mentalidad "everything as code

• Serás responsable de comunicar de manera clara y oportuna cualquier cambio, mejora o beneficio introducido en los procesos existentes o nuevos que impacten en el modelo de Gobierno y en el de relación con los equipos de Producto, Plataformas, Arquitectura y SREs.

FORMACION Y CONOCIMIENTOS REQUERIDOS

Nivel avanzado, con experiencia demostrable, trabajando con AWS y AZURE, con profundo conocimiento de modelos de HA y DR, orientados a la resiliencia de Aplicaciones Cloud

Nivel avanzado, con experiencia demostrable, trabajando con Kafka, MongoDB, Kubernetes (openShift, AKS, EKS, ...)

Nivel experto en procesos ITIL

Conocimientos sólidos en automatización, preferentemente orientado a infraestructuras (Terraform, Ansible) pero muy valorable el trabajo con SDKs o APIs de Hiper-escalares y OpenShift.

Nivel avanzado en administración, implementación y operación de Dynatrace y el stack LGTM (Loki, Grafana, Tempo y Mimir). Manejo del estándar OpenTelemetry:

Administración de Acceso y Políticas, Diseño de Dashboards y Configuración de Alertas, Monitorización y Optimización de Rendimiento, explotación de logs...)

EXPERIENCIA PROFESIONAL REQUERIDA PARA EL PUESTO

5 años en puestos relacionados con procesos ITIL, de los cuales 3 como Service Delivery Manager

5 años en coordinación de equipos multidisciplinares y gestión de proyectos con múltiples dependencias en entornos cloud

3 años en diseño de arquitecturas de aplicaciones, plataformas y servicios Cloud, etc. preferentemente en modelos Serverless o PaaS y sus despliegues y operación automatizado

3 años en Implantación, explotación y uso de herramientas de observabilidad como Dynatrace y stack LGTM (Loki, Grafana, Tempo y Mimir). Manejo del estándar OpenTelemetry

HABILIDADES CLAVE

• Liderazgo estratégico y capacidad de influencia transversal

• Visión tecnológica y orientación a resultados.

• Excelentes habilidades de comunicación y colaboración.

• Capacidad analítica y de toma de decisiones en entornos complejos.

INDICADORES DE ÉXITO

• Evolución tecnológica de la plataforma alineada con los objetivos estratégicos.

• Nivel de adopción y reutilización de la plataforma por parte de los equipos de negocio y tecnología.

• Mejora en la eficiencia operativa y calidad técnica de las soluciones desplegadas.

Habilidades y soft-skills

• Capacidad de trabajo en equipos multidisciplinares, en un entorno dinámico.

• Habilidades de comunicación.

• Pensamiento analítico, pragmático y orientación a resolución de problemas.

• Proactividad, autonomía y ganas de enseñar, aprender y pensar en equipo.

• Buen rollo, buen humor, y ganas de divertirse

Need a visa? No sponsorship mentioned here. Browse visa jobs