Regístrate para acceder a todas las funciones de nuestro servicio
  • Búsqueda de ofertas de trabajo
  • Favoritas
  • Crear CV
    Nuevo
  • Salario
  • Alertas de empleo

Site Reliability Engineer

Completa

Plenit

About this role

The challenge
As a Site Reliability Engineer , you will play a critical role in ensuring the stability, availability, resilience, and performance of our cloud platform in production.
Your mission will cover two complementary areas. First, you will work proactively to prevent incidents by identifying operational risks, anticipating capacity constraints, improving production changes, resolving recurring problems, and driving reliability improvements, migrations, and corrective initiatives. Second, when incidents occur, you will help restore service quickly by coordinating technical diagnosis, controlling the situation, communicating impact, and leading recovery efforts.
This is a senior, hands-on role for someone with a strong operational mindset who is comfortable working in complex production environments, investigating technical issues, making decisions under pressure, and taking ownership of platform reliability.
You will be part of the Site Reliability Engineering team, working in a practical environment focused on keeping the platform stable while continuously reducing operational risk and manual effort.
A key part of the role will be to identify what could affect the platform before it becomes an incident. You will assess risks, make them visible, define mitigation or correction plans, assign ownership, and follow actions through to completion. You will also contribute to capacity planning, production scaling, infrastructure improvements, migrations, observability, automation, and operational readiness.
When incidents happen, you will act as a technical reference and coordination point. You will help establish control, assess impact, accelerate diagnosis, involve the right teams, communicate clearly, and restore service as quickly and safely as possible.
You will work closely with engineering, systems, networking, storage, database, security, and product teams. Your operational perspective will help ensure that reliability, scalability, recoverability, and operability are considered in platform changes and evolution.
Participation in an on-call rotation and response to critical incidents outside regular working hours are part of the role.

Requirements that are important for us
We are looking for a senior Site Reliability Engineer with strong experience operating critical production environments and a proven ability to identify risks, troubleshoot complex issues, and drive long-term reliability improvements.
Relevant experience and expected outcomes
  • Operating critical production infrastructures while ensuring availability, stability, performance, and recoverability.
  • Identifying technical and operational risks and driving mitigation, correction, or contingency plans.
  • Maintaining a prioritized backlog of risks, recurring problems, capacity constraints, and reliability improvements.
  • Leading incident response processes, including impact assessment, technical diagnosis, coordination, communication, and service restoration.
  • Performing root cause analysis and implementing corrective and preventive actions.
  • Managing recurring problems and ensuring they remain visible, owned, prioritized, and followed through to resolution.
  • Monitoring demand, system limits, growth trends, bottlenecks, and saturation points.
  • Contributing to capacity planning and executing infrastructure scaling in production environments.
  • Reviewing infrastructure changes, deployments, maintenance activities, and migrations from a reliability perspective.
  • Defining implementation, validation, rollback, and contingency plans for production changes.
  • Improving observability through metrics, logs, traces, alerts, dashboards, and service-health indicators.
  • Improving runbooks, diagnostic playbooks, recovery procedures, and operational documentation.
  • Identifying repetitive tasks, unnecessary escalations, and manual processes that should be automated or simplified.
  • Working with systems that support web applications, including a solid understanding of DNS, TCP/IP, and load balancing.
  • Administering technologies such as NGINX, Apache, load balancers, databases, and related web infrastructure.
  • Strong experience with Linux systems and working knowledge of Windows environments.
  • Good understanding of networking, virtualization, storage, databases, and infrastructure dependencies.
  • Familiarity with Kubernetes and containerized platforms is highly valuable.
  • Experience with cloud infrastructure is highly valuable.
Key skills and expected impact
  • Strong troubleshooting capabilities and the ability to analyze metrics, logs, traces, events, and system behaviour.
  • Broad technical knowledge across infrastructure, applications, networks, storage, and databases.
  • Strong risk-based thinking, with the ability to distinguish urgent work from important preventive work.
  • Ability to turn operational concerns into concrete actions, owners, deadlines, and measurable outcomes.
  • Ability to remain calm under pressure while acting decisively and proactively.
  • Strong ownership of platform stability and service recovery.
  • Strong coordination skills during incidents, migrations, changes, and technical escalations.
  • Clear communication with technical teams, stakeholders, and affected parties.
  • Ability to challenge unsafe changes or insufficient operational preparation constructively.
  • Strong documentation habits and commitment to shared operational knowledge.
  • Continuous-improvement mindset focused on reducing incidents, recovery time, operational effort, and manual intervention.
  • Motivation to act as a key technical reference during complex and high-impact scenarios.
Tools
  • Observability tools for metrics, logs, traces, alerting, and dashboards.
  • Infrastructure and system-administration tools across Linux and Windows environments.
  • Networking, virtualization, storage, and database technologies.
  • Web infrastructure tools such as load balancers, NGINX, Apache, and proxies.
  • Cloud infrastructure and configuration-management platforms.
  • Kubernetes and container orchestration technologies.
  • Incident-management and on-call tools.
  • Operational documentation tools, runbooks, playbooks, and incident procedures.
  • Capacity-planning, performance-analysis, and infrastructure-scaling tools.
  • Automation and scripting tools.
What success looks like
Success in this role means that operational risks are identified early, capacity issues are anticipated, production changes are safer, recurring problems are permanently addressed, and the platform becomes progressively more observable, resilient, scalable, and easier to operate.
When incidents occur, they are controlled quickly, communicated clearly, and resolved with reduced recovery time. The objective is not only to respond better, but to reduce the number, frequency, and impact of incidents over time.

About us

Plenit (formerly Jotelulu) is the Operating Platform for IT Service Providers. We help MSPs, ISVs, VARs, and IT resellers run, scale, and grow their businesses on top of our own sovereign infrastructure, with footprint across Europe and the Americas.

We are a European cloud platform with a clear conviction: digital sovereignty matters, and the IT channel deserves a partner that speaks their language, builds for their reality, and stays close to the ground. We are not enterprise-y. We are simple, pragmatic, and obsessed with our partners’ success.

Five years in, we are 130 people, four countries, and growing fast — which is exactly why we are bringing this role on now.



 
We’re passionate about technology, but we’re also a bit geeky—we love Star Wars, video games, 90s movies, and pop culture. We work hard and with passion, but we also have a great time together. At Plenit, you’ll find a balance between work and fun, where teammates often become friends, and boredom is rare—there’s always something new, unexpected, and memorable happening.

What We Offer

At Plenit, we aim to create an environment where people can grow, enjoy what they do, and bring their best selves to work. That’s why we’ve designed a benefits package that reflects who we are, how we work, and what we value as a company.

Our Culture

Our culture is how we work, how we relate to one another, and how we build something meaningful together.
It’s not about “cool” values for a website — it’s about real behaviors we live every day: a positive atmosphere, accountability, trust, and effectiveness.

We are committed to:

A great working environment built on teamwork, optimism, joy and the drive to do things well.
Responsibility and ownership, because everyone knows what needs to be done and finds the way to make it happen.
Trust, as the foundation of teamwork and our relationships with partners.
Effectiveness, focusing on results and constantly challenging ourselves to ensure we’re on the right path.

Special Days
You’ll enjoy:
December 24th and 31st off,
Your birthday off — because we believe your day should be yours.




Competitive Salary
We offer a competitive salary aligned with the market and reviewed periodically to ensure fairness and recognition.




Continuous Learning
We support your professional growth through internal training, learning resources, and development plans tailored to your career path.




Private Health Insurance
You’ll have access to private health coverage, giving you peace of mind and support in your day-to-day life.




Informal Events
We believe in celebrating achievements, strengthening team bonds, and having fun along the way. We regularly organize informal events and activities to enjoy time together and reinforce our team spirit.
Oferta de empleo publicada 21 días atrás
Ofertas similares que pueden interesarteSegún la Site Reliability Engineer en Madrid
  •  ...This position is at Indra The selection process will be fully managed by Indra. -- # Site Reliability Engineer SRE and DevOps Specialist **Ubicación:** Madrid, MD, ES **Perfil profesional:** ATM **Experiencia requerida:** **Modalidad del puesto:** Híbrido ## **Site... 
    Ofertas de empleo recomendadas
    Práctica
    Tiempo completo
    Empleo permanente
    Trabajo híbrido

    Indra

    Madrid
    Hace un mes
  •  ...Site Reliability Engineer (SRE) You’ll own the reliability, scalability, and operational excellence of the systems that power our platform. You’ll partner closely with Engineering, Security, and Product to build resilient infrastructure, improve developer experience,... 
    Ofertas de empleo recomendadas
    Remoto

    Randstad (Schweiz) AG

    Madrid
    3 días atrás
  •  ...Exoscale, a leading Swiss/European cloud provider, seeks a Site Reliability Engineer to design and maintain its base systems and hypervisor infrastructure across Linux, KVM, and networking components. Join a distributed European team, contribute to automation, security... 
    Ofertas de empleo recomendadas
    Remoto

    cloudControl

    Madrid
    4 días atrás
  •  ...transportation and urban mobility. We provide engineering that keeps the world moving,...  ...are looking for an experienced Senior Site Reliability Engineer (d/f/m) who will take ownership of...  ...engineering teams to build resilient, reliable solutions. Drive Monitoring & Observability... 
    Ofertas de empleo recomendadas
    Remoto
    Horario flexible

    thyssenkrupp Elevator

    Madrid
    1 día atrás
  •  ...Technical Director Job Summary Join us at Electronic Arts, where you will help shape the future of global gaming. As a Site Reliability Engineer for our Worldwide Localization Team, you will guide the reliability, performance, and scalability of the infrastructure... 
    Ofertas de empleo recomendadas
    Trabajar en la oficina
    Trabajo híbrido
    3 días a la semana

    Electronic Arts

    Madrid
    3 días atrás
  •  ...Site Reliability Engineer - Compute System & Network (f/m/d) Full time | Exoscale | Spain Posted On 07/31/2026 Job Information Number of Positions 1 Assigned Recruiter(s) Pierre-Matthieu Alamy,Mathias Fiedler Hiring Manager Pierre-Matthieu Alamy Technology... 
    Tiempo completo
    Trabajar en la oficina
    Desde casa
    Remoto
    Horario flexible

    cloudControl

    Madrid
    4 días atrás
  •  ...protects the environment. This is a place where you can be proud to work and do something that matters. What Will You Do? As a Data Engineer Intern at Procter & Gamble, you will build and optimize ETL pipelines while implementing data quality and validation processes.... 
    Práctica

    Experimentation Jobs

    Madrid
    4 días atrás
  •  ...About this role The challenge As a Site Reliability Engineer , you will play a critical role in ensuring the stability, availability, resilience, and performance of our cloud platform in production. Your mission will cover two complementary areas. First, you will... 
    Tiempo completo
    Inicio inmediato

    Plenit

    Madrid
    21 días atrás
  •  ...millions worldwide. Explore exciting career opportunities at ThetaRay - where innovation meets purpose. We are looking for a Site Reliability Engineer (SRE) to join our growing Global Support organization. The person is responsible for resolving technical issues and taking... 

    Thetaray

    Madrid
    4 días atrás
  • de 60000 a 100000 €/año

     ...Location: Madrid, M, ES We’re looking for a  Site Reliability Engineer (SRE)to join our Logging & Monitoring squad. You’ll ensure the  reliability, scalability, and security of our observability platforms, while building automated solutions that deliver seamless experiences... 
    Tiempo completo
    Contrato
    Trabajo híbrido

    Swiss Re

    Madrid
    7 días atrás
  •  ...welcomes diverse applicants. As part of its ongoing efforts to grow its infrastructure footprint Exoscale is hiring a Site Reliability Engineer. The site reliability engineer plays a critical role in ensuring constant availability of the Exoscale platform. The... 
    Tiempo completo
    Desde casa
    Remoto
    Horario flexible

    Exoscale

    Madrid
    12 días atrás
  •  ...Site Reliability Engineer (m/f/d) Key responsibilities As a Site Reliability Engineer within Advanced Analytics (DA3) in the Chief Data & AI Office at Allianz Partners, you will join the platform engineering team to own the reliability and operational health of... 
    Tiempo completo
    Empleo permanente
    Contrato
    Trabajar en la oficina

    AP Solutions GmbH

    Madrid
    Hace 2 meses
  •  ...… Fancy joining us? Our Product & Engineering teams are both based in Madrid, with a...  ...over the taxi app service industry! Site Reliability Engineers at Cabify work on improving...  ...about us! As a Site Reliability Engineer, you will be: Evolving our... 
    Trabajar en la oficina
    Remoto
    Trabajo híbrido
    Horario flexible
    1 día a la semana
    3 días a la semana

    Cabify

    Madrid
    Hace 2 meses
  •  ...still got a long way to go… Fancy joining us? Our Product & Engineering teams are both based in Madrid, with a strong remote culture...  ...have big plans to take over the taxi app service industry! Site Reliability Engineers at Cabify work on improving all aspects of our platform... 
    Tiempo completo
    Remoto
    1 día a la semana

    Cabify

    Madrid
    Hace un mes
  •  ...fundamentales para garantizar la calidad y disponibilidad de los servicios digitales. ¿Cuál será tu misión? Como Site Reliability Engineer (SRE), tu misión será mejorar la fiabilidad, escalabilidad y eficiencia operativa de las plataformas cloud de producción mediante... 
    Aprendiz
    Práctica
    Trabajo híbrido
    Turno de mañana

    Telefónica S.A.

    Madrid
    6 días atrás
  •  ...HITACHI EUROPE S.A. (SPAIN) **Profession (Job Category):** Engineering & Science **Job Schedule:** Full time **Remote:** No **Job...  ...Description:** **Summary** Hitachi Europe S.A. is searching for a Site Service Engineer for its Proton Therapy project in Madrid. The Site Service... 
    Tiempo completo
    Trabajar en la oficina
    Remoto
    Trabajo por turnos

    Hitachi

    Madrid
    Hace un mes
  • QUARK, parte de Sener, busca un/a Ingeniero/a Eléctrico/a de Obra para liderar la ejecución eléctrica en Data Centers. Trabajarás directamente en obra, coordinando con diseño, contratistas y cliente, asegurando calidad, seguridad y cumplimiento normativo. Se requiere...
    Contratista
    Aprendiz

    Sener en Ingeniería

    Madrid
    4 días atrás
  • Una empresa de ingeniería busca un profesional en topografía para unirse a su equipo técnico en Madrid. El candidato será responsable de garantizar la precisión geométrica y supervisar los parámetros geotécnicos en proyectos relevantes. Se requiere titulación en Ingeniería...
    Indefinido

    ConfiARTE

    Madrid
    1 día atrás
  • Se necesitan operarios para trabajos en altura en el sector de la construcción, con experiencia previa en este tipo de labores. Los puestos se desarrollarán en varias obras ubicadas en Madrid. Las funciones principales incluyen rehabilitación de fachadas: pintura, restauración...
    Indefinido
    Tiempo completo
    Empleo permanente
    Inicio inmediato

    Doméstiko.com

    Madrid
    Hace un mes
  • La firma busca un Arquitecto/a Técnico o Ingeniero/a con sólida experiencia en el diseño, cálculo y certificación de estructuras de andamiaje para proyectos de Obra Pública. Debe dominar AutoCAD y demostrar capacidad analítica y de resolución de problemas in situ, con ...

    Colegio Oficial de Aparejadores, Arquitectos Técnicos e Inge...

    Leganés, Comunidad de Madrid
    3 días atrás
  •  ...Sr Mechanical Engineer, TIPM, Global Building Design & Engineering Job ID: 10483105 | Amazon EU SARL (Spain Branch) Amazon is currently looking to hire an experienced Mechanical Engineer to join our team as Senior Technical Infrastructure Program Manager (TIPM),... 
    Contratista
    Contrato

    Amazon

    Madrid
    2 días atrás
  •  ...At Mott MacDonald, we believe engineering is more than just infrastructure—it's about shaping a future that’s resilient, inclusive, and sustainable...  ...and optioneering exercises toidentifyoptimalsolutions Site visits to assess conditions and support construction phases... 
    Práctica
    Temporal
    Empleo permanente
    Trabajar en la oficina
    Horario flexible

    Mott MacDonald

    Madrid
    3 días atrás
  • Responsibilities Diseño y cálculo de estructuras metálicas y de hormigón armado para plantas industriales. Desarrollo de ingeniería básica y de detalle en proyectos industriales y/o de generación de energía. Diseño de plataformas, estructuras de proceso, pipe-racks...

    Jobtailor

    Madrid
    1 día atrás
  • ¿Eres una persona junior con vocación técnica y quieres dejar tu huella en grandes proyectos de construcción? ¿Te ilusiona formar parte de una empresa que trabaja en proyectos emblemáticos como estadios, grandes edificios o centros comerciales? ¿Buscas aprender de profesionales...
    Aprendiz
    Trabajar en la oficina
    Lunes a jueves
    Jornada intensiva

    OCA Global

    Madrid
    Hace un mes
  •  ...looking for a proactive and technically skilled Backbone Build Engineer . Join our team and work with Arelion , one of the world'...  ...-1 Internet providers, operating one of the largest and most reliable global IP backbone networks. With more than 30 years of experience... 
    Tiempo completo
    Horario flexible

    Trinetix

    Madrid
    15 días atrás
  • This position is at FCC The selection process will be fully managed by FCC. -- # INGENIERO/A OBRAS MARÍTIMAS 26/0099 **Localidad**: Madrid **Provincia**: Madrid **País**: España **Empresa**: FCC INFRAESTRUCTURAS **Nº Vacantes (puestos)**: 1 ### Funciones FCC Construcción...
    Tiempo completo

    FCC

    Madrid
    Hace un mes
  • Descrizione del lavoro En Veolia Servicios Lecam precisamos incorporar en Madrid un/a Ingeniero/a de Oficina Técnica para las obras de mantenimiento de instalaciones. Sus funciones principales serán: Realizar ofertas técnicas y económicas de instalaciones mecánicas...
    Indefinido
    Práctica
    Empleo permanente
    Trabajar en la oficina
    Inicio inmediato

    Veolia

    Madrid
    5 días atrás
  •  ...The Manager of Structural & Geotechnical Engineering will be based in Madrid, Spain. In this...  ...placement and performance based on specific site conditions. Lead value engineering...  ..., higher performance, and greater reliability, helping our customers capture the full... 
    Horario flexible

    Nextracker

    Madrid
    2 días atrás
  • STANDBY busca Ingeniero / Responsable de Obra en Instalaciones Eléctricas, Clima y Contraincendios para liderar proyectos en urbanizaciones, edificios y locales en Madrid. El/la candidato/a ideal tendrá experiencia en cálculo y diseño de instalaciones eléctricas, climatización...
    Indefinido
    Inicio inmediato

    STANDBY.es | Selección de Personal

    Madrid
    2 días atrás
  •  ...Jobtailor seeks an experienced Structural Engineer to design and calculate metallic and reinforced concrete structures for industrial plants, including foundations and platforms. You will perform basic and detailed engineering, supervise junior engineers, and coordinate... 

    Jobtailor

    Madrid
    1 día atrás

¿Quieres recibir más ofertas?

Suscríbete y recibe ofertas similares para Site Reliability Engineer. ¡Entérate antes que nadie!