Senior Site Reliability Engineer en Spain, España - Jobeax
Descripción de la vacante
Senior Site Reliability Engineer en Spain, España
Airalo
RemotoTrabaja desde cualquier lugar
Tiempo completoJornada semanal estándar
ContratoTemporal o freelance
£94,000 - £130,000 a year
Spain
Senior Site Reliability Engineer en Spain, España is listed on Jobeax. Browse 80,000+ vacancies available.
About the role
At Airalo, we're making it easier for people to stay connected wherever they travel. As the world's first eSIM store, we help millions of travelers access affordable mobile data in 200+ countries and regions around the world.
Today, we're a team of 400+ people across 60+ countries, building a product used by travelers every day. We've grown quickly, but we've worked hard to keep what matters: trust, ownership, and the freedom for people to do great work without unnecessary layers or bureaucracy.
We're fully remote by design, genuinely global, and united by a shared mission to make travel simpler for everyone.
We are looking for a Senior Site Reliability Engineer to join our growing engineering team.
We are a company that values SRE principles and practices. We believe in empowering our SREs to make data-driven decisions, automate operational tasks, and continuously improve the reliability of our systems. We foster a blameless culture where everyone is encouraged to learn from mistakes and share knowledge. If you are passionate about building and maintaining highly reliable systems, we would love to hear from you!
What you'll do
Lead the design of scalable, fault-tolerant and self-healing systems in a multi-region AWS environment.
Define and track Service Level Objectives (SLOs) and Service Level Indicators (SLIs) to drive architectural decisions and error budget policies.
Conduct blameless post-incident reviews to uncover systemic root causes and implement long-term preventive measures.
Identify patterns of manual work and lead the development of internal tools/automation to permanently eliminate them.
Develop and maintain automated runbooks and playbooks for common operational tasks and complex incident response.
Shift from simple monitoring to deep observability, ensuring high cardinality data leads to proactive actionable insights.
Proactively identify and mitigate operational risks through chaos engineering and architecture reviews.
Work with software engineers to design systems for reliability, scalability, and maintainability from the early stages of the SDLC.
Continuously evaluate and optimize system performance, capacity, and cost efficiency.
Beyond just participating, you will refine the on-call experience to reduce alert fatigue, improve MTTR, and ensure sustainable rotation health.
What we're looking for
Bachelor's degree in Computer Engineering or a similar discipline.
5+ years of experience as a Site Reliability Engineer or in a similar role.
3+ years of experience with AWS services including strong knowledge of container orchestration.
2+ years of Kubernetes experience
Deep understanding of observability principles and tools such as: Prometheus, Datadog, OpenTelemetry and similar.
Experience with leading incident management and complex postmortem analysis.
Experience and interest in managing infrastructure as code (Terraform).
Experience with chaos engineering and other techniques for testing system resilience.
Experience with CI/CD tools such as GitHub Actions for automated delivery.
Proficiency in at least one programming language (Python, Go, Java, etc.) for building automation and internal tooling.
Ability to work independently and collaboratively in a fast-paced environment.
Team player and open to new ideas.
Good communication skills and fluency in English.
Nice to have
Prior experience with Scrum and other agile methods.
Certification in relevant areas such as AWS Certified DevOps Engineer, Certified Kubernetes Administrator (CKA), or similar.
Prior experience with Telco Core Networks (e.g., 5G/LTE Packet Core, IMS, Signaling) and low-latency networking.
Experience with AI-driven SRE tools for anomaly detection and improvements
Contributions to open-source SRE projects or communities.
Prior work experience in telecommunications.
Deep understanding of eSIM and GSMA related technologies and services.
Languages
Languages: English is our main working language day to day, so you'll need to be comfortable communicating in it both in meetings and async.
What you'll get
Benefits: Learn more about our benefits here in this link - https://jobeax.com/link/TWLNQH5o0x3BATYD
Paid Rotation: We offer standby fees + overtime pay.
How you'll work
Location: Remote, anywhere in Spain or the UK.
Contract:
Spain: Full-time, permanent contrato indefinido via Deel (our employer of record in Spain)
UK: Full-time, permanent
Participating in our on-call rotation is a core expectation of this role. It's essential for maintaining 24/7 service reliability across our global operations, ensuring our systems remain resilient and our customers experience uninterrupted service, regardless of time zone or geography.
Delayed Start: No on-call duties for your first 6 months.
Rest & Recovery: Guaranteed rest periods and flexible hours following night incidents.
Shared Load: Rotations are split (Weekdays vs. Weekends) to minimize fatigue.
About the company & team
Curious what it's actually like to work at Airalo? We've put together a guide for each of the hubs we hire from most, covering our benefits, remote culture, employment setup, and what our team values most:
Life at Airalo · España: https://jobeax.com/link/dOQ1M09DESDbasBl
Life at Airalo · United Kingdom: https://jobeax.com/link/LPeBWHDmx9whDKH8
We started Airalo to make staying connected effortless, wherever you are in the world. Today, millions of travelers rely on our eSIMs, and many of the people building the product were customers first.
Our team is united by a shared belief that great work happens when people are trusted to own what they do. As Airalo continues to grow, so do the opportunities to learn, take on new challenges, and make a real impact.
For many of our teammates in Spain, it's the combination of stability, flexibility, and genuine work-life balance that makes Airalo feel like home.
We started Airalo to make staying connected effortless, wherever you are in the world. Today, millions of travelers rely on our eSIMs, and many of the people building the product first discovered us as customers.
Our team spans 60+ countries, but what people often notice quickly is how work actually feels here. There's trust to own what you do, space to focus without constant layers of process, and a clear sense of what you're responsible for from day one.
As Airalo continues to grow, people tend to stay for a mix of things: the chance to work on a product used globally, the autonomy to shape how they work, and the opportunity to keep taking on new challenges without getting stuck in one lane.
Hiring process
If you are interested in this position, please apply via the link.
On-Call
Please refer to the On-Call Policy in the Airalo Handbook for full details: https://jobeax.com/link/Jl1pAfWV9YPjHW7g
Additional information
By applying, you acknowledge and agree that, in case of successful application, Airalo may request to run background checks as a condition for entering into an agreement with you. Rest assured that these checks will only occur upon your prior consent and at the end of the selection process, and will be strictly limited to what is allowed under the laws that are applicable to you. All data that you share or that we collect in connection with such checks will be processed in accordance with our Privacy Policy, available here:
https://jobeax.com/link/SBeygKENu4X1abuq
We sincerely thank all applicants in advance for submitting their interest in this opportunity. Airalo is an equal-opportunity employer and values diversity, equity & inclusion. We do not discriminate on the basis of race, religion, national origin, gender, sexual orientation, age, marital status, veteran status, or disability status. We are committed to providing reasonable accommodations upon request for individuals with disabilities throughout our job interview process.
... seamless, resilient, and scalable platform around the clock. We are looking for an experienced, proactive and innovative professional that is keen to join as a Senior Site Reliability Engineer! The mission of Nexthink's SRE team is to strengthen our infrastructure and enhance our ability to deploy, monitor, and scale systems ...
Want your engineering skills to enable real scientific breakthroughs? We’re looking for a Senior Site Reliability Engineer to join our highly skilled IT Operations team at EMBL-EBI. At EMBL-EBI, our IT & Technical Services department underpins groundbreaking research that improves human and planetary health. As part of ...
... systems-minded engineer who cares deeply about reliability, scalability, and production excellence? Join Lodgify as a Senior Site Reliability Engineer and help our engineering teams build and operate services that are reliable, observable, scalable, and resilient by design. In this role, you will improve the reliability of our ...
... solid set of product ideas lined up ready for innovative engineers to tackle. And of course we have big plans to take over the taxi app service industry! Site Reliability Engineers at Cabify work on improving all aspects of our platform and have an impact across the whole organisation. They are a blend of systems engineers ...
About the role As a Senior SRE at Remote, you'll work with a high degree of autonomy on complex reliability and platform problems, owning the plan and execution of features and projects within our SRE/Platform domain. You'll contribute to the platform's architecture and reliability strategy, translating ambiguous requirements ...
... users worldwide. Our commitment to reliability is a key foundation of our product and our dedication to exceeding customer availability expectations is a core engineering focus. As a Senior Site Reliability Engineer, you'll join our SRE team based in Europe to ensure our production systems are not only operational but also ...
About the role We're looking for a Senior Site Reliability Engineer to join the Software Logistics Team (CI/CD) within our Platform Engineering Segment. Platform Engineering's mission is to provide easy-to-use, self-service platforms that enable other segments to build, deploy, and monitor their business applications with ...
... reinventing credit to make it more honest and friendly, giving consumers the flexibility to buy now and pay later without any hidden fees or compounding interest. Site Reliability Engineering at Affirm is a small, yet crucial, team that helps our Engineering partners to "Operate What They Own" with excellence to protect their ...
... Elastic's complete, cloud-based solutions for search, security, and observability help organizations deliver on the promise of AI. As a Principal Platform Engineer focused on capacity, you will play a crucial role in managing and optimizing our compute resources, ensuring that our Elastic Cloud Hosted and Serverless workloads ...
About the role We're hiring a Senior Site Reliability Engineer to join our Platform team and help us build secure, reliable, and scalable systems as Maze grows. This role sits at the intersection of cloud infrastructure, reliability, and security . You'll work across our platform to improve how we operate production systems, ...
... resilience - Persistent in identifying systemic issues, understanding failure modes in distributed systems, and driving solutions to completion Performance and reliability focus - Analytical mindset toward observing end-to-end service performance, system health, and user impact in production environments Dynamic environment ...
About the role We are looking for a Site Reliability Engineer (SRE) passionate about infrastructure reliability, automation, and the development of scalable production systems. What you'll do - Own and improve production infrastructure reliability and stability - Prepare, execute, and support deployments and infrastructure ...
... Migrations - Analyze and plan complex migrations. What are we looking for? - Bachelor's degree in Computer Engineering, Electronics Engineering, Telecommunications Engineering, or a related field. - 3+ years of experience in telecommunications or related fields. - Experience working in Site Reliability Engineering, DevOps, Infrastructure ...
... evolución y mejora continua de plataformas complejas (M2M), queremos conocerte. ¿Te unes a nuestro #WelcomeHome? ¿Qué buscamos? Experiencia sólida de 3 a 6 años como Site Reliability Engineer (SRE), DevOps Engineer o en roles similares , trabajando sobre entornos productivos críticos. Experiencia en soporte técnico y funcional ...
About the role We are seeking a Site Reliability Engineer to join the Observability group inside our Platform Engineering domain. Platform Engineering’s goal is to provide easy to use, self-service platforms to enable other segments to easily build, deploy and monitor their business applications. And Observability’s role ...
... grow together. You will be part of a culture that values trust, accountability, and shared success where your work truly matters. Your Impact Join a team of senior engineers operating in a large-scale, multi-cloud production environment supporting tens of thousands of enterprise customers worldwide. This is not a typical ...
... are invited, and ideas matter. A team where everyone makes play happen. Reports to: Technical Director Hiring process External careers site, Internal careers site Locations : Madrid, Spain Required Qualifications - 3+ years of experience as a Site Reliability Engineer, DevOps Engineer, or Cloud Infrastructure Engineer. ...
... place for you? What you'll do - You will report to the Manager of Customer Reliability, work as an important member of the technical support team providing senior level technical knowledge on our customer issues. - Work with and foster relationships with the development engineering team to ensure tracking on engineering ...
About the role Senior Site Reliability Engineer - Observability We are looking for a Senior Site Reliability Engineer with deep observability expertise to strengthen the reliability and visibility of Shopmonkey's production infrastructure. This is a hands-on contract role for an experienced SRE who has built production-grade ...