Senior Site Reliability Engineer en España - Jobeax
Descripción de la vacante
Senior Site Reliability Engineer en España
RemotoTrabaja desde cualquier lugar
HíbridoCombinación de oficina y remoto
Tiempo completoJornada semanal estándar
ContratoTemporal o freelance
€60,000 - €85,500 a year
Spain
Senior Site Reliability Engineer en España is listed on Jobeax. Browse 80,000+ vacancies available.
About the job
At Nexthink, we empower our customers with industry-leading solutions to enable continuous improvement of employee experience. We deliver unmatched visibility across all environments, so IT teams can consistently see, diagnose, and fix digital workplace issues. As a SaaS provider, our commitment is to deliver a seamless, resilient, and scalable platform around the clock.
We are looking for an experienced, proactive and innovative professional that is keen to join as a Senior Site Reliability Engineer! The mission of Nexthink's SRE team is to strengthen our infrastructure and enhance our ability to deploy, monitor, and scale systems effectively and reliably. They work closely with over 50 Product Engineering teams that develop our products and services, as well as with the Technical Platform Engineering, Security and Architecture teams to understand the reliability requirements, design and implement solutions, and promote them for adoption and usage.
Join our vibrant team of diverse and experienced engineers where cutting-edge technology meets innovation. Be a part of Nexthink's Digital Employee Experience technological revolution, ensuring our global customers enjoy a seamless user experience. Apply now and become a key player in our dynamic SRE organisation.
What you'll do
Implement and manage cloud-native systems (AWS) using best-in-class tools and automation.
Operate and enhance Kubernetes clusters, deployment pipelines, and service meshes to support rapid delivery cycles.
Design, build, and maintain the infrastructure powering our multi-tenant SaaS platform with reliability, security, and scalability in mind.
Define and maintain SLOs, SLAs, and error budgets, and proactively address availability and performance issues.
Develop infrastructure-as-code (Terraform or similar) for repeatable and auditable provisioning.
Build internal platform tools and automation to support provisioning, monitoring, and operational efficiency.
Monitor infrastructure and applications ensuring high-quality user experiences.
Participate in a shared on-call rotation, responding to incidents, troubleshooting outages, and driving timely resolution and communication.
Act as an Incident Commander during the on-call duty and coordinate cross-team responses effectively to maintain an SLA.
Drive and refine incident response processes, reducing Mean Time to Detect (MTTD) and Mean Time to Recovery (MTTR).
Diagnose and resolve complex issues independently, minimizing the need for external escalation.
Work closely with software engineers to embed observability, fault tolerance, and reliability principles into service design.
Automate runbooks, health checks, and alerting to support reliable operations with minimal manual intervention.
Support automated testing, canary deployments, and rollback strategies to ensure safe, fast, and reliable releases.
Contribute to security best practices, compliance automation, and cost optimization.
What we're looking for
Minimum Bachelor’s degree in Computer Science or equivalent practical experience.
5+ years of experience as a Site Reliability Engineer or Platform Engineer with strong knowledge of software development best practices.
Strong hands-on experience with public cloud services (AWS, GCP, Azure) and supporting SaaS product.
Strong programming or scripting skills (e.g., Python, Go, Bash...), and experience with infrastructure-as-code (e.g. Terraform).
Proficiency with Kubernetes, container-based deployment (e.g., Docker) and related ecosystems (e.g., Helm).
Experience with managing monitoring solutions (e.g. Datadog).
Comfortable participating in a rotating on-call schedule, managing critical incidents, and leading post-incident reviews.
At ease with operating and managing production systems, striking the right balance between urgency and methodology.
Strong system-level troubleshooting skills and a proactive mindset toward incident prevention.
Deep understanding of Linux systems, networking, and common troubleshooting practices.
Solid understanding of the network stack (e.g., TCP/IP, VPN, etc.), cloud architectures (VPC, subnets, firewalls, load balancers), service mesh (e.g., Istio) and storage (e.g., S3, EBS, etc).
Knowledge of zero-downtime deployment strategies, blue/green and canary releases.
Exposure to compliance standards such as SOC 2, ISO 27001, or HIPAA. FedRAMP experience is a big plus.
Experience with chaos engineering or resilience testing practices.
Excellent problem-solving skills, collaborative mindset, and a strong grasp of agile, iterative development.
Self-driven, highly organised, and capable of independently managing priorities.
Curiosity to learn new things and discover new technologies.
Strong communication, presentation, and team collaboration skills.
Excellent written and verbal skills in English.
The prior experience with any of the above-mentioned tools is a bonus, but not a must! We encourage you to apply even if you do not meet every single requirement. We welcome candidates with different level of background and experience. If you are excited about this role, please apply and our recruiters will assess your application.
What you'll get
Additional Information
We are the pioneers and trailblazers of a global IT Market Category (DEX) that is shaping the future of how the world works, giving our customers’ IT Teams total digital visibility across their enterprise. Our innovative solutions integrate real-time analytics, automation, and employee feedback across all endpoints. This enables our IT teams to solve complex technical challenges, create ever more productive workplaces, and deliver happy, satisfied employees in the digital workplace.
With over 1000 employees across 5 continents, Nexthink operates as One Team, connecting, collaborating and innovating to continuously grow. We call our employees ‘Nexthinkers’ and our commitment to diversity, inclusion, and equity is second to none. We currently have over 75 nationalities working with us, from all cultures and backgrounds, speaking many different languages.
IIf you are looking for a change and like a nice atmosphere, lots of challenges, and having fun while working, this is a great opportunity for you! Check what we offer:
💼 Permanent Contract and a competitive compensation package.
📍 Beautiful office with a view of Lake Geneva, conveniently located next to the Prilly-Malley train station
🏡 Hybrid work model balancing office and remote work, with a structured approach for new hires to foster connections and onboarding.
🏖️ Flexible Hours and unlimited vacation (employees have unlimited paid time off on top of the 25 days of holidays we offer) plus 3 company-paid volunteer days.
🤸 Free access to a fitness centre inside the building.
🚞 Reimbursement of the half-fare travel card for public transport.
🧑🏫 Reimbursement up to 50% of the cost of French classes.
🍉 Fresh fruit, cookies, and soft drinks as well.
🤝 Regular company and team events like Voluntary Days, Pizza talks, Team Building activities, hosting Meetups at the office and more!
📣 Bonuses for referring successful hires after three months of continuous employment.
🚚 We offer a relocation package to people who are coming from another country.
Please note that not all the benefits listed above are available for temporary, contract, and internship roles. To ensure you have the most up-to-date information, we recommend checking with your Recruitment Partner.
The base salary for this role is €60,000 – €85,500 gross per year, with a total on-target earnings (OTE) range of €66,000 – 93,000€ including an annual performance bonus. You'll also be part of our broader total rewards package — including benefits tailored to where you live and how you work best.
We set our pay ranges using objective criteria: the scope and level of the role, the skills it takes to do it well, and the relevant market data. Ranges are reviewed every year to remain competitive and fair. We're transparent about this because we think you deserve to know what you're working towards from day one. In accordance with the EU Pay Transparency Directive (2023/970), we publish salary ranges on every Nexthink role. We won't ask what you currently earn or your previous salary. What matters to us is what this role is worth and whether it works for you. Nexthinkers come from all kinds of backgrounds, and that's what makes us stronger. We welcome applications from everyone.
About the company & team
Nexthink is the leader in digital employee experience management software. The company provides IT leaders with unprecedented insight allowing them to see, diagnose and fix issues at scale impacting employees anywhere, with any application or network, before employees notice the issue. As the first solution to allow IT to progress from reactive problem solving to proactive optimization, Nexthink enables its more than 1,300 customers to provide better digital experiences to more than 18 million employees. Dual headquartered in Lausanne, Switzerland and Boston, Massachusetts, Nexthink has 9 offices worldwide.
Want your engineering skills to enable real scientific breakthroughs? We’re looking for a Senior Site Reliability Engineer to join our highly skilled IT Operations team at EMBL-EBI. At EMBL-EBI, our IT & Technical Services department underpins groundbreaking research that improves human and planetary health. As part of ...
... systems-minded engineer who cares deeply about reliability, scalability, and production excellence? Join Lodgify as a Senior Site Reliability Engineer and help our engineering teams build and operate services that are reliable, observable, scalable, and resilient by design. In this role, you will improve the reliability of our ...
... layers or bureaucracy. We're fully remote by design, genuinely global, and united by a shared mission to make travel simpler for everyone. We are looking for a Senior Site Reliability Engineer to join our growing engineering team. We are a company that values SRE principles and practices. We believe in empowering our SREs ...
... solid set of product ideas lined up ready for innovative engineers to tackle. And of course we have big plans to take over the taxi app service industry! Site Reliability Engineers at Cabify work on improving all aspects of our platform and have an impact across the whole organisation. They are a blend of systems engineers ...
About the role As a Senior SRE at Remote, you'll work with a high degree of autonomy on complex reliability and platform problems, owning the plan and execution of features and projects within our SRE/Platform domain. You'll contribute to the platform's architecture and reliability strategy, translating ambiguous requirements ...
... users worldwide. Our commitment to reliability is a key foundation of our product and our dedication to exceeding customer availability expectations is a core engineering focus. As a Senior Site Reliability Engineer, you'll join our SRE team based in Europe to ensure our production systems are not only operational but also ...
About the role We're looking for a Senior Site Reliability Engineer to join the Software Logistics Team (CI/CD) within our Platform Engineering Segment. Platform Engineering's mission is to provide easy-to-use, self-service platforms that enable other segments to build, deploy, and monitor their business applications with ...
... reinventing credit to make it more honest and friendly, giving consumers the flexibility to buy now and pay later without any hidden fees or compounding interest. Site Reliability Engineering at Affirm is a small, yet crucial, team that helps our Engineering partners to "Operate What They Own" with excellence to protect their ...
... Elastic's complete, cloud-based solutions for search, security, and observability help organizations deliver on the promise of AI. As a Principal Platform Engineer focused on capacity, you will play a crucial role in managing and optimizing our compute resources, ensuring that our Elastic Cloud Hosted and Serverless workloads ...
About the role We're hiring a Senior Site Reliability Engineer to join our Platform team and help us build secure, reliable, and scalable systems as Maze grows. This role sits at the intersection of cloud infrastructure, reliability, and security . You'll work across our platform to improve how we operate production systems, ...
... resilience - Persistent in identifying systemic issues, understanding failure modes in distributed systems, and driving solutions to completion Performance and reliability focus - Analytical mindset toward observing end-to-end service performance, system health, and user impact in production environments Dynamic environment ...
About the role We are looking for a Site Reliability Engineer (SRE) passionate about infrastructure reliability, automation, and the development of scalable production systems. What you'll do - Own and improve production infrastructure reliability and stability - Prepare, execute, and support deployments and infrastructure ...
... Migrations - Analyze and plan complex migrations. What are we looking for? - Bachelor's degree in Computer Engineering, Electronics Engineering, Telecommunications Engineering, or a related field. - 3+ years of experience in telecommunications or related fields. - Experience working in Site Reliability Engineering, DevOps, Infrastructure ...
... evolución y mejora continua de plataformas complejas (M2M), queremos conocerte. ¿Te unes a nuestro #WelcomeHome? ¿Qué buscamos? Experiencia sólida de 3 a 6 años como Site Reliability Engineer (SRE), DevOps Engineer o en roles similares , trabajando sobre entornos productivos críticos. Experiencia en soporte técnico y funcional ...
About the role We are seeking a Site Reliability Engineer to join the Observability group inside our Platform Engineering domain. Platform Engineering’s goal is to provide easy to use, self-service platforms to enable other segments to easily build, deploy and monitor their business applications. And Observability’s role ...
... grow together. You will be part of a culture that values trust, accountability, and shared success where your work truly matters. Your Impact Join a team of senior engineers operating in a large-scale, multi-cloud production environment supporting tens of thousands of enterprise customers worldwide. This is not a typical ...
... are invited, and ideas matter. A team where everyone makes play happen. Reports to: Technical Director Hiring process External careers site, Internal careers site Locations : Madrid, Spain Required Qualifications - 3+ years of experience as a Site Reliability Engineer, DevOps Engineer, or Cloud Infrastructure Engineer. ...
... place for you? What you'll do - You will report to the Manager of Customer Reliability, work as an important member of the technical support team providing senior level technical knowledge on our customer issues. - Work with and foster relationships with the development engineering team to ensure tracking on engineering ...
About the role Senior Site Reliability Engineer - Observability We are looking for a Senior Site Reliability Engineer with deep observability expertise to strengthen the reliability and visibility of Shopmonkey's production infrastructure. This is a hands-on contract role for an experienced SRE who has built production-grade ...