Senior Manager - Reliability & Root Cause Analysis en España - Jobeax
Descripción de la vacante
Senior Manager - Reliability & Root Cause Analysis en España
Tiempo completoJornada semanal estándar
Full-time hours
Spain
Senior Manager - Reliability & Root Cause Analysis en España is listed on Jobeax. Browse 80,000+ vacancies available.
About the role
Job Description SummaryWe are seeking an experienced Manager - Reliability & Root Cause Analysis to lead a multidisciplinary engineering team within the Fleet Intelligence & Reliability organization. The team is responsible for identifying systemic fleet risks, determining the physical and systemic causes of complex failures, and ensuring that corrective and preventive actions deliver measurable and sustainable reliability improvement across the global Solar and Storage installed fleet.
This role owns the technical governance of Root Cause Analyses for complex, high-impact, recurring, and fleet-wide issues. The manager will ensure that investigations are evidence-based, technically rigorous, and completed with clear accountability from initial containment through corrective-action validation and fleet-risk closure.
The leader will manage engineers across relevant disciplines, which may include power electronics, electrical systems, mechanical systems, embedded software, plant controls, and systems engineering. The position requires close partnership with Fleet Performance & Analytics, Product Engineering, Quality, Manufacturing, Sourcing, Project Engineering, Field Operations, and commercial teams to convert fleet experience into product and operational improvement.Job DescriptionRoles and Responsibilities
What you'll do
Lead, coach, and develop a multidisciplinary engineering team responsible for fleet reliability, failure investigation, and Root Cause Analysis.
Own and continuously improve the RCA governance process, including issue prioritization, ownership assignment, technical reviews, milestones, escalation, corrective-action tracking, validation, and closure criteria.
Lead investigations of complex, high-impact, systemic, and recurring failures affecting solar inverters, battery energy storage systems, plant controls, power conversion equipment, and associated interfaces.
Ensure investigations are grounded in verified evidence, operating data, event logs, inspection results, failure reproduction, laboratory analysis, engineering calculations, and appropriate structured problem-solving methods.
Differentiate failure symptoms, direct physical causes, contributing factors, systemic causes, and organizational causes, ensuring that conclusions are supported by evidence.
Drive immediate containment, short-term mitigation, and permanent corrective and preventive actions with the relevant engineering and operational owners.
Prioritize investigations using safety exposure, customer impact, availability loss, recurrence, fleet population, warranty exposure, and financial consequence.
Maintain a fleet-level view of failure modes and systemic risk across products, projects, regions, configurations, software versions, suppliers, and operating environments.
Partner with Fleet Performance & Analytics to use fleet data, trends, event signatures, and risk models to detect emerging issues, quantify exposure, and verify corrective-action effectiveness.
Translate field and fleet learning into recommendations for product requirements, design changes, validation plans, software releases, manufacturing processes, supplier controls, maintenance strategies, and technical documentation.
Provide technical leadership during critical fleet events, customer escalations, and executive-level issue reviews.
Ensure customer-facing and internal RCA reports are accurate, evidence-based, clear, and delivered in accordance with committed milestones.
Develop and maintain reusable engineering knowledge, including RCA standards, failure libraries, lessons learned, troubleshooting guides, technical advisories, and investigation templates.
Support the definition and deployment of Change, Modification, and Upgrade actions resulting from fleet reliability findings.
Establish and monitor meaningful metrics such as RCA cycle time, milestone adherence, recurrence rate, corrective-action closure, investigation backlog, and quantified fleet-risk reduction.
Promote a culture of safety, technical rigor, accountability, collaboration, transparency, and continuous improvement.
What we're looking for
Required Qualifications
Bachelor's degree in Electrical Engineering, Mechanical Engineering, Systems Engineering, Control Systems Engineering, Power Electronics, or a related technical field.
Demonstratable experience in engineering, reliability, failure analysis, product engineering, installed-base engineering, or a related technical function.
Demonstrated experience leading complex technical investigations or RCAs involving multidisciplinary systems.
Demonstrated experience leading engineers, technical teams, or significant cross-functional engineering programs.
Experience with power generation, renewable energy, battery energy storage, power conversion, industrial automation, or comparable complex industrial equipment.
Ability and willingness to travel internationally as required.
Nice to have
Desired Characteristics
Advanced degree in engineering, reliability, systems engineering, or a related discipline.
Previous people-leadership experience within a global engineering or technical organization.
Strong knowledge of solar inverters, battery energy storage systems, power electronics, plant controls, electrical protection, thermal management, and industrial communication networks.
Experience applying structured investigation methodologies such as Fault Tree Analysis, 5 Whys, Ishikawa analysis, 8D, Kepner-Tregoe, or Failure Mode and Effects Analysis.
Knowledge of reliability engineering, failure mechanisms, accelerated testing, design validation, reliability growth, and corrective-action effectiveness assessment.
Experience analyzing operational data, waveforms, alarms, event logs, failed components, inspection evidence, and intervention history.
Strong systems thinking, technical judgment, and ability to make risk-based decisions with incomplete or evolving information.
Proven ability to influence Product Engineering, Quality, Manufacturing, Sourcing, Operations, and commercial teams in a matrixed environment.
Strong customer-facing skills and the ability to communicate sensitive technical findings with clarity, credibility, and appropriate transparency.
Ability to manage multiple high-priority investigations across a global fleet.
Excellent written and verbal communication skills.
Self-starting attitude with the ability to establish structure, set priorities, develop people, and drive issues through sustainable closure.
... seamless, resilient, and scalable platform around the clock. We are looking for an experienced, proactive and innovative professional that is keen to join as a Senior Site Reliability Engineer! The mission of Nexthink's SRE team is to strengthen our infrastructure and enhance our ability to deploy, monitor, and scale systems ...
... members who want to join us on our mission to lead cloud security globally. Does this sound like the right place for you? What you'll do - You will report to the Manager of Customer Reliability, work as an important member of the technical support team providing senior level technical knowledge on our customer issues. - Work ...
Want your engineering skills to enable real scientific breakthroughs? We’re looking for a Senior Site Reliability Engineer to join our highly skilled IT Operations team at EMBL-EBI. At EMBL-EBI, our IT & Technical Services department underpins groundbreaking research that improves human and planetary health. As part of ...
Drive Train Expert / Técnico/a de Vibraciones – Mechanical Field Team (Lugo) ¿Eres especialista en mecánica de grandes componentes con foco en vibraciones? ¿Te interesa el diagnóstico avanzado y las reparaciones complejas en aerogeneradores?
... cloud-native services CI/CD and DevOps expertise - Advanced knowledge of CI/CD pipelines, automated testing, and deployment automation in cloud environments Root cause analysis and resilience - Persistent in identifying systemic issues, understanding failure modes in distributed systems, and driving solutions to completion ...
... the field of power conversion, energy storage, cooling, electrical distribution, and control software into modular enclosures. About the Role We are seeking a Senior Design Reliability & Safety Engineer to join our Data Center team of quality and customer satisfaction. This role is ideal for professionals with a strong background ...
About the role We are looking for a Site Reliability Engineer (SRE) passionate about infrastructure reliability, automation, and the development of scalable production systems. What you'll do - Own and improve production infrastructure reliability and stability - Prepare, execute, and support deployments and infrastructure ...
What you'll do - Service Reliability and Optimization: Focus on capacity planning and launch reviews for services before they go live. Perform blameless postmortems and proactive identification of potential outages to foster iterative improvements - Accountability/Problem Solving: Resolves complex problems in a global Kubernetes-based ...
... posibilidades de conseguir una entrevista leyendo la siguiente descripción general de este puesto antes de presentar su candidatura. Si tienes experiencia como Site Reliability Engineer (SRE) , trabajando en entornos cloud, con foco en observabilidad, automatización y operación 24/7, y te interesa participar en la evolución y mejora ...
... the field of power conversion, energy storage, cooling, electrical distribution, and control software into modular enclosures. About the Role We are seeking a Senior Design Reliability & Safety Engineer to join our Data Center team of quality and customer satisfaction. This role is ideal for professionals with a strong background ...
... PPAP, APQP, and supplier development initiatives. - Strong root cause analysis and problem-solving skills utilizing methodologies such as 8D, Fishbone, Pareto Analysis, 5-Why, and Failure Analysis. - Experience with manufacturing data analysis using tools such as Python, SQL, Power BI, Excel, Minitab, JMP, or similar platforms. ...
... technical authority for complex customer challenges, owning escalations across Azure, AWS, and ERP-integrated environments. You will lead incident response, drive root cause analysis, and partner cross-functionally to improve system resilience and customer experience. What you’ll do: Troubleshoot and resolve advanced issues ...
... expert-level capability in system-level debugging across hardware, operating system, drivers, firmware, and applications. The Principal Engineer drives deep root cause analysis, enforces structured problem-solving methodologies, and delivers scalable, permanent solutions that improve overall product quality and reliability. ...
You'll most likely need Spanish to apply. This job was automatically translated to English, see original posting . About the role 🚀 We are looking for a Senior Test Manager! | Technology Migration Project What you'll do Strategy & Innovation: Define and document the comprehensive testing concept and integrate methodologies ...
... M&A • Ex-CFO/M&A @Curatible (exited to Blackstone) • Ex-President of the Board @SotremoSA (exited) • Co-founder/CFO @SoftOne (exited) Your Role Most Product Manager roles are built around specs and handoffs, waiting on a designer's mockup, an engineer's sprint slot, a data analyst's report. This role is built differently: ...
... EverGuard - an internal AI designed to ensure everything we build is safe, ethical, and human-first . And we're only just getting started! Your Role Most Product Manager roles are built around specs and handoffs, waiting on a designer's mockup, an engineer's sprint slot, a data analyst's report. This role is built differently: ...
About the job This job was automatically translated to English, see original posting . At SKLUM, we are looking for a Brand Manager to join our PR and Communications team at our offices in Villalonga (10 minutes from Gandía). In this role, you will be responsible for defining brand strategy and positioning, ensuring consistency ...
... understand our development roadmap and know how to use new functionalities. What we're looking for - 3+ years of experience in iGaming (either as an Account Manager or in a similar client-facing platform role). - Advanced commercial negotiation skills with a proven ability to build, explain, and defend financial proposals ...
... role. - Interview with the Hiring Manager: Assess your experience building products, including relevant domain context where applicable. - Interview with a Senior Business Leader: Assess how you think and act as a product and business operator in complex real-world scenarios. - Interview with the CPO (or Senior Product ...