Director, Data & Storage Reliability Engineering

Engineer

Director, Data & Storage Reliability Engineering

Apply Now

- $0.00

  • Date posted
    September 10, 2026
  • Expiration date
    December 10, 2026
  • Application ends
    December 10, 2026

Our Client Currently looking for Director, Data & Storage Reliability Engineering

Qualifications

 

To be successful in this role you have:

  • Experience in leveraging or critically thinking about how to integrate AI into work processes, decision-making, or problem-solving. This may include using AI-powered tools, automating workflows, analyzing AI-driven insights, or exploring AI’s potential impact on the function or industry.
  • Strong product mindset with demonstrated experience treating technical capabilities as products with roadmaps, priorities, customers, adoption goals, and measurable business outcomes.
  • Experience translating production insights, customer pain points, operational challenges, reliability risks, and platform telemetry into prioritized engineering investments and long-term roadmaps.
  • Experience partnering closely with production operations, customer escalation teams, reliability organizations, and software engineering teams to drive systemic improvements based on operational learnings.
  • Experience defining product strategies, developing roadmaps, prioritizing investments, and aligning stakeholders across multiple organizations without direct authority.
  • Experience operating a portfolio of engineering investments, balancing short-term customer needs with long-term reliability, performance, scalability, and resilience objectives.
  • 15+ years of experience in software engineering, platform engineering, reliability engineering, infrastructure engineering, database engineering, distributed systems, product management, or large-scale SaaS environments.
  • 8+ years of engineering leadership experience, including leading managers and globally distributed teams.
  • Extensive experience leading Reliability Engineering, Platform Engineering, Database Engineering, Infrastructure Engineering, Production Engineering, Performance Engineering, or related technical organizations.
  • Deep expertise in distributed systems, databases, storage technologies, cloud infrastructure, and large-scale SaaS architectures.
  • Strong understanding of reliability engineering principles, observability, scalability, resiliency, operational excellence, and performance engineering.
  • Experience building and operating observability, telemetry, diagnostics, reliability, or performance capabilities at scale.
  • Proven experience identifying systemic issues and converting operational insights into strategic engineering improvements.
  • Experience partnering closely with Product Management organizations to influence roadmaps and deliver customer-centric outcomes.
  • Experience driving engineering initiatives through data, metrics, customer impact analysis, and measurable business outcomes.
  • Experience leveraging AI technologies to improve decision-making, analytics, engineering workflows, operational efficiency, reliability insights, automation, or customer outcomes.
  • Exceptional communication, stakeholder management, and leadership skills.
  • Bachelor’s degree in Computer Science, Engineering, or a related technical field, or equivalent practical experience.

Desired Skills 

  • Previous Product Management experience in a platform, infrastructure, cloud, database, storage, or SaaS environment.
  • Experience applying product management disciplines such as roadmap planning, prioritization, customer-centric thinking, outcome measurement, and portfolio management to engineering organizations.
  • Experience operating large-scale enterprise database and storage platforms supporting mission-critical workloads.
  • Experience building and scaling Reliability Engineering, Performance Engineering, Platform Engineering, SRE, or Production Engineering organizations.
  • Experience with observability platforms, telemetry systems, diagnostics frameworks, and production analytics.
  • Experience with migration readiness, resiliency validation, reliability testing, operational risk reduction, and large-scale cloud transformations.
  • Experience leveraging AI technologies to improve anomaly detection, forecasting, incident analysis, prioritization, and engineering productivity.
  • Strong understanding of distributed systems architecture, cloud platform operations, and hyperscale environments.
  • Experience developing executive-facing reliability scorecards, engineering metrics, and business impact reporting.
  • Experience influencing platform architecture, database strategy, storage strategy, and long-term engineering roadmaps.
  • Experience with Linux-based production environments and large-scale cloud infrastructure.
  • Experience supporting enterprise database technologies such as MySQL, MariaDB, PostgreSQL, Oracle, SQL Server, or cloud-native database platforms.
  • Familiarity with ServiceNow platform architecture and large-scale SaaS operations.
  • Are you interested in this position?

     

    Apply by clicking on the “Apply Now” button below!

     

    #AlbionarcJobs#FintechJobs

    #AsiaJobs#MiddleEastCareers

    #TechTalent#FintechRecruitment

    #FinanceOpportunities#

     

Apply Now

- $0.00

Select your currency