Our Client Currently looking for Director, Data & Storage Reliability Engineering
Qualifications
To be successful in this role you have:
- Experience in leveraging or critically thinking about how to integrate AI into work processes, decision-making, or problem-solving. This may include using AI-powered tools, automating workflows, analyzing AI-driven insights, or exploring AI’s potential impact on the function or industry.
- Strong product mindset with demonstrated experience treating technical capabilities as products with roadmaps, priorities, customers, adoption goals, and measurable business outcomes.
- Experience translating production insights, customer pain points, operational challenges, reliability risks, and platform telemetry into prioritized engineering investments and long-term roadmaps.
- Experience partnering closely with production operations, customer escalation teams, reliability organizations, and software engineering teams to drive systemic improvements based on operational learnings.
- Experience defining product strategies, developing roadmaps, prioritizing investments, and aligning stakeholders across multiple organizations without direct authority.
- Experience operating a portfolio of engineering investments, balancing short-term customer needs with long-term reliability, performance, scalability, and resilience objectives.
- 15+ years of experience in software engineering, platform engineering, reliability engineering, infrastructure engineering, database engineering, distributed systems, product management, or large-scale SaaS environments.
- 8+ years of engineering leadership experience, including leading managers and globally distributed teams.
- Extensive experience leading Reliability Engineering, Platform Engineering, Database Engineering, Infrastructure Engineering, Production Engineering, Performance Engineering, or related technical organizations.
- Deep expertise in distributed systems, databases, storage technologies, cloud infrastructure, and large-scale SaaS architectures.
- Strong understanding of reliability engineering principles, observability, scalability, resiliency, operational excellence, and performance engineering.
- Experience building and operating observability, telemetry, diagnostics, reliability, or performance capabilities at scale.
- Proven experience identifying systemic issues and converting operational insights into strategic engineering improvements.
- Experience partnering closely with Product Management organizations to influence roadmaps and deliver customer-centric outcomes.
- Experience driving engineering initiatives through data, metrics, customer impact analysis, and measurable business outcomes.
- Experience leveraging AI technologies to improve decision-making, analytics, engineering workflows, operational efficiency, reliability insights, automation, or customer outcomes.
- Exceptional communication, stakeholder management, and leadership skills.
- Bachelor’s degree in Computer Science, Engineering, or a related technical field, or equivalent practical experience.
Desired Skills
- Previous Product Management experience in a platform, infrastructure, cloud, database, storage, or SaaS environment.
- Experience applying product management disciplines such as roadmap planning, prioritization, customer-centric thinking, outcome measurement, and portfolio management to engineering organizations.
- Experience operating large-scale enterprise database and storage platforms supporting mission-critical workloads.
- Experience building and scaling Reliability Engineering, Performance Engineering, Platform Engineering, SRE, or Production Engineering organizations.
- Experience with observability platforms, telemetry systems, diagnostics frameworks, and production analytics.
- Experience with migration readiness, resiliency validation, reliability testing, operational risk reduction, and large-scale cloud transformations.
- Experience leveraging AI technologies to improve anomaly detection, forecasting, incident analysis, prioritization, and engineering productivity.
- Strong understanding of distributed systems architecture, cloud platform operations, and hyperscale environments.
- Experience developing executive-facing reliability scorecards, engineering metrics, and business impact reporting.
- Experience influencing platform architecture, database strategy, storage strategy, and long-term engineering roadmaps.
- Experience with Linux-based production environments and large-scale cloud infrastructure.
- Experience supporting enterprise database technologies such as MySQL, MariaDB, PostgreSQL, Oracle, SQL Server, or cloud-native database platforms.
- Familiarity with ServiceNow platform architecture and large-scale SaaS operations.
-
Are you interested in this position?
Apply by clicking on the “Apply Now” button below!
#AlbionarcJobs#FintechJobs
#AsiaJobs#MiddleEastCareers
#TechTalent#FintechRecruitment
#FinanceOpportunities#
