JobBobsReal-time global job discoveryLive

Senior Cloud Resilience Architect

Blinkhealth · India · 2026-06-30

executive
Apply on the employer's site

About this role

<div class="content-intro"><p><strong>Company Overview:<br><br></strong><a href="https://www.blinkhealth.com/">Blink Health </a>is the fastest growing healthcare technology company that builds products to make prescriptions accessible and affordable to everybody.  Our two primary products – BlinkRx and Quick Save – remove traditional roadblocks within the current prescription supply chain, resulting in better access to critical medications and improved health outcomes for patients. <br><br><a href="https://www.blinkhealth.com/blinkrx">BlinkRx </a>is the world’s first pharma-to-patient cloud that offers a digital concierge service for patients who are prescribed branded medications. Patients benefit from transparent low prices, free home delivery, and world-class support on this first-of-its-kind centralized platform. With BlinkRx, never again will a patient show up at the pharmacy only to discover that they can’t afford their medication, their doctor needs to fill out a form for them, or the pharmacy doesn’t have the medication in stock. <br><br>We are a highly collaborative team of builders and operators who invent new ways of working in an industry that historically has resisted innovation. Join us!</p></div><h3><strong>Responsibilities</strong></h3> <ul> <li><strong>Evaluate and mature the organization’s disaster recovery posture</strong>, including recovery objectives (RTO/RPO), dependency mapping, and failure domain analysis across applications, data, and infrastructure.<br><br></li> <li>Define, document, and <strong>establish disaster recovery standards and best practices</strong> across cloud infrastructure, platforms, and application architectures.<br><br></li> <li>Partner with SRE, platform, security, and product engineering teams to <strong>design and implement resilient, fault-tolerant systems</strong>, progressing from backup-based recovery to <strong>multi-region and active-active architectures</strong>.<br><br></li> <li>Lead the <strong>disaster recovery roadmap</strong>, balancing technical feasibility, cost, risk, and business priorities.<br><br></li> <li>Design and recommend <strong>reference architectures</strong> for disaster recovery patterns, including pilot-light, warm standby, hot standby, and active-active.<br><br></li> <li>Drive adoption of <strong>active-active disaster recovery</strong> for critical systems, including traffic management, data replication, consistency models, and automated failover.<br><br></li> <li>Define and operationalize <strong>testing strategies</strong> for DR, including game days, chaos testing, and regular recovery exercises.<br><br></li> <li>Establish clear <strong>documentation, runbooks, and escalation paths</strong> to ensure recoverability is well understood and not dependent on individuals.<br><br></li> <li>Evaluate and recommend <strong>platform upgrades, cloud services, and tooling</strong> that improve resilience, recovery speed, and reliability.<br><br></li> <li>Serve as a <strong>technical authority and advisor</strong> on disaster recovery and resilience for leadership and engineering teams.<br><br></li> <li>Provide architectural guidance, design reviews, and mentorship to engineers implementing DR-related changes.<br><br></li> <li>Partner with security and compliance teams to ensure DR strategies meet regulatory, audit, and data protection requirements.<br><br></li> </ul> <h3><strong>Desired Experience</strong></h3> <ul> <li>Bachelor’s or Master’s degree in Computer Science or equivalent practical experience.<br><br></li> <li>8+ years of experience in cloud infrastructure, platform engineering, SRE, or reliability-focused architecture roles.<br><br></li> </ul> <h4><strong>Disaster Recovery & Resilience</strong></h4> <ul> <li>Deep understanding of <strong>disaster recovery concepts</strong> including RTO/RPO, blast radius reduction, failure domains, and dependency isolation.<br><br></li> <li>Proven experience designing and implementing <strong>multi-region and multi-availability zone architectures</strong>.<br><br></li> <li>Hands-on experience moving systems toward <strong>active-active or highly available architectures</strong>.<br><br></li> <li>Strong grasp of <strong>data replication strategies</strong>, consistency tradeoffs, and recovery patterns for databases and stateful systems.<br><br></li> </ul> <h4><strong>Cloud & Platform Engineering</strong></h4> <ul> <li>Extensive experience with major cloud providers (<strong>AWS preferred</strong>, GCP/Azure acceptable).<br><br></li> <li>Strong understanding of managed cloud services and their DR characteristics and limitations.<br><br></li> <li>Experience with <strong>Kubernetes-based platforms</strong>, including regional failover, workload portability, and cluster recovery strategies.<br><br></li> <li>Familiarity with global traffic management, DNS, load balancing, and service mesh patterns.<br><br></li> </ul> <h4><strong>Automation…

Skills asked for

Similar jobs

Apply on the employer's site

Your next role is already in here.

Search live openings from thousands of employers, save the ones worth a second look, and let JobBob keep watch for the rest.