Principal Cloud Kubernetes Architect Disaster Recovery

Aviso de fuente externaen BMA Group Global

We are seeking a Principal Cloud, Kubernetes & Disaster Recovery Architect to serve as a senior technical authority responsible for defining enterprise strategy, governance, and architect...

Fuente externa - sin verificarhace 11 díasVigente hasta: 21 oct 2026

Salario

No especificado

Ubicación

San Juan, PR

Tipo de empleo

Tiempo completo

Modalidad

No especificado

Principal Cloud Kubernetes Architect Disaster Recovery

San Juan, PR

Descripción del empleo

We are seeking a Principal Cloud, Kubernetes & Disaster Recovery Architect to serve as a senior technical authority responsible for defining enterprise strategy, governance, and architecture across hybrid cloud infrastructure, Kubernetes platforms, Infrastructure as Code, and Business Continuity/Disaster Recovery (BCDR).
This is a senior individual contributor role for a highly experienced infrastructure or platform engineering professional who can establish technical standards, design resilient environments, and advise leadership on cloud strategy, operational risk, scalability, and disaster readiness.
Key ResponsibilitiesDefine enterprise Kubernetes and container platform strategy across hybrid, multi-cloud, and on-premises environments.Develop reference architectures and standardized deployment patterns for microservices, data platforms, and analytics workloads.Establish Infrastructure as Code governance and reusable automation using Terraform, ArgoCD, Helm, GitOps, and CI/CD pipelines.Design hybrid cloud infrastructure strategies that support scalability, performance, capacity planning, and cost optimization.Define and lead enterprise-wide BCDR strategy across Azure, AWS, GCP, and on-premises infrastructure.Architect multi-region and multi-cloud recovery solutions, including active-active, active-passive, pilot-light, and warm-standby models.Establish recovery objectives, governance frameworks, documentation standards, ownership models, and testing programs.Lead disaster recovery simulations, failover testing, risk assessments, and continuous improvement initiatives.Define resilience strategies for Kubernetes clusters, workloads, data platforms, backups, and critical infrastructure.Establish enterprise observability standards across metrics, logs, tracing, dashboards, alerting, and operational reporting.Define service reliability standards, including SLOs, SLIs, SLAs, capacity planning, and self-healing capabilities.Implement security standards such as RBAC, encryption, secure recovery procedures, and zero-trust architecture.Partner with Security, Risk, Audit, Compliance, SRE, Platform, Infrastructure, and Data teams.Provide technical guidance to executive leadership and mentor engineering teams.
Required QualificationsBachelor’s degree in Computer Science, Computer Engineering, Information Systems, or a related field; master’s degree preferred.12+ years of experience in SRE, infrastructure engineering, cloud architecture, or platform engineering.Deep experience designing and operating Kubernetes platforms in enterprise environments.Strong understanding of hybrid cloud, multi-cloud, and on-premises infrastructure architectures.Proven leadership of enterprise Business Continuity and Disaster Recovery strategy and execution.Hands-on experience with cloud disaster recovery solutions such as Azure Site Recovery, AWS disaster recovery services, or comparable multi-cloud technologies.Expertise in Terraform, ArgoCD, Helm, GitOps, and modern CI/CD practices.Strong knowledge of distributed systems, resilience engineering, high availability, backup, replication, and workload recovery.Experience establishing RTO and RPO requirements and leading full-scale disaster recovery testing.Understanding of regulatory and security frameworks such as NIST, ISO 27001, HIPAA, SOC 2, and FDA requirements.Experience supporting audits and producing evidence-based technical documentation, runbooks, recovery procedures, and architecture diagrams.Strong communication skills with the ability to advise leadership and collaborate across technical and business teams.Full professional fluency in English is required; bilingual English and Spanish is preferred.
This opportunity is ideal for a senior technical leader who combines strong Kubernetes and cloud architecture expertise with enterprise-level ownership of infrastructure resilience and disaster recovery.

¿Es tuya esta vacante?

Reclámala gratis y recibe candidatos con video en CazVid.

Empleos similares