Weekday AI
Weekday AI·2 days ago

Senior platform engineer - cloud & kubernetes

Apply now

Location

onsite, Abu Dhabi, United Arab Emirates

Salary

$32k – $55k/yr

Commitment

Full Time

Level

Senior (5+ years)

Required skills

Microsoft AzureKubernetesAKSTerraformCI/CDDockerHelmAzure DevOpsObservabilityPrometheusGrafanaLokiLinuxBash scriptingGitOpsDisaster Recovery

Job Description

This role is for one of the Weekday's clients

Salary range: Rs 2600000 - Rs 4500000 (ie INR 26-45 LPA)

Experience: 5+ yrs

Location: Abu Dhabi, United Arab Emirates

Job Type: Full-time

We are looking for an experienced Senior Platform Engineer to design, implement, secure, and operate scalable cloud-native platforms using Microsoft Azure and Kubernetes. The role focuses on building highly available, resilient, secure, and well-governed production environments while driving automation, observability, infrastructure-as-code, and platform engineering best practices.

The ideal candidate will have strong hands-on expertise across Azure, AKS, Kubernetes, Terraform, CI/CD, cloud governance, security, observability, and Disaster Recovery, along with the ability to provide technical leadership across complex platform environments.

Key Responsibilities

  • Design, deploy, and manage Microsoft Azure infrastructure and AKS/Kubernetes platforms.
  • Implement and maintain Disaster Recovery, backup, restore, and business continuity solutions.
  • Build infrastructure automation using Terraform, Helm, Azure DevOps, and CI/CD pipelines.
  • Manage Kubernetes networking, ingress, storage, namespaces, resource limits, security, and platform configurations.
  • Implement observability solutions using Prometheus, Grafana, Loki, and related monitoring technologies.
  • Manage TLS certificates, secrets, access controls, RBAC, and platform security mechanisms.
  • Establish cloud and Kubernetes governance covering resource organisation, naming, tagging, security policies, and operational standards.
  • Define Infrastructure-as-Code and CI/CD governance, including reusable Terraform modules, state management, code reviews, approval gates, environment promotion, and artifact management.
  • Participate in architecture and technical design reviews for platforms, applications, integrations, and infrastructure changes.
  • Drive security and compliance readiness through vulnerability remediation, access reviews, security baselines, audit controls, and policy enforcement.
  • Define and monitor platform availability, SLIs, SLOs, capacity, performance, and operational health.
  • Lead incident and problem management, including root-cause analysis, corrective actions, and prevention of recurring issues.
  • Perform capacity planning, performance optimisation, and cloud cost optimisation / FinOps activities.
  • Own DR testing, RTO/RPO validation, backup and recovery standards, and periodic recovery exercises.
  • Plan and execute platform migrations, infrastructure upgrades, and cloud transformation initiatives.
  • Evaluate emerging platform technologies, conduct POCs, and establish approved patterns for production adoption.
  • Maintain technical documentation including HLDs, LLDs, architecture diagrams, SOPs, runbooks, troubleshooting guides, and DR procedures.
  • Provide technical leadership, mentoring, and knowledge sharing across platform engineering teams.

What Makes You a Great Fit

  • 5+ years of experience in platform engineering, cloud infrastructure, DevOps, SRE, or related roles.
  • Strong hands-on expertise in Microsoft Azure and Kubernetes, particularly AKS.
  • Strong experience with Docker, Helm, Terraform, and Azure DevOps / CI/CD.
  • Good understanding of Azure networking, including VNets, NSGs, Private Endpoints, and Azure Firewall.
  • Experience implementing observability using Prometheus, Grafana, Loki, or similar platforms.
  • Strong knowledge of Linux, Bash scripting, Git, and GitOps practices.
  • Proven experience with Disaster Recovery, backup/restore, RTO/RPO planning, and recovery testing.
  • Strong understanding of Kubernetes security, governance, RBAC, policy enforcement, and Azure Policy.
  • Experience establishing reusable Infrastructure-as-Code patterns and platform engineering standards.
  • Strong understanding of CI/CD governance, release controls, environment promotion, and source-control standards.
  • Experience working with databases such as PostgreSQL, MongoDB, MySQL, or Azure SQL.
  • Strong knowledge of incident management, problem management, RCA, capacity planning, performance optimisation, and cloud cost management.
  • Experience working in security-, compliance-, audit-, or governance-controlled environments.
  • Ability to create and maintain HLDs, LLDs, architecture diagrams, operational runbooks, SOPs, and technical documentation.
  • Strong experience participating in architecture and technical design reviews.
  • Excellent troubleshooting, analytical, and problem-solving capabilities.
  • Strong technical leadership, mentoring, communication, and cross-functional collaboration skills.
  • Experience in banking or other regulated industries would be an advantage.
  • Strong understanding of high availability, production operations, platform modernisation, and cloud transformation initiatives.

Ready to join the team?

Apply now