Principal DevOps Engineer

🏢 Landmark Group
📍 United Arab EmiratesFull-timeHybrid
📅 Posted: 3w ago🔄 Updated: 3w ago
CV%
✨ AI Summary
Landmark Group is seeking a Principal DevOps Engineer to lead their cloud infrastructure, platform engineering, DevOps, SRE, MLOps, AIOps, and DevSecOps initiatives. This hands-on leadership role is responsible for building highly scalable, secure, and automated platforms to accelerate software delivery. The ideal candidate will possess deep technical expertise, strong architectural and operational leadership, and a drive for platform modernization and cloud transformation. Key responsibilities include defining and driving the organization's platform strategy, designing cloud-native architectures, building self-service capabilities, and collaborating with various teams to establish best practices. The role involves designing and implementing CI/CD pipelines, IaC standards, and automating infrastructure provisioning. SRE duties include defining SLOs/SLIs, leading incident management, and building proactive monitoring frameworks. Experience with Kubernetes, containerization (Docker, EKS, GKE, ECS), and major cloud platforms (GCP, AWS, Azure) is crucial. Security and DevSecOps integration throughout the software delivery lifecycle is also a key focus. The position requires excellent communication, stakeholder management, and leadership skills, with a preference for candidates with experience in Internal Developer Platforms and multi-cloud/hybrid-cloud environments.
Required Skills
Information Technology
SREOpenTelemetry
Engineering, Construction & Trades
Structural Engineering
Other
AIOpsPlatform Engineering
🎁 Benefits & Perks
Opportunity to define and lead the DevOps, SRE, and Platform Engineering vision; Ownership of cloud infrastructure supporting large-scale e-commerce, logistics, supply chain and enterprise applications; Exposure to modern cloud-native, AI, platform engineering, and reliability engineering practices; Ability to influence architecture, engineering culture, and operational excellence; Work alongside senior technology leaders on strategic transformation initiatives.
Requirements
The ideal candidate will have a Bachelor's or Master's degree in a relevant field and at least 10 years of experience in DevOps, MLOps, AIOps, SRE, Infrastructure, Cloud, or Platform Engineering. Proven experience designing and operating large-scale production platforms, with deep expertise in cloud platforms (especially GCP, with AWS/Azure experience being desirable) and Kubernetes is required. Strong skills in Infrastructure as Code (Terraform, Ansible), CI/CD, Linux administration, and distributed systems are essential, along with excellent communication and leadership abilities.
Description
We are looking for a Principal DevOps Engineer to lead our cloud infrastructure, platform engineering, DevOps, SRE, MLOps, AIOps and DevSecOps initiatives. This is a hands-on leadership role responsible for building highly scalable, secure, resilient, and automated platforms that enable engineering teams to deliver software faster and more reliably.The ideal candidate combines deep technical expertise with strong architectural and operational leadership, driving platform modernization, automation, cloud transformation, and engineering excellence across the organization.Key ResponsibilitiesPlatform Engineering & Architecture● Define and drive the organization's DevOps, MLOps, AIOPs, SRE, Platform Engineering, and Infrastructure strategy.● Design highly available, scalable, secure, and resilient cloud-native architectures capable of supporting rapid business growth.● Build and maintain self-service platform capabilities that improve developer productivity and deployment velocity.● Collaborate with architects, engineering leaders, security teams, and product teams to establish engineering best practices and operational standards.● Lead cloud transformation programs, and platform optimization efforts.DevOps & Automation● Design and implement enterprise-grade CI/CD pipelines supporting multiple development teams and environments.● Establish Infrastructure as Code (IaC) standards using Terraform, Ansible, CloudFormation, and similar technologies.● Automate infrastructure provisioning, deployments, configuration management, and operational workflows.● Drive GitOps adoption and deployment automation across cloud and Kubernetes environments.● Standardize release management, environment management, and deployment governance processes.Site Reliability Engineering (SRE)● Define and implement Service Level Objectives (SLOs), Service Level Indicators (SLIs), and error budgets.● Drive reliability, performance, scalability, availability, and disaster recovery initiatives.● Lead incident management, root cause analysis, postmortems, and operational excellence programs.● Build proactive monitoring, alerting, observability, and capacity planning frameworks.● Optimize infrastructure performance, cost efficiency, and resource utilization across cloud environments.Cloud & Container Platforms● Design, build, and manage Kubernetes platforms across cloud environments.● Lead containerization initiatives using Kubernetes, Docker, EKS, GKE, ECS, and related technologies.● Architect multi-region, highly available cloud infrastructure on GCP, AWS, and Azure.● Implement cloud-native networking, service mesh, API gateways, and distributed systems best practices.● Establish backup, disaster recovery, business continuity, and resilience strategies.Security & DevSecOps● Embed security controls throughout the software delivery lifecycle.● Implement DevSecOps practices including vulnerability management, secrets management, policy-as-code, and infrastructure security.● Partner with security teams to ensure compliance, governance, and risk mitigation.● Drive cloud security best practices, identity and access management, encryption, and threat detection initiatives.Observability & Operations● Build centralized observability platforms covering metrics, logs, traces, events, and user experience monitoring.● Implement monitoring and logging solutions using Grafana, Prometheus, ELK, Dynatrace, Splunk, OpenTelemetry, and similar tools.● Lead production readiness reviews and operational health assessments.● Establish operational KPIs and continuously improve platform reliability and efficiency.Leadership & Mentorship● Provide technical leadership and mentorship to DevOps, SRE, Platform Engineering, and Infrastructure teams.● Define engineering standards, operational practices, and platform governance models.● Drive adoption of modern engineering practices across the organization.● Act as the technical escalation point for critical production and infrastructure challenges.Required Qualifications● Bachelor's or master's degree in computer science, Information Technology, Software Engineering, or related discipline.● 10+ years of experience in DevOps, MLOps, AIOps, Site Reliability Engineering, Infrastructure Engineering, Cloud Engineering, or Platform Engineering.● Proven experience designing and operating large-scale, mission-critical production platforms.● Deep expertise in cloud platforms, with strong hands-on experience in Google Cloud Platform (GCP). Experience with AWS and Azure is desirable.● Strong experience with Kubernetes, Docker, container orchestration, and microservices architectures.● Extensive experience implementing Infrastructure as Code using Terraform, Ansible, CloudFormation, or similar tools.● Strong experience building and managing enterprise CI/CD platforms.● Expert-level Linux administration and troubleshooting skills.● Experience managing distributed systems, event-driven architectures, and high-volume production workloads.● Strong understanding of networking, load balancing, DNS, CDN, security, and distributed computing concepts.● Excellent communication, stakeholder management, and leadership skills.Preferred Qualifications● Experience building Internal Developer Platforms (IDP) and self-service engineering platforms.● Experience implementing SRE frameworks including SLOs, SLIs, error budgets, and reliability engineering practices.● Experience with multi-cloud and hybrid-cloud architectures.● Experience supporting AI/ML platforms, MLOps, data platforms, and large-scale analytics workloads.● Experience implementing service mesh technologies such as Istio or Linkerd.● Relevant certifications such as Google Professional Cloud Architect, Google Professional DevOps Engineer, AWS Solutions Architect Professional, AWS DevOps Engineer Professional, Certified Kubernetes Administrator (CKA), or Certified Kubernetes Security Specialist (CKS).What You'll Gain● Opportunity to define and lead the DevOps, SRE, and Platform Engineering vision for a rapidly growing technology organization.● Ownership of cloud infrastructure supporting large-scale e-commerce, logistics, supply chain and enterprise applications.● Exposure to modern cloud-native, AI, platform engineering, and reliability engineering practices.● Ability to influence architecture, engineering culture, and operational excellence across the organization.● Work alongside senior technology leaders on strategic transformation initiatives.
✨ Premium Match Details
Deep-dive CV analysis, customized Cover Letters, and Interview prep!
📊 Match Analysis
Insights against your active CV
📊
Personalized Match Analysis
Upload your CV to see exact matching percentages, detailed skills mapping, and gap analysis for this role.
🎯 Overalli74%
⚡ Skillsi85%
View Breakdown
Ontology Match: 85.0
Matched:✓ Requirements Matching✓ Ontology Skills Mapping
📜 Eligibilityi49%
View Breakdown
Local: 19600%
🏗️ Career Fiti91%
View Breakdown
Seniority: 91.0
📋 Requirementsi67%
View Breakdown
Domain: 67.0
🔥 Motivationi78%
View Breakdown
Title Fit: 78.00