Site Reliability Engineer

🏢 superbot (bot)
📍 United Arab EmiratesFull-timeOn-site
📅 Posted: Yesterday🔄 Updated: Yesterday
CV%
✨ AI Summary
As a Site Reliability Engineer, you will be responsible for the reliability, scalability, and resilience of critical services by partnering with product and platform teams. Your key responsibilities include defining and tracking SLIs/SLOs/SLAs, leading post-mortems, and managing incidents to reduce MTTR. You will design, build, and maintain cloud infrastructure on AWS using Terraform, and scale and optimize the Kubernetes-based container platform. Strengthening observability through Prometheus and Grafana, accelerating CI/CD pipelines, and acting as an internal reliability advisor to engineering teams are also crucial aspects of this role. The ideal candidate will have 4+ years of experience in SRE, DevOps, or platform engineering, with hands-on Kubernetes and Terraform experience, proficiency in scripting/programming languages like Python, Go, or Bash, and experience with observability tools. An AWS certification is also required.
Required Skills
Information Technology
CI/CDGo
🎁 Benefits & Perks
Competitive Compensation, performance-based bonuses, accommodation, meal allowances, assistance with work visa processing, generous holiday and New Year bonuses, MacBook and iPhone.
Requirements
Requires 4+ years in an SRE, DevOps, or platform engineering role with production ownership at scale. Must have hands-on Kubernetes experience, infrastructure-as-code fluency with Terraform, and observability stack experience with Prometheus and Grafana. Proficiency in at least one scripting or programming language (Python, Go, or Bash) for automation is necessary. AWS certification (Solutions Architect, DevOps Engineer, or SysOps Administrator) is also required.
Description
About The RoleYou'll sit at the intersection of software engineering and infrastructure, partnering with product and platform teams to keep our systems observable, scalable, and resilient. From shaping our on-call culture to driving infrastructure-as-code adoption, you'll have real influence over how BOT builds and operates software at scale — for 500+ people and growing.Key ResponsibilitiesOwn reliability across critical services — define and track SLIs/SLOs/SLAs, lead blameless post-mortems, and drive down MTTR through systematic incident management.Design, build, and maintain cloud infrastructure on AWS (primary) using Terraform, ensuring environments are reproducible, version-controlled, and auditable.Scale and optimize our Kubernetes-based container platform — capacity planning, resource tuning, autoscaling, and cluster lifecycle management.Strengthen observability end-to-end: instrument services with Prometheus and Grafana, build actionable alerting, and reduce alert noise through continuous tuning.Accelerate CI/CD pipelines (GitHub Actions / GitLab CI) to support frequent, safe deployments — shift reliability left by embedding checks into the delivery workflow.Partner with engineering teams as an internal reliability advisor — run game days, advocate for SRE best practices, and help developers build services that operate well from the start.Job Requirements4+ years in an SRE, DevOps, or platform engineering role with production ownership at scale.Hands-on Kubernetes experience — deployment, scaling, networking, and troubleshooting in a production environment.Infrastructure-as-code fluency with Terraform (or Pulumi) across a major cloud provider (AWS preferred).Observability stack experience — Prometheus, Grafana, and/or equivalent tools with a track record of building meaningful dashboards and alerts.Proficiency in at least one scripting or programming language (Python, Go, or Bash) for automation and tooling.AWS certification (Solutions Architect, DevOps Engineer, or SysOps Administrator).What We OfferCompetitive Compensation: Enjoy a salary package tailored to your skills and experience, along with performance-based bonuses.Comprehensive Benefits: We support your well-being with accommodation, meal allowances, and assistance with work visa processing.Work-Life Balance: Unwind with generous holiday and New Year bonuses.Top-Tier Equipment: Stay productive with the latest tools, including a MacBook and iPhone.Thriving Culture: Immerse yourself in a dynamic, inclusive work environment that fosters growth.
✨ Premium Match Details
Deep-dive CV analysis, customized Cover Letters, and Interview prep!
📊 Match Analysis
Insights against your active CV
📊
Personalized Match Analysis
Upload your CV to see exact matching percentages, detailed skills mapping, and gap analysis for this role.
🎯 Overalli74%
⚡ Skillsi85%
View Breakdown
Ontology Match: 85.0
Matched:✓ Requirements Matching✓ Ontology Skills Mapping
📜 Eligibilityi49%
View Breakdown
Local: 19600%
🏗️ Career Fiti91%
View Breakdown
Seniority: 91.0
📋 Requirementsi67%
View Breakdown
Domain: 67.0
🔥 Motivationi78%
View Breakdown
Title Fit: 78.00