DevOps Technical Lead

Link Development · Cairo, Egypt · Posted 2026-07-14

Role FocusAll positions are Senior / Lead Site Reliability Engineers, focused on enabling engineering teams to build stable, scalable platforms and to perform "boring" (i.e., predictable, low-risk) product releases.In essence: the SRE function acts as a release and reliability enabler across the organization — ensuring predictable deployments, strong operational standards, and high system resilience. Key ResponsibilitiesAct as embedded SRE within one or more product/engineering teamsOwn service reliability end-to-end: development, deployment, and productionBuild and operate cloud-native platforms based on KubernetesImplement Infrastructure as Code and CI/CD automationDevelop observability solutions (monitoring, logging, alerting)Lead incident response and root cause analysisDefine and enforce operational guardrails and release frameworksDrive continuous improvement in performance, resilience, and cost efficiencyEnable safe, predictable, and standardized releases across teamsCoach engineering teams toward higher operational maturity (not just "keep the lights on") Cloud HyperscalerStrong hands-on experience with at least one major hyperscaler (AWS, Azure, or GCP).AWS strongly preferred.KubernetesHands-on experience: cluster setup, application deployment, stateful and stateless workloads, autoscalingInfrastructure as CodeSkilled in at least one of: Terraform, Ansible, Chef, PuppetCI/CDExperience setting up and managing pipelines with Jenkins and/or GitLab CIObservabilityProficient implementing monitoring/logging solutions, e.g. Prometheus, Grafana, ELK stackProgrammingWorking familiarity with at least one of: Node.js, Golang, JavaIdeal Profile (Experience & Soft Skills)Deep hands-on experience in SRE / DevOps / Platform Engineering (please indicate years)Proven experience leading or managing engineering teams or technical squadsStrong ownership mindset; able to operate independently with minimal supervisionExperience in distributed systems and production-critical environmentsStrong incident management experience, including on-call rotationsAbility to influence across teams without formal/direct authorityPrior experience working in embedded team structures (i.e., not a siloed central ops team)Strong communication skills in English (spoken and written)Additional ContextA key focus of this engagement is scaling reliability practices across multiple teams. Candidates are expected not only to operate systems, but to actively:Define and evangelize operational standardsImprove release processes across teamsCoach engineering teams toward higher operational maturity

Apply for this role

Skills mentioned in this role

GoNode.jsAWSAzureGCPTerraform

Preparing for a Software & IT role

  • Public work matters more than titles: a GitHub or portfolio the hiring manager can click through in under a minute usually beats an extra bullet on your CV.
  • Name the exact stack in your CV so a keyword screener + a skim-reader both find it — languages, frameworks, cloud, databases.
  • For interviews: expect a technical screen (coding or system design) before the on-site loop. Practice explaining trade-offs out loud, not just writing correct code.
  • For English-only teams verify your written English is sharp — most Egypt-based tech scaleups run interview loops entirely in English.

Other open roles at Link Development

See all 9 open roles at Link Development →

Related jobs in Software & IT