Company logo hidden

Senior Site Reliability Engineer, Vulcan (AI Security product)

Unlock employer United Arab Emirates Direct to Company Under an hour ago · 08 Sep 2026

Financial

  • Estimate: $60k - $120k*
  • Zero income tax location

Accessibility

  • Office Only
  • Apply from abroad
  • Relocation Support
  • Visa Provided

Requirements

  • Experience: Senior
  • English: Professional

Position

About the Role
We are seeking a hands-on infrastructure engineer to take ownership of the deployment, migration, and troubleshooting of on-premise Kubernetes environments for enterprise and government clients, particularly in airgapped, high-security data center environments where remote access is restricted. This client-facing, on-site role positions you as the technical authority—responsible for executing complex infrastructure changes correctly the first time, diagnosing failures independently under pressure, and maintaining clear communication with client stakeholders. Real ownership is a crucial aspect of this role as you will need to deeply understand the systems and make sound judgment calls when things do not go as planned, without relying on remote support. You will also play a vital role in supporting our Cybersecurity business, Vulcan, which offers GenAI cybersecurity solutions by providing red and blue team services to ensure compliance and security.

Ready to apply for roles like this?

Unlock the company name and direct application link. Subscribers get instant access to fresh jobs across Dubai, Abu Dhabi and Riyadh, many with visa support.

Unlock employer & apply directly

Responsibilities

  • Plan and execute on-prem Kubernetes cluster deployments, upgrades, and infrastructure migrations (including IP re-addressing, certificate rotation, and cluster reconfiguration) in both production and airgapped environments.
  • Diagnose and resolve failures independently on-site.
  • Own the full infrastructure stack from end-to-end: Kubernetes control plane and data plane, PostgreSQL (primary/replica replication), distributed storage (e.g., SeaweedFS/Ceph/similar), private container registries, and centralized logging (ELK or equivalent).
  • Validate deployment tooling (scripts, installers, automation) in lab/staging environments before any client-facing execution.
  • Clearly communicate the technical work directly to client stakeholders on-site: explain status, failures, and remediation plans.
  • Travel to client data centers (including airgapped/restricted-access sites) as required, sometimes on short notice, for deployment and go-live support.
  • Write clear and structured runbooks, decision trees, and incident reports for others (including less experienced engineers) to follow under pressure.
  • Proactively escalate risks to internal leadership, rather than waiting until after issues arise.

Requirements
Technical:

  • 5-6 years of hands-on experience with Kubernetes in production, including at least one on-premise (not purely cloud-managed) deployment.
  • Solid understanding of etcd internals, quorum, peer membership, and failure recovery, not limited to kubectl-level familiarity.
  • Experience with kubeadm-based cluster bootstrapping and certificate management (SANs, CA rotation, renewal).
  • Working knowledge of PostgreSQL replication, Linux networking fundamentals (DNS, NTP, firewalls), and container registries (Docker Distribution or similar).
  • Comfortable working entirely from the Linux command line, writing and debugging bash scripts, and reading unfamiliar automation tooling under time pressure.
  • Experience with at least one distributed storage system (SeaweedFS, Ceph, MinIO, or similar) is a strong plus.
  • Experience with GPU-enabled Kubernetes nodes (NVIDIA device plugin, container toolkit) is a plus, though not required.

Working Style:

  • Demonstrated ability to work independently in high-pressure, high-stakes environments without live support.
  • Strong incident communication skills; able to explain technical failures to non-technical stakeholders factually and calmly, without over-promising or minimizing issues.
  • A track record of validating changes in test environments prior to production involvement, and the judgment to insist on this even under deadline pressure.
  • Comfortable with travel, including to secure/restricted facilities where personal devices, internet access, or remote assistance may not be available.

Nice to Have:

  • Previous consulting, systems integration, or professional services experience, ideally with enterprise or government accounts.
  • Experience specifically in the GCC/Middle East region, or with government-sector clients.
  • Security background; the ability to reason about access controls, credential handling, and airgapped operational discipline is advantageous given the environments involved.

Interview Process

  • HR phone interview: 1 hour
  • Online interview: 1.5 - 2 hours, meet with the hiring manager
  • Online interview: 1 hour, meet with the hiring team

Why Join Us?

  • Innovative Environment: Be part of a company at the forefront of technology focused on providing security in GenAI, with opportunities to work on groundbreaking projects.
  • Growth Opportunities: Advance your career with our development programs and growth-focused culture.
  • Dynamic Team: Join a multi-cultural and dynamic team of dedicated professionals who inspire and support each other.
  • Compensation: Competitive salary and benefits package, commensurate with experience and performance.
Apply Direct

Jobs you might like   View all jobs

About Technology, Information and Internet Company

Company details are hidden. Subscribe to view full company profile.

Ready to apply for this role?

Apply Direct