Hamsa builds AI models that master Arabic dialects. Our platform covers speech-to-text, text-to-speech, translation and spoken language understanding, and powers voice agents and phone agents deployed by enterprises across MENA — in banking, telecom, government services, call centres, healthcare and hospitality.
That means our infrastructure carries live human conversations. When a phone agent is mid-call with a customer in Emirati or Levantine dialect, there is no retry, no queue, no "please refresh." Latency, availability and audio quality are product features, not ops metrics. If that sounds like an interesting engineering problem rather than a stressful one, keep reading.
About the role
We're hiring a Senior DevOps Engineer to own how our platform is built, deployed, secured and operated across AWS and Google Cloud. You'll be the most senior infrastructure voice on the team — setting the standards that backend, ML and integration engineers build on, and making the architectural calls rather than executing someone else's.
You'll work closely with our ML team on model serving and GPU capacity, with our voice team on real-time media and telephony infrastructure, and with our enterprise clients' technical teams on private cloud and self-hosted deployments.
What you'll do
- Own the architecture of our multi-cloud estate across AWS (EKS, VPC, IAM, CloudWatch, GuardDuty) and Google Cloud (GKE, Cloud Run, Vertex AI, Artifact Registry, GPU compute)
- Design and maintain the Terraform codebase so every environment — including client-dedicated ones — is reproducible and reviewable
- Run production Kubernetes: cluster architecture, upgrades, node pools, autoscaling for spiky voice traffic, network policies, RBAC and tenant isolation
- Build and operate the serving infrastructure for our STT and TTS models, including GPU node management, autoscaling and cost control
- Support the real-time voice stack — media servers, SIP and VoIP integrations, and the network path that keeps call quality high
- Design CI/CD and GitOps workflows with proper quality gates and safe rollout strategies: canary, blue/green, fast rollback
- Own observability end to end — metrics, logs, traces and alerting that actually help during a live incident, including call-quality and latency signals
- Lead incident response for infrastructure issues and write the post-mortem afterwards
- Package and support private cloud and on-premise deployments for enterprise and government clients, including the security documentation those engagements require
- Drive cloud cost discipline across compute, GPU and egress
- Implement security controls across IAM, network segmentation, secrets management and audit logging, and support client security reviews and compliance work
- Mentor mid-level engineers, review infrastructure changes, and document what you build
What we're looking for
- 5+ years in DevOps, Cloud, Platform or Site Reliability Engineering, with at least 3 years on production systems you were responsible for
- Deep hands-on experience with AWS or Google Cloud, and working competence in the other — we run both
- 3+ years operating Kubernetes in production, not just deploying to it
- 3+ years with Terraform, including module design and state management across environments
- Strong Linux administration, networking (VPC, DNS, TLS, load balancing) and production troubleshooting under pressure
- Proficiency in Python or Go, plus solid shell scripting
- Experience designing CI/CD pipelines for containerised microservices
- Working knowledge of Prometheus, Grafana, or equivalent observability tooling
- Comfortable working in Arabic and English
Nice to have
- GPU workloads or ML model serving — Vertex AI, SageMaker, Triton, vLLM, or self-hosted inference
- Real-time or low-latency systems: WebRTC, LiveKit, media servers, streaming, or VoIP and SIP
- Experience delivering on-premise or air-gapped deployments to enterprise or government clients
- Exposure to ISO 27001, SOC 2, or regional data residency requirements
- Certifications: AWS Solutions Architect or DevOps Engineer Professional, Google Professional Cloud DevOps Engineer, CKA, HashiCorp Terraform Associate
Why Hamsa
- Arabic voice AI is being built now, and you'd be building the infrastructure underneath it rather than integrating someone else's API
- Real ownership — this is a role where your architectural decisions stick
- Enterprise-grade problems at startup speed, with clients whose systems people actually depend on