What You'll Do
Build and operate cloud infrastructure across AWS, Azure and GCP using Terraform and Terragrunt.
Manage Kubernetes platforms and application delivery using Docker, Helm, Argo CD and GitOps.
Own and improve CI/CD workflows, including builds, automated testing, preview environments, image publishing, deployments and rollbacks.
Monitor production systems through metrics, logs and traces; improve alerting, runbooks, capacity planning and incident response.
Strengthen platform reliability, security, networking, access management, secrets handling, backups and disaster recovery.
Build automation and self-service tools that make it easier and safer for engineering teams to deploy and operate services.
Partner with product and AI teams to productionize workloads and integrate model providers and supporting services.
What We're Looking For
5+ years of hands-on experience in DevOps, SRE or platform engineering.
Strong practical experience with AWS, Terraform, Kubernetes, Docker and CI/CD systems.
Experience with Helm, Argo CD and GitOps-based delivery.
Solid understanding of Linux, networking, cloud security and production observability.
Ability to automate operational work using Python, Bash or similar tools.
Experience with Azure or GCP is an advantage.
Experience supporting AI workloads, model integrations or data-intensive systems is a plus.
Clear communication in Hebrew and English, strong ownership and a collaborative approach.
Comfortable operating in a fast-moving startup and supporting production systems.