Zipher is building the Autonomous Execution Layer for cloud data and AI workloads. Backed by $50M in funding , we dynamically orchestrate clusters, predict bottlenecks, and auto-heal infrastructure in real time — with zero human intervention . Our platform runs in production at global enterprise customers, including Fortune 500 companies , delivering mission-critical resilience and sub-second optimization We are looking for a DevOps Engineer to build and own the platform our autonomous execution engine runs on. You will own how Zipher ships, scales, and stays up — the delivery pipelines, the Kubernetes infrastructure, and the reliability of systems enterprise customers depend on around the clock.
What You’ll Do
* Architect and own GitOps-based delivery — ArgoCD, Helm, Terraform, GitHub Actions — so the engine ships to production safely, many times a day
* Build and operate the Kubernetes platform , including on-demand environments spun up per pull request and torn down automatically
* Design deployment safety into the platform: canary analysis, blue/green rollouts, automated rollback, and an observability stack (Prometheus, Grafana, OpenTelemetry) that surfaces failure first
* Partner closely with backend and data engineers to make production infrastructure secure, reproducible, and cost-aware across AWS accounts and enterprise deployments
* Drive reliability end to end: define SLOs , own incident response, and raise the operational bar for mission-critical services
What We Offer
* Own the platform behind a new category of autonomous cloud infrastructure High ownership from day one : real architectural influence, direct exposure to founders, and responsibility for mission-critical systems
* A small, technical, high-velocity team that values curiosity, speed, rigor, and engineering craftsmanship Top-of-market compensation and meaningful equity
Ready to own the platform that lets an autonomous execution engine run in production? Hit Apply.
Requirements: What You’ll Bring 6+ years in DevOps, Platform, or Infrastructure Engineering , including ownership of production environments for a real product at scale
* Deep hands-on experience with Kubernetes, Helm, and Terraform , with GitOps (ArgoCD or equivalent) as your default way to ship
* Strong production experience on AWS — EKS, IAM, VPC networking, and managed services such as Lambda, S3, DynamoDB, or Kinesis — plus scripting in Python, Go, or Bash
* Real depth in observability and operations : metrics, tracing, log aggregation, alerting, and SLOs you defined and defended
* A high-agency, engineering-first mindset : you enjoy ambiguous, high-leverage problems and take responsibility for reliability, performance, and security
Nice to Have
* Experience building ephemeral environments with Crossplane, Terraform, or a home-grown control plane
* Experience operating data and streaming infrastructure at scale , such as Kafka, Spark, EMR, MongoDB Atlas, or Snowflake
* Familiarity with security and compliance in a fast-moving startup: SOC 2, SSO/MFA, secret scanning, IaC policy enforcement, and cloud cost optimization
* Experience in an elite IDF technology unit or another high-performance engineering environment
This position is open to all candidates.