remotely.living

DevOps

OptiPlay · Remote - Countries of Europe or Ukraine · full-time · 2026-08-31

Apply for this job

Job description

We are launching our entire platform from scratch — fully cloud-native, automated, and built on top of DevOps and Infrastructure-as-Code principles.

Our stack: AWS, Kubernetes, Terraform, GitOps, PostgreSQL, ClickHouse, Redis, RabbitMQ, microservices, Node.js, Pixi.js.

We are a small and fast-growing engineering team, and we are looking for a Senior DevOps Engineer who will own the whole cloud infrastructure and shape the technical direction of the platform.

🚀 What you will build

- Build our entire AWS cloud infrastructure from scratch following automation-first, HA, and security-by-design principles;

- Design and operate Kubernetes environments; create and maintain custom Helm charts;

- Implement Infrastructure-as-Code using modular Terraform structures;

- Design and maintain CI/CD pipelines and GitOps delivery workflows;

- Build secure networking architecture: VPC, VPN, WAF, firewalls, ingress controllers;

- Set up monitoring, logging, alerting, and tracing (Prometheus / Grafana / Loki / Tempo);

- Deploy and maintain infrastructure for PostgreSQL, ClickHouse, Redis, RabbitMQ;

- Define solutions for container registry and artifact management;

- Ensure system scalability, reliability, observability, and security;

- Participate in architectural discussions and influence OptiPlay’s platform engineering strategy.

🧩 What makes you a great match

- 4+ years of experience in DevOps, Cloud, or Platform Engineering;

- AWS production experience: VPC, VPN, WAF, firewalls, multi-environment setups, networking fundamentals;

- Kubernetes: strong understanding of internals, Helm, container runtime ecosystem;

- Terraform: modular IaC design, reusable infrastructure patterns, infra from scratch;

- CI/CD & GitOps: GitLab CI, ArgoCD, automated deployment workflows;

- Monitoring & Observability: Prometheus, Grafana, Loki, Tempo, or similar.

- Datastores & Messaging: PostgreSQL, ClickHouse, Redis, RabbitMQ.

- Networking & Proxying: NGINX, HAProxy, Traefik, ingress controllers.

- Automation & Scripting: Bash, Python (or similar).

- Artifact Management: modern registries like ECR, Harbor, Nexus, etc.

- Experience designing infrastructure for high-load, distributed, real-time systems.

📄 Nice to Have

- Terragrunt

- SSL4SaaS

- Experience with GCP or Azure

- Multi-cluster or multi-region architecture

- Deep Kubernetes ecosystem experience (operators, CRDs, service mesh)

🎁 Benefits

- 21 vacation days + 5 extra day-offs annually;

- 12 paid sick days;

- Fully remote format — work from anywhere you feel productive;

- Flexible schedule: start your day anytime between 08:00–11:00 CET;

- Fixed budget for health insurance and gym/fitness;

- Provided all required work equipment;

- Zero bureaucracy and direct communication with founders and C-level;

- Minimal meetings, async-friendly workflow;

- Startup energy: fast motion, creativity, and tight-knit communication;

- Business trips and team meetups several times per year;

- Multiple salary payout options (flexible formats);

- …And many more perks unlocked after we hit break-even.