Senior DevOps Engineer (Infrastructure)
WFA Digital Insight
Zeely’s AI‑driven marketing platform demands a DevOps leader who can tame a sprawling multi‑cloud environment. The Senior DevOps Engineer will own the full lifecycle of infrastructure across AWS, GCP, Cloudflare and Hetzner, translating high‑level reliability goals into Terraform modules and Kubernetes patterns that keep the platform responsive for over a million users. What sets this role apart is the blend of hands‑on cloud engineering with strategic cost‑optimization, as Zeely expects the engineer to embed FinOps thinking while scaling CI/CD pipelines and observability stacks. Collaboration is baked in: the position works side‑by‑side with product engineers to improve developer experience and with on‑call teams to sharpen incident response. Candidates should be comfortable navigating both the tactical details of EKS and the broader architecture vision.
Job Description
Join us as a Senior DevOps Engineer and take ownership of evolving and scaling the infrastructure behind Zeely’s AI-driven platform.
We’re looking for a senior infrastructure engineer who thrives in ownership, complexity, and scale — someone who can design resilient systems, define infrastructure standards, and shape the long-term architecture of our platform. In this role, you’ll work closely other DevOps engineers & engineering teams to improve reliability and developer experience across a complex multi-cloud environment, scaling our CI/CD and observability stack, leading infrastructure integrations, and optimizing cloud costs while strengthening monitoring and alerting.
Zeely AI is an all-in-one AI markeing platform that helps small businesses and marketing teams grow faster and achieve their goals through high-performing content and automation. Our platform enables users to create ad creatives, UGC videos, viral content, banners, and more, as well as launch effective paid and organic campaigns across Meta channels.
In simple terms, we’ve made world-class AI marketing tools and solutions accessible to businesses that previously could only get this level of support from large brands with expensive marketing teams.
A few things that define where we are today:
4 years on the market — a stable, profitable company backed by top international investors;
4x growth over the past year;
1.5M+ users already use Zeely, and that number keeps growing every day;
#1 in AIGC in the US market;
200+ specialists on the team;
US-focused, with a global user base.
Your Key Responsibilities:
Maintain and optimize AWS infrastructure, including EKS, RDS, SQS, S3, OpenSearch and related services, to ensure reliability, scalability and efficient resource usage.
Manage and scale GCP environments supporting AI services, data pipelines and analytics workloads, including services such as BigQuery.
Maintain supporting infrastructure across Cloudflare and Hetzner where applicable, including networking, edge services and production traffic routing.
Drive Infrastructure as Code practices with Terraform across cloud providers to ensure consistency, maintainability and clear infrastructure ownership.
Maintain and improve Kubernetes-based production environments, ensuring high availability, autoscaling and stable deployment processes.
Refine monitoring, logging and alerting to improve signal quality, incident response and infrastructure visibility.
Implement, maintain and improve APM and observability tooling.
Ensure high availability and resilience of production infrastructure under growing product workloads.
Contribute to the development of our multi-cloud infrastructure strategy with a focus on AWS and GCP.
Collaborate with engineering teams to improve CI/CD pipelines, infrastructure reliability and developer experience.
Continuously optimize cloud infrastructure costs through FinOps practices while maintaining performance and reliability.
Participate in on-call rotations, ensuring timely response to critical infrastructure incidents when they occur.
What You Need to Join Us:
Must have:
6+ years of hands-on commercial DevOps experience.
5+ years of practical experience with AWS, including services such as EKS, EC2, RDS, SQS, S3, OpenSearch and related infrastructure components.
Strong Kubernetes administration and troubleshooting skills, preferably with EKS.
Advanced experience with Terraform, including modules, state management and multi-provider infrastructure.
Hands-on experience with GCP, ideally including BigQuery, Compute Engine, data pipelines or related services.
Strong Linux, networking and security fundamentals.
Experience with production infrastructure ownership, monitoring, alerting and incident response.
English — Upper-Intermediate / B2 or higher, with the ability to communicate with external providers and technical stakeholders.
Fluent Ukrainian for day-to-day internal communication with the team.
Availability to work within Central European Time / Eastern European Time zones.
Nice to have:
Experience with Cloudflare, R2, caching or edge services.
Experience with Hetzner infrastructure.
Experience optimizing OpenSearch / ELK stack performance.
Experience deploying, supporting or monitoring infrastructure for AI/ML workloads in AWS or GCP.
Experience with FinOps or cloud cost optimization.
Previous experience owning DevOps initiatives or mentoring other engineers.
Why It’s Exciting to Work With Us:
Build AI-powered tools with real impact, helping over one million small businesses grow, compete, and launch ads without agencies or complex marketing setups.
Join a strong A-level team of experienced professionals who set a high bar and push each other to do better work.
Remote-first approach: work from anywhere, with access to our offices in Warsaw and Kyiv.
We take care of the essentials. Paid time off, sick leaves, and financial support with private entrepreneurship matters.
Regular performance reviews to ensure your growth is transparent and rewarded.
Recruitment Process:
Recruiter interview → Teсh Interview → Interview with COO / Co-Founder → Reference check → Job offer.
You’ll influence core product decisions, and help build a fast-growing, financially stable AI platform — not as a contributor on the sidelines, but as an owner of critical parts of the product.
Send us your CV — we’d love to get to know you better! 💚
How to Stand Out
- Highlight concrete examples of Terraform modules you built and how they improved consistency across clouds.
- Prepare a brief case study of a Kubernetes incident you owned, focusing on detection, response and post‑mortem actions.
- Demonstrate familiarity with both AWS and GCP by mentioning specific services you managed in each environment.
- Include any FinOps or cost‑optimization results you achieved, even if they are percentages or dollar amounts you can share.
- During interviews, ask about Zeely’s multi‑cloud roadmap to show strategic interest beyond day‑to‑day tasks.
- Negotiate remote‑work allowances early; many remote teams budget for home‑office upgrades.
- Watch for vague on‑call expectations—clarify rotation frequency and escalation procedures before accepting.
This is a remote position listed on WFA Digital, the platform for professionals who work from anywhere. Browse more remote jobs across all categories.