Why This Job is Featured on The SaaS Jobs
This Senior CloudOps Engineer role stands out in SaaS because it is anchored in observability, reliability, and multi cloud operations, the operational backbone of any subscription product with always on expectations. The mandate spans logging pipelines, monitoring, incident response, and uptime targets, indicating work that sits close to production outcomes rather than internal tooling. With infrastructure spanning Kubernetes, service oriented components, and both cloud and on prem contexts, it reflects the hybrid reality many mature SaaS platforms still run.
From a SaaS career perspective, this kind of remit builds durable expertise in operating customer facing systems at scale: designing telemetry that supports rapid diagnosis, improving fault tolerance, and automating repeatable infrastructure changes through IaC. Experience with mature stacks like Prometheus, Graylog, Terraform, and GitLab CI also transfers well across SaaS companies where platform reliability and cost control depend on strong operational instrumentation.
The role is best suited to engineers who prefer ownership, can translate OKRs into concrete infrastructure work, and are comfortable collaborating with architects and tech leads on incremental improvements. It will appeal to those who like structured delivery rhythms and can participate in on call rotations while steadily raising the quality of production signals and response practices.
The section above is editorial commentary from The SaaS Jobs, provided to help SaaS professionals understand the role in a broader industry context.
Job Description
About the Role:
As a Cloud Ops Engineer at Wrike, you have advanced skills in supporting cloud and on-prem data center infrastructure with security in mind. You know how to work with containers, networking, monitoring, automation, and debugging a reasonably complex infrastructure. You feel comfortable defining your own work based on the team OKRs. You can also help others do so when necessary. You are used to proposing meaningful improvements to the existing infrastructure in alignment with architects and tech leads, and you can drive the execution
For this position, we are seeking an Observability Hero who will be responsible for the fault tolerance, scalability, and performance of logging pipelines and monitoring
Your Impact:
- Managing the Wrike product infrastructure
- Implementing reliable solutions to ensure a product uptime SLA of 99.9%
- Working with GCP, AWS and other cloud providers in the IaC paradigm
- Introducing and supporting new infrastructure services
- Actively participating in incident response and management, including on-call duties
- Developing and maintaining professional connections within and outside of the team
Your Qualifications:
- PostgreSQL as DB platform
- Kafka and rabbitmq for messaging
- Kubernetes and ArgoCD (Service-oriented architecture)
- Nginx, HAproxy and Istio for load balancing
- GCP, AWS and Cloudflare are our main cloud providers
- Puppet, Ansible and Terraform for defining everything as a code
- Python to automate everything
- Prometheus (VictoriaMetrics) and Zabbix for monitoring
- Graylog, Logstash, Fluentd for logging
- Jenkins and Gitlab-CI for building pipelines
- Upper Intermediate English skills
Standout Qualities:
- A suitable candidate must have extensive hands-on experience supporting and optimizing the Observability stack
- Intermediate knowledge of the following areas: Data networks, Security, Databases, Cloud providers, Process automation, Kubernetes
- Advanced experience with any Cloud Provider management using IAC (AWS/GCP/Azure).
- Advanced Linux administration skills with experience in maintaining highly available infrastructure for web application stack
Team Dynamics:
- Your manager will be Dmitrii Vlasov, SysOps Team Lead
- We have two dozen folks in the SysOps Department, consisting of four teams distributed in Prague, Nicosia, Rennes and Tallinn
Our Work Style:
- Scrum based processes (daily standups, biweekly plannings, quarterly plannings)
- Hybrid mode (1-2 days per week)
- On-call shifts in production and development environments
What’s Next?
- Introduction Call with a Recruiter
- Technical interview
- Cultural interview
Why Join Wrike?
- 30 days of paid vacation + seniority leave + Recovery Days (RTT)
- Parental Leave: 16 Weeks Maternity / 4 Weeks Paternity + Daycare Spot
- Health Insurance (Employees + Dependents)
- Life Insurance Plan
- Home Allowance (Hybride 20€/month or Full remote 50€/month)
- Public transportation: Coverage of 75% of subscription fee / Sustainable mobility package 200€/year
- CSE Specific Benefits i.e. vacation vouchers,negotiated fees etc.
Your recruitment buddy will be Alexandra Vorobyova, Lead Recruiter.
#LI-AV1