Why This Job is Featured on The SaaS Jobs
In SaaS, observability is often the difference between meeting customer expectations and repeatedly firefighting production issues. This Senior CloudOps Engineer role is featured because it centers on designing and owning an end to end observability approach across distributed systems, with a clear emphasis on Kubernetes and multi cloud infrastructure. The listing also signals a mature operational environment through explicit uptime targets and structured incident response practices.
For a long term SaaS career, this kind of mandate builds durable platform engineering skills that translate across product companies: defining standards, setting instrumentation strategy, and turning telemetry into operational decisions. Working in an IaC paradigm across major cloud providers, alongside service mesh and messaging components, creates broad exposure to the reliability concerns that emerge as SaaS architectures expand. The responsibility to propose improvements in alignment with architects and tech leads also reflects a role that blends hands on engineering with system level thinking.
This position fits engineers who prefer ownership over narrowly scoped ticket work and who are comfortable making tradeoffs in production systems. It will suit someone who enjoys debugging complex infrastructure, collaborating across time zones, and participating in on call as part of maintaining service health.
The section above is editorial commentary from The SaaS Jobs, provided to help SaaS professionals understand the role in a broader industry context.
Job Description
About the Role:
As a Cloud Ops Engineer at Wrike, you have advanced skills in supporting cloud and on-prem data center infrastructure with security in mind. You know how to work with containers, networking, monitoring, automation, and debugging a reasonably complex infrastructure. You feel comfortable defining your own work based on the team OKRs. You can also help others do so when necessary. You are used to proposing meaningful improvements to the existing infrastructure in alignment with architects and tech leads, and you can drive the execution
For this position, we are seeking an Observability Hero with a deep background in Infrastructure Engineering to design, build, and own our observability strategy. In this role, you will define the processes, standards, and architecture required to ensure comprehensive visibility across our distributed systems, heavily focused on modern Cloud and Kubernetes environments
Your Impact:
- Designing, building, and owning our end-to-end observability strategy (metrics, logs, and distributed traces) for our distributed cloud infrastructure
- Managing the Wrike product infrastructure
- Implementing reliable solutions to ensure a product uptime SLA of 99.9%
- Working with GCP, AWS and other cloud providers in the IaC paradigm
- Introducing and supporting new infrastructure services
- Actively participating in incident response and management, including on-call duties
- Developing and maintaining professional connections within and outside of the team
Your Qualifications:
- Kafka and rabbitmq for messaging
- Kubernetes and ArgoCD (Service-oriented architecture)
- Nginx, HAproxy and Istio for load balancing
- GCP, AWS and Cloudflare are our main cloud providers
- Puppet, Ansible and Terraform for defining everything as a code
- Python to automate everything
- Prometheus (VictoriaMetrics) and Zabbix for monitoring
- Graylog, Logstash, Fluentd for logging
- Jenkins and Gitlab-CI for building pipelines
- Upper Intermediate English skills
Standout Qualities:
- A suitable candidate must have extensive hands-on experience supporting and optimizing the Observability stack
- Intermediate knowledge of the following areas: Data networks, Security, Databases, Cloud providers, Process automation, Kubernetes
- Advanced experience with any Cloud Provider management using IAC (AWS/GCP/Azure).
- Advanced Linux administration skills with experience in maintaining highly available infrastructure for web application stack
Team Dynamics:
- Your manager will be Dmitrii Vlasov, SysOps Team Lead
- We have two dozen folks in the SysOps Department, consisting of four teams distributed in Prague, Nicosia, Rennes and Tallinn
Our Work Style:
- Scrum based processes (daily standups, biweekly plannings, quarterly plannings)
- Hybrid mode (1-2 days per week)
- On-call shifts in production and development environments
What’s Next?
- Introduction Call with a Recruiter
- Technical interview
- Cultural interview
Why Join Wrike?
- 5 Weeks of paid vacation
- Sick Leave Compensation
- 5 Paid Uncertified Sick Days
- 2 weeks fully paid w/ medical certificate, additional
- 4 weeks paid at 80% salary rate
- Parental Leave (fully paid): 18 Weeks Maternity / 4 Week Paternity
- 2 Volunteer Days
- Meal Vouchers (CZK 220 per working day)
- Annual Prague Travel Card (Lítačka)
- Hybrid Working Model
- Benefit budget with flexible options, including a MultiSport card, Canadian Medical membership, contributions to a pension savings plan and additional choices available through Benefit Plus
Your recruitment buddy will be Alexandra Vorobyova, Lead Recruiter.
#LI-AV1