Why This Job is Featured on The SaaS Jobs
Modern SaaS products run on distributed, containerised infrastructure where customer experience is tightly linked to reliability. This Senior CloudOps Engineer focus on observability sits at that intersection, covering metrics, logs, and traces across Kubernetes and multi cloud environments. The remit suggests a platform maturity where instrumentation, standards, and shared visibility are treated as first class system design concerns rather than afterthoughts.
For a SaaS career, ownership of an end to end observability strategy builds durable leverage. The work touches incident response, uptime targets, infrastructure as code, and the practical realities of operating messaging, load balancing, and CI pipelines in production. Experience gained here transfers well across SaaS companies because the same patterns recur: defining service health, reducing mean time to recovery, and turning operational signals into engineering priorities.
This role tends to suit engineers who like combining hands on infrastructure work with setting conventions that other teams adopt. It will appeal to professionals comfortable navigating ambiguity via OKRs, collaborating with architects and tech leads, and participating in on call as part of operating a live SaaS platform. It also fits someone who enjoys deep technical ownership in a non remote, multi site setup.
The section above is editorial commentary from The SaaS Jobs, provided to help SaaS professionals understand the role in a broader industry context.
Job Description
About the Role:
As a Cloud Ops Engineer at Wrike, you have advanced skills in supporting cloud and on-prem data center infrastructure with security in mind. You know how to work with containers, networking, monitoring, automation, and debugging a reasonably complex infrastructure. You feel comfortable defining your own work based on the team OKRs. You can also help others do so when necessary. You are used to proposing meaningful improvements to the existing infrastructure in alignment with architects and tech leads, and you can drive the execution
For this position, we are seeking an Observability Hero with a deep background in Infrastructure Engineering to design, build, and own our observability strategy. In this role, you will define the processes, standards, and architecture required to ensure comprehensive visibility across our distributed systems, heavily focused on modern Cloud and Kubernetes environments
Your Impact:
- Designing, building, and owning our end-to-end observability strategy (metrics, logs, and distributed traces) for our distributed cloud infrastructure
- Managing the Wrike product infrastructure
- Implementing reliable solutions to ensure a product uptime SLA of 99.9%
- Working with GCP, AWS and other cloud providers in the IaC paradigm
- Introducing and supporting new infrastructure services
- Actively participating in incident response and management, including on-call duties
- Developing and maintaining professional connections within and outside of the team
Your Qualifications:
- Kafka and rabbitmq for messaging
- Kubernetes and ArgoCD (Service-oriented architecture)
- Nginx, HAproxy and Istio for load balancing
- GCP, AWS and Cloudflare are our main cloud providers
- Puppet, Ansible and Terraform for defining everything as a code
- Python to automate everything
- Prometheus (VictoriaMetrics) and Zabbix for monitoring
- Graylog, Logstash, Fluentd for logging
- Jenkins and Gitlab-CI for building pipelines
- Upper Intermediate English skills
Standout Qualities:
- A suitable candidate must have extensive hands-on experience supporting and optimizing the Observability stack
- Intermediate knowledge of the following areas: Data networks, Security, Databases, Cloud providers, Process automation, Kubernetes
- Advanced experience with any Cloud Provider management using IAC (AWS/GCP/Azure).
- Advanced Linux administration skills with experience in maintaining highly available infrastructure for web application stack
Team Dynamics:
- Your manager will be Dmitrii Vlasov, SysOps Team Lead
- We have two dozen folks in the SysOps Department, consisting of four teams distributed in Prague, Nicosia, Rennes and Tallinn
Our Work Style:
- Scrum based processes (daily standups, biweekly plannings, quarterly plannings)
- Hybrid mode (1-2 days per week)
- On-call shifts in production and development environments
What’s Next?
- Introduction Call with a Recruiter
- Technical interview
- Cultural interview
Benefits & Perks:
- 25 calendar days of paid vacation
- Sick Leave Compensation (5 Paid Uncertified Sick Days)
- Parental Leave: 18 Weeks Maternity / 4 Week Paternity
- 2 Volunteer Days
- Medical Insurance (Employees + Dependents)
- Hybrid Working Model
- School Allowance (Up to €600/month for school aged kids)
- Simcard w/ Unlimited Internet Access for active employees
- Office Lunch Allowance (via Wolt) on Wednesdays / Thursdays
Your recruitment buddy will be Alexandra Vorobyova, Lead Recruiter.
#LI-AV1