Why This Job is Featured on The SaaS Jobs
Observe by Snowflake sits in a distinctly SaaS-native category: observability delivered as a cloud platform, with heavy emphasis on telemetry pipelines, correlation, and long-term analytics. Featuring an Infrastructure Engineer role here highlights how foundational reliability and developer enablement are to product credibility in SaaS, especially when the service itself is used to monitor other production systems. The mention of open formats and a data-lake approach also signals a platform that must balance cost, performance, and operability at scale.
For a SaaS career, this role builds durable experience across the operating layer of a subscription product: cloud infrastructure, CI/CD, security standards, and incident response. These are the mechanics behind uptime, safe releases, and predictable performance, which translate across most B2B SaaS environments. Working close to engineering teams also tends to sharpen judgment around tradeoffs between platform evolution and day-to-day operational demands.
This position suits an engineer who enjoys automation, systems thinking, and owning reliability outcomes end to end. It aligns well with someone who wants hands-on responsibility in AWS, Kubernetes, and Infrastructure-as-Code, and who is comfortable participating in on-call while driving post-incident improvements and longer-term platform resilience.
The section above is editorial commentary from The SaaS Jobs, provided to help SaaS professionals understand the role in a broader industry context.
Job Description
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done.
Observe by Snowflake is an AI-powered observability platform built on the Snowflake Data Cloud and engineered for scale. We ingest and store logs, metrics, traces, and events on an open, scalable data lake using open formats like Apache Iceberg, delivering deep correlation and long-term analytics at dramatically lower cost. A dynamic Knowledge Graph and chat-based AI SRE provide rich context and guided workflows so teams can move from detection to root cause and resolution significantly faster.
The Infrastructure team at Observe by Snowflake is responsible for building, scaling, and operating the development and production environments that power our observability platform. We are a small, highly collaborative team with a broad scope, focused on delivering reliable infrastructure while continuously improving the systems that support our engineers and customers.
What You’ll Do
Design, build, and operate scalable cloud infrastructure in AWS supporting a high-scale observability platform.
Improve system reliability, performance, and operational visibility across development and production environments.
Develop and maintain CI/CD pipelines and internal tooling to improve developer productivity and deployment safety.
Identify and mitigate security risks, and help maintain internal security standards and compliance requirements.
Build infrastructure that supports high availability, scalability, and operational resilience
Participate in an on-call rotation, contributing to incident response and post-incident improvements.
Partner closely with engineering teams to ensure infrastructure supports evolving product and platform needs.
What We’re Looking For
2+ years of experience in Infrastructure Engineering, Site Reliability Engineering (SRE), DevOps, or related roles.
Experience operating container orchestration platforms such as Kubernetes
Hands-on experience managing cloud infrastructure using Infrastructure-as-Code tools such as Terraform, Ansible, or similar.
Strong programming skills in Go, Python, or similar languages, with a focus on automation and systems development.
Experience supporting production systems at scale, with a focus on reliability and operational excellence.
Strong problem-solving skills and the ability to balance short-term operational needs with long-term infrastructure design.
Experience with AWS, GCP, and Azure
Nice to Have
Experience operating large-scale distributed systems.
Familiarity with observability platforms, telemetry pipelines, or monitoring infrastructure.
Experience improving developer platform tooling or internal infrastructure platforms.
Experience working in high-growth or rapidly evolving engineering environments.
Every Snowflake employee is expected to follow the company’s confidentiality and security standards for handling sensitive data. Snowflake employees must abide by the company’s data security plan as an essential part of their duties. It is every employee's duty to keep customer information secure and confidential.
Snowflake is growing fast, and we’re scaling our team to help enable and accelerate our growth. We are looking for people who share our values, challenge ordinary thinking, and push the pace of innovation while building a future for themselves and Snowflake.
How do you want to make your impact?
For jobs located in the United States, please visit the job posting on the Snowflake Careers Site for salary and benefits information: careers.snowflake.com