Senior Software Engineer - Observability
Snowflake- Compensation
- $200k–$288k Published range
- Location
- Hybrid - Bellevue, WA, 3 days/week in-office Remote eligibility
- Employment
- Full-time Senior
Role skills
Apply on jobs.ashbyhq.com
About the job
Snowflake is powering the era of the agentic enterprise, seeking AI-native thinkers. The External Observability Platform team in Bellevue, WA is hiring a Senior Software Engineer to own key components of the AI-native External Observability Platform. This role involves building high-throughput, low-latency infrastructure for ingesting, processing, and serving petabytes of telemetry data with strict SLA guarantees.
Key Responsibilities
- Design and implement key components of Snowflake’s External Observability platform handling trillions of events per day across multi-cloud deployments (AWS, Azure, GCP).
- Write scalable, reliable, and testable backend services to process time-series data, high-cardinality metrics, and distributed traces at massive scale.
- Practice infrastructure-as-code (IaC) using tools like Terraform to deliver self-healing, automated telemetry pipelines.
- Collaborate with Application Engineering, Security, and Customer Support teams to make systems measurable and resolve cross-organizational dependencies.
- Serve as an expert troubleshooter for critical system failure modes, conducting post-mortems and building automation.
Minimum Qualifications
- 7-12 years of professional experience building infrastructure and backend distributed systems at scale using languages such as Go, C++, Java, or Rust.
- Proven track record of developing hyper-scale distributed systems in public cloud environments (AWS, Azure, or GCP).
- Deep CS fundamentals (data structures, algorithms, concurrency, storage engines, distributed consensus, networking).
- Strong experience with infrastructure-as-code tools such as Terraform or Pulumi.
- Hands-on expertise with time-series databases, distributed tracing frameworks (OpenTelemetry, Jaeger), or log streaming systems (Kafka, Flink, ClickHouse, Prometheus/Thanos).
Preferred Qualifications
- Massive scale experience (HPC or global installations processing petabytes of telemetry per day).
- Customer-facing observability tools, APIs, and analytics dashboards.
- Deep understanding of Linux kernel tuning, eBPF, or low-level network optimization.
- Multi-cloud expertise across AWS, Azure, and GCP.
Compensation: $200K – $287.5K.
Location: Bellevue, WA (Hybrid: 3 days/week in-office).
Apply via the provided application link.
Skills & tags
What you can verify before applying
Compare the essentials before you leave: pay, remote scope, employment type, source, and the employer apply destination.