Skip to main content

Senior Software Engineer - Observability

Snowflake
Hybrid - Bellevue, WA, 3 days/week in-officeUpdated 3d ago
Compensation
$200k–$288k
Published range
Location
Hybrid - Bellevue, WA, 3 days/week in-office
Remote eligibility
Employment
Full-time
Senior
Role family
Engineering
B2B SaaS
Role skills
Apply on jobs.ashbyhq.com
Job actionsApply now
Job actionsApply now

About the job

Snowflake is powering the era of the agentic enterprise, seeking AI-native thinkers. The External Observability Platform team in Bellevue, WA is hiring a Senior Software Engineer to own key components of the AI-native External Observability Platform. This role involves building high-throughput, low-latency infrastructure for ingesting, processing, and serving petabytes of telemetry data with strict SLA guarantees.

Key Responsibilities

  • Design and implement key components of Snowflake’s External Observability platform handling trillions of events per day across multi-cloud deployments (AWS, Azure, GCP).
  • Write scalable, reliable, and testable backend services to process time-series data, high-cardinality metrics, and distributed traces at massive scale.
  • Practice infrastructure-as-code (IaC) using tools like Terraform to deliver self-healing, automated telemetry pipelines.
  • Collaborate with Application Engineering, Security, and Customer Support teams to make systems measurable and resolve cross-organizational dependencies.
  • Serve as an expert troubleshooter for critical system failure modes, conducting post-mortems and building automation.

Minimum Qualifications

  • 7-12 years of professional experience building infrastructure and backend distributed systems at scale using languages such as Go, C++, Java, or Rust.
  • Proven track record of developing hyper-scale distributed systems in public cloud environments (AWS, Azure, or GCP).
  • Deep CS fundamentals (data structures, algorithms, concurrency, storage engines, distributed consensus, networking).
  • Strong experience with infrastructure-as-code tools such as Terraform or Pulumi.
  • Hands-on expertise with time-series databases, distributed tracing frameworks (OpenTelemetry, Jaeger), or log streaming systems (Kafka, Flink, ClickHouse, Prometheus/Thanos).

Preferred Qualifications

  • Massive scale experience (HPC or global installations processing petabytes of telemetry per day).
  • Customer-facing observability tools, APIs, and analytics dashboards.
  • Deep understanding of Linux kernel tuning, eBPF, or low-level network optimization.
  • Multi-cloud expertise across AWS, Azure, and GCP.

Compensation: $200K – $287.5K.

Location: Bellevue, WA (Hybrid: 3 days/week in-office).

Apply via the provided application link.

Skills & tags

What you can verify before applying

Compare the essentials before you leave: pay, remote scope, employment type, source, and the employer apply destination.