Skip to main content

Senior Site Reliability Engineer

Reddit
Hybrid - San Francisco, CAUpdated 7h ago
Base salary
$191k–$267k
Published base salary range
Location
Hybrid - San Francisco, CA
Remote eligibility
Employment
Full-time
Senior
Role family
Infrastructure
Media / Creator
Role skills
Apply on job-boards.greenhouse.io
Job actionsApply now
Job actionsApply now

About the job

About Reddit

Reddit is a community of communities, built on shared interests, passion, and trust. It is home to the most open and authentic conversations on the internet, with 100,000+ active communities and approximately 130 million daily active unique visitors.

About the Role

The Ads organization powers Reddit's advertising platform, enabling advertisers to reach engaged communities while helping Reddit grow its business. The Ads Reliability team partners with Ads Engineering to improve reliability, scalability, operational excellence, and developer productivity across Reddit's advertising ecosystem. We are looking for a Senior Site Reliability Engineer to build, operate, and scale the critical systems behind Reddit Ads.

What You'll Do

  • Partner with Ads Engineering teams to improve reliability, scalability, and operational excellence of ad-serving, auction, targeting, measurement, and billing systems.
  • Design, build, and maintain infrastructure, tooling, and automation that improve service reliability and engineering productivity.
  • Improve observability through monitoring, alerting, tracing, logging, and dashboards.
  • Participate in on-call rotations and lead incident response efforts for critical production systems.
  • Run root cause analysis and drive corrective actions following incidents.
  • Collaborate with software engineers throughout the service lifecycle, from design reviews through production operations.
  • Drive adoption of SRE best practices including SLIs, SLOs, error budgets, capacity planning, and operational readiness reviews.
  • Reduce operational toil through automation and self-service tooling.
  • Help define and measure advertiser-critical user journeys such as campaign creation, ad delivery, reporting, and billing.
  • Scale Ads systems to support continued traffic growth, increased advertiser demand, and evolving business requirements.

What We're Looking For

  • 5+ years of experience in Site Reliability Engineering, Infrastructure Engineering, or related roles operating large scale distributed systems.
  • Strong experience supporting high traffic, user facing production environments.
  • Strong cross-functional collaboration skills to lead and influence operational excellence.
  • Good understanding of modern distributed systems, scale engineering, and cloud-native architectures.
  • Strong software engineering skills in languages like general-purpose backend languages like Go.
  • Demonstrated ability to troubleshoot complex issues across applications, infrastructure, networking, and services.
  • Experience with observability platforms, monitoring systems, alerting, and incident response.
  • Experience driving automation and operational improvements.

Benefits

  • Comprehensive Health benefits
  • 401k Matching
  • Workspace benefits for your home office
  • Personal & Professional development funds
  • Family Planning Support
  • Flexible Vacation & Reddit Global Days Off
  • 4+ months paid Parental Leave
  • Paid Volunteer time off

Pay Transparency

The base salary range for this position is $190,800—$267,100 USD. In addition to base salary, this job is eligible to receive equity in the form of restricted stock units, and depending on the position offered, it may also be eligible to receive a commission. Reddit offers a wide range of benefits to U.S.-based employees, including medical, dental, and vision insurance, 401(k) program with employer match, generous time off for vacation, and parental leave.

Skills & tags

What you can verify before applying

Compare the essentials before you leave: pay, remote scope, employment type, source, and the employer apply destination.