Skip to main content

Senior Site Reliability Engineer (Capacity) - Platform Infrastructure

Elastic
Remote - SpainUpdated 11d ago
Base salary
€76k–€102k
Published base salary range
Location
Remote - Spain
Remote eligibility
Employment
Full-time
Senior
Role family
Infrastructure
B2B SaaS
Apply on jobs.elastic.co
Job actionsApply now
Job actionsApply now

About the job

About the Role

As a Principal Platform Engineer focused on capacity, you will play a crucial role in managing and optimizing our compute resources, ensuring that our Elastic Cloud Hosted and Serverless workloads can scale seamlessly. You’ll collaborate closely with our control plane and cross-functional platform engineering teams, addressing the real-world challenges of cloud scaling and resource allocation. Our diverse team spans EMEA and NASA, and together we tackle the complexities of building scalable systems, making a meaningful impact on how we serve our customers.

What You Will Be Doing

  • Assess current and future capacity requirements based on workload demands to ensure seamless scaling of resources.
  • Develop and maintain accurate capacity models that predict resource needs and align with business objectives.
  • Collaborate with teams to implement proactive measures that prevent capacity shortages and bottlenecks.
  • Implement effective strategies for optimizing resource usage across our cloud environments.
  • Ensure that compute resources are utilized efficiently to enhance performance and support seamless scalability.
  • Analyze capacity metrics and trends to guide effective resource allocation decisions.
  • Develop insightful reporting tools that provide clear visibility into capacity and performance, helping to optimize our compute resources.
  • Operate an autoscaling framework that accommodates various customer workloads seamlessly.
  • Optimize infrastructure performance across over 60 regions in Elastic Cloud.
  • Collaborate with development teams to implement scaling best practices effectively.

What You Bring

  • 5+ years with cloud infrastructure and capacity management
  • Knowledge of performance monitoring and optimization techniques
  • Understanding of cloud scaling challenges and solutions
  • Proficiency with incident investigation and troubleshooting processes
  • Experience with compute auto-scaling processes and capacity reservations across the three major CSPs
  • Solid software and platform engineering background
  • Worked with the three major cloud service providers and navigated compute capacity scaling issues

Compensation

Compensation for this role is in the form of base salary. This role does not have a variable compensation component. The typical starting salary range for this role is: €76.000—€101.800 EUR.

Benefits

  • Competitive pay based on the work you do here and not your previous salary
  • Health coverage for you and your family in many locations
  • Ability to craft your calendar with flexible locations and schedules for many roles
  • Generous number of vacation days each year
  • Increase your impact - We match up to $2000 (or local currency equivalent) for financial donations and service
  • Up to 40 hours each year to use toward volunteer projects you love
  • Embracing parenthood with minimum of 16 weeks of parental leave

Application Instructions

Apply via the provided job link.

Skills & tags

What you can verify before applying

Compare the essentials before you leave: pay, remote scope, employment type, source, and the employer apply destination.