Skip to main content
Posted 26 August, 2026

Senior Site Reliability Engineer (Capacity) - Platform Infrastructure

Elastic
Ireland Full Time
Reference: 102_699059_8155560

What is The Role
As a Principal Platform Engineer focused on capacity, you will play a crucial role in managing and optimizing our compute resources, ensuring that our Elastic Cloud Hosted and Serverless workloads can scale seamlessly. You'll collaborate closely with our control plane and cross-functional platform engineering teams, addressing the real-world challenges of cloud scaling and resource allocation. Our diverse team spans EMEA and NASA, and together we tackle the complexities of building scalable systems, making a meaningful impact on how we serve our customers.
What You Will Be Doing
  • Assess current and future capacity requirements based on workload demands to ensure seamless scaling of resources. Develop and maintain accurate capacity models that predict resource needs and align with business objectives. Collaborate with teams to implement proactive measures that prevent capacity shortages and bottlenecks.
  • Implement effective strategies for optimizing resource usage across our cloud environments. Ensure that compute resources are utilized efficiently to enhance performance and support seamless scalability.
  • Analyze capacity metrics and trends to guide effective resource allocation decisions. Develop insightful reporting tools that provide clear visibility into capacity and performance, helping to optimize our compute resources.
  • Operate an autoscaling framework that accommodates various customer workloads seamlessly. Optimize infrastructure performance across over 60 regions in Elastic Cloud. Collaborate with development teams to implement scaling best practices effectively.
What You Bring
  • 5+ years with cloud infrastructure and capacity management
  • Knowledge of performance monitoring and optimization techniques
  • Understanding of cloud scaling challenges and solutions
  • Proficiency with incident investigation and troubleshooting processes
  • experience with compute auto-scaling processes and capacity reservations across the three major CSPs
  • solid software and platform engineering background
  • worked with the three major cloud service providers and navigated compute capacity scaling issues

Sign up for Job Alerts