
About the Role
As a Senior Site Reliability Engineer (Capacity) at Elastic, you will play a crucial role in managing and optimizing compute resources to ensure Elastic Cloud Hosted and Serverless workloads scale seamlessly. You will collaborate with control plane and platform engineering teams to solve complex cloud scaling and resource allocation challenges.
What You Will Be Doing
- Capacity Modeling: Assess current and future requirements to ensure seamless scaling. Develop and maintain accurate models that predict resource needs and align with business objectives.
- Resource Optimization: Implement strategies to optimize compute usage across cloud environments, enhancing performance and scalability.
- Data-Driven Decisions: Analyze capacity metrics and trends to guide resource allocation. Develop reporting tools for clear visibility into capacity and performance.
- Autoscaling Operations: Operate an autoscaling framework for diverse customer workloads. Optimize infrastructure performance across over 60 regions in Elastic Cloud.
What You Bring
- 5+ years of experience with cloud infrastructure and capacity management.
- Proficiency in performance monitoring and optimization techniques.
- Deep understanding of cloud scaling challenges and solutions.
- Experience with compute auto-scaling processes and capacity reservations across major Cloud Service Providers (CSPs).
- Strong background in software and platform engineering.
- Proven ability to navigate compute capacity scaling issues across major cloud providers.
- Proficiency with incident investigation and troubleshooting processes.
Benefits
- Competitive pay based on market benchmarks.
- Flexible locations and schedules.
- Health coverage for you and your family.
- Generous annual vacation days.
- Financial donation matching up to $2,000.
- Up to 40 hours per year for volunteer projects.
- Minimum of 16 weeks of parental leave.
Timezone overlap
UTC+0–+3
Benefits
Open to
Europe
Sign in to track applications and earn points.