
About the Role
As a Senior Site Reliability Engineer (Capacity) at Elastic, you will play a crucial role in managing and optimizing compute resources to ensure Elastic Cloud Hosted and Serverless workloads scale seamlessly. You will collaborate with control plane and platform engineering teams to solve complex cloud scaling and resource allocation challenges.
What You Will Be Doing
- Capacity Modeling: Assess current and future requirements to develop accurate models that predict resource needs and align with business objectives.
- Resource Optimization: Implement strategies to optimize compute usage across cloud environments, enhancing performance and scalability.
- Metrics & Reporting: Analyze capacity trends to guide resource allocation decisions and build reporting tools for clear visibility.
- Autoscaling: Operate an autoscaling framework across 60+ Elastic Cloud regions and collaborate with development teams on scaling best practices.
What You Bring
- 5+ years of experience in cloud infrastructure and capacity management.
- Solid software and platform engineering background.
- Proficiency with compute auto-scaling processes and capacity reservations across major Cloud Service Providers (CSPs).
- Strong knowledge of performance monitoring, optimization techniques, and incident troubleshooting.
- Experience navigating complex compute capacity scaling issues.
Benefits
- Competitive pay based on role responsibilities.
- Flexible locations and schedules.
- Health coverage for you and your family.
- Generous annual vacation allowance.
- Financial donation matching (up to $2,000) and 40 hours of volunteer time per year.
- Minimum of 16 weeks of parental leave.
Timezone overlap
UTC+0–+3
Benefits
Open to
Europe
Sign in to track applications and earn points.