
About the Role
As a Senior Site Reliability Engineer (Capacity) at Elastic, you will play a crucial role in managing and optimizing compute resources to ensure our Elastic Cloud Hosted and Serverless workloads scale seamlessly. You will collaborate with control plane and platform engineering teams to solve complex cloud scaling and resource allocation challenges.
What You Will Be Doing
- Capacity Modeling: Assess current and future requirements to develop accurate models that predict resource needs and align with business objectives.
- Resource Optimization: Implement strategies to optimize compute usage across cloud environments, enhancing performance and scalability.
- Data-Driven Decisions: Analyze capacity metrics and trends to guide resource allocation and build reporting tools for better visibility.
- Autoscaling Operations: Operate and optimize an autoscaling framework across 60+ Elastic Cloud regions, collaborating with development teams on best practices.
What You Bring
- 5+ years of experience in cloud infrastructure and capacity management.
- Proficiency with compute auto-scaling processes and capacity reservations across major Cloud Service Providers (CSPs).
- Strong background in platform engineering and software development.
- Expertise in performance monitoring, optimization techniques, and incident troubleshooting.
- Experience navigating complex compute capacity scaling issues across multiple cloud environments.
Benefits & Perks
- Competitive pay based on role responsibilities.
- Health coverage for you and your family.
- Flexible work schedules and locations.
- Generous annual vacation allowance.
- Financial donation matching (up to $2,000) and 40 hours of volunteer time off per year.
- Minimum of 16 weeks of parental leave.
Timezone overlap
UTC+0–+3
Benefits
Open to
Europe
Sign in to track applications and earn points.