Senior Site Reliability Engineer
Guarda esta oferta y sigue tu búsqueda
Crea una cuenta gratis para guardar empleos, crear alertas y volver a esta oferta desde tu panel.
Overview
As Principal Platform Engineer focused on capacity, you optimize Elastic Cloud Hosted and Serverless workloads to scale seamlessly. You'll partner with control plane and cross-functional teams to tackle cloud scaling challenges and resource allocation at a global scale. Your work directly supports customers by ensuring reliable, efficient compute across many regions. This role offers a meaningful impact on how Elastic delivers scalable AI-powered search and observability.
Compensaciones / Beneficios base salary
health coverage for you and family
flexible locations and schedules
generous vacation days
donations matching up to $2000
volunteering time (up to 40 hours/year) versus parity of benefits across regions
Responsabilidades Assess current and future capacity needs and develop predictive models
Collaborate to implement proactive capacity planning to avoid shortages
Optimize resource usage across cloud environments for performance and scalability
Analyze metrics and build reporting tools for visibility into capacity and performance
Operate an autoscaling framework for diverse customer workloads
Optimize infrastructure performance across 60+ Elastic Cloud regions
Work with development teams to implement scaling best practices
Requisitos principales 5+ years in cloud infrastructure and capacity management
Knowledge of performance monitoring and optimization techniques
Understanding of cloud scaling challenges and solutions
Proficiency in incident investigation and troubleshooting processes
Experience with compute auto-scaling and capacity reservations across the three major CSPs
Solid software and platform engineering background
Experience with all three major cloud service providers and navigating compute capacity scaling issues
problem-solving mindset
cross-functional collaboration
proactive communication
cloud infrastructure management
capacity planning and modeling
performance monitoring and tuning