Site Reliability Engineer

Hace 1 día

Catalonia, España Emburse, Inc. Jornada completa
  • Develop software and software fixes to integrate internal systems. Ensure code quality, test and distribute code updates, and monitor the health and stability of the servers
  • Meet and beat Key Performance Indicators, SLAs, maintain an error budget and adhere to it
  • Identify, evaluate, and execute preventative measures to minimize and avoid impact to the customer experience
  • Employ deep troubleshooting skills to improve the availability, performance, and security for CR and Emburse, ensure services are designed with 24/7 availability and operational readiness and rigor
  • Coding and Automation of Applications on Cloud Platforms
  • Work with Engineering leadership to build shared services that meet the requirements and need of the platform and application teams
  • Work with Cloud Platform and Operations leaders to develop narratives, backlog grooming, epic planning and overall sprint planning processes
  • Ensure the platform holds a high degree of reliability, at least four 9s
  • Define non-functional requirements as part of the product lifecycle to influence the new designs, standards, and methods for scalable, highly available distributed systems
  • Own technically intricate issues that cross between DevOps, Databases, Networking, Code, Infrastructure and people; drive them to satisfactory completion
  • Work closely with product different stakeholders to align Operational priorities and planning with the product and engineering roadmap
  • Prepare and present engineering related documents to key stakeholders
  • Provide recommendations and feedback in review sessions, design reviews and review sessions
  • Mentor SRE I and II's
  • Assist guiding more junior engineers in best practices
  • Conduct and assist with investigation, test and deployment activities, identify and mitigate risks in development activities

Benefits

  • Flexible spending accounts
  • Generous paid time off for every employee and flexible work schedules
  • Paid leave for new parents
  • Volunteering opportunities
  • Health savings accounts (HSA)
  • Medical, dental, vision, disability and life insurance plans
  • Generous match to pre- or post-tax retirement savings accounts
  • Financial planning services
  • Program in partnership with eCornell to grow their knowledge and their careers
  • Delicious supply of snacks and beverages
  • Quarterly outings and Friday get-togethers to build a strong community at work

Experience with full lifecycle of SaaS implementations as well as Infrastructure as codeMinimum of 7 years' experience in an engineering role requiredExperience working with Ansible and Terraform tools hightly desirableStrong scripting skills. OOP is a plusExperience working with an off shore teamStrong Kubernetes experienceExcellent follow-up and project management skillsProven ability to create and maintain new toolsPreferred AWS with optional Azure Cloud experienceBachelor's degree in Computer Science or a STEM field requiredExcellent technical skills. Up to 70% of the job is hands on in a distributed Linux environmentObservability backgroundDeep understanding of infrastructure as code, scripting, self-healing, containers, DevOps tooling, distributed systems higly desiredLiaise between other teams to help prioritize and align prioritiesExcellent written and verbal communication skills, in EnglishExcellent troubleshooting skills