Staff Engineer — Performance, Reliability

Hace 5 horas

Barcelona, Cataluña, España Jobrapido Jornada completa
OverviewUsted podría ser el solicitante perfecto para este trabajo. Lea toda la información asociada y asegúrese de presentar su candidatura.
As Staff Engineer on Factorials DX and Performance team, you will shape how we measure and improve performance and service health across the engineering organization. You’ll partner with product, infrastructure, and DX leaders to raise the bar on observability, load validation, and AI-assisted workflows. The role combines hands-on engineering with technical leadership to drive scalable, reliable systems for 15,000+ customers and 1M+ users. This is a cross-cutting role with broad impact, focused on building robust signals and practices across teams.
Compensaciones / BeneficiosPrivate health insuranceWellhub health and fitness programCobee expense managementLanguage classesBreakfast in office and organic fruitPet friendly officeResponsabilidadesDefine and evolve SLIs and SLOs for critical product journeysStandardize observability, dashboards, and service health visibility across teamsInvestigate bottlenecks across application, database, and system layersDrive improvements in latency, throughput, scalability, and reliabilityBuild more structured load testing workflows for critical pathsHelp teams validate system behavior under realistic traffic and concurrencyAnalyze capacity and behavior under peak load and growthDefine practices and tooling to prevent performance regressions before productionCollaborate with product and infrastructure teams to align on performance prioritiesDesign AI-assisted workflows to support metric interpretation, anomaly analysis, and incident investigationRequisitos principalesStrong hands-on experience improving performance, scalability, and reliability in complex systemsExperience defining or operating SLIs, SLOs, and service health frameworksStrong knowledge of observability practices and tools such as DatadogExperience investigating production bottlenecks xqziphu across application, database, and distributed system layersExperience building or improving load testing, benchmarking, or performance validation workflowsExperience diagnosing tail latency, throughput issues, and performance variability in productionBroad experience with cloud-based production systemsStrong communication skills and cross-team alignmentProactive mindset and strong ownership mentalityStrong communication and technical writingCross-team collaborationProactive ownershipDatadog observabilitySLIs/SLOs and service health frameworksPerformance optimization and load testing