Senior Site Reliability Engineer

12 ore fa

biassono, lombardia, Italia Docebo Tempo pieno
Overview

As Senior Site Reliability Engineer II at Docebo, you safeguard reliability across key product domains powering AI-driven learning experiences used by millions. You partner with engineering to set SLIs/SLOs, drive long-term reliability roadmaps, and lead incident responses during critical outages. You influence design and deployment patterns to balance feature velocity with rock-solid reliability and mentor engineers to elevate best practices. This role offers impact at scale within a fast-growing learning-technology company that values collaboration and innovation. Join a mission-driven team that prioritizes trust, diverse perspectives, and continuous improvement.

Retribuzione / Benefits
  • health benefits
  • paid vacation days
  • two Docebo Days
  • floating holidays
  • birthday off
  • Employee Share Purchase Plan (ESPP) at 15% discount
Responsabilità
  • Partner with engineering teams to set reliability objectives (SLIs/SLOs) and build roadmaps for core services
  • Lead cross-service incident responses during outages and translate chaos into durable solutions
  • Drive adoption of reliability patterns, deployment safety, and scalability across the company
  • Build and refine observability tools to catch issues before production
  • Influence product design and system architecture to balance feature velocity with reliability
  • Mentor senior engineers, elevate best practices, and contribute to recruiting
Requisiti fondamentali
  • 6–10 years of SRE/DevOps/production engineering experience with distributed systems
  • Cloud architecture and container orchestration expertise (AWS preferred) and Infrastructure as Code
  • Proficient in at least one major programming language and automation of repetitive tasks
  • Proven track record in large-scale incident management and reliability programs
  • Strong cross-functional communication across engineering, product, and customer-facing teams
  • cross-functional communication
  • leadership and mentorship
  • problem-solving under pressure
  • AWS
  • container orchestration
  • Infrastructure as Code