Senior Site Reliability Engineer

6 ore fa

biassono, lombardia, Italia Docebo Tempo pieno
Overview

As Senior Site Reliability Engineer II at Docebo, you safeguard reliability across key product domains powering AI-driven learning experiences used by millions. You partner with engineering to set SLIs/SLOs, drive long-term reliability roadmaps, and lead incident responses during critical outages. You influence design and deployment patterns to balance feature velocity with rock-solid reliability and mentor engineers to elevate best practices. This role offers impact at scale within a fast-growing learning-technology company that values collaboration and innovation. Join a mission-driven team that prioritizes trust, diverse perspectives, and continuous improvement.

Retribuzione / Benefits
  • health benefits
  • paid vacation days
  • two Docebo Days
  • floating holidays
  • birthday off
  • Employee Share Purchase Plan (ESPP) at 15% discount
Responsabilità
  • Partner with engineering teams to set reliability objectives (SLIs/SLOs) and build roadmaps for core services
  • Lead cross-service incident responses during outages and translate chaos into durable solutions
  • Drive adoption of reliability patterns, deployment safety, and scalability across the company
  • Build and refine observability tools to catch issues before production
  • Influence product design and system architecture to balance feature velocity with reliability
  • Mentor senior engineers, elevate best practices, and contribute to recruiting
Requisiti fondamentali
  • 6–10 years of SRE/DevOps/production engineering experience with distributed systems
  • Cloud architecture and container orchestration expertise (AWS preferred) and Infrastructure as Code
  • Proficient in at least one major programming language and automation of repetitive tasks
  • Proven track record in large-scale incident management and reliability programs
  • Strong cross-functional communication across engineering, product, and customer-facing teams
  • cross-functional communication
  • leadership and mentorship
  • problem-solving under pressure
  • AWS
  • container orchestration
  • Infrastructure as Code