Remote
Senior Site Reliability Engineer
About this role
Artificial Intelligence. Actual Impact. At Docebo, we’re using AI to change how people learn at work—and we mean actually change it. We’re an AI-powered learning platform that helps organizations create, deliver, and manage training all in one place. But our real mission goes deeper: we help teams move faster, work smarter, and focus on the work that truly matters. Our platform is built with intelligent, time-saving tools that personalize learning, eliminate busywork, and turn training from a checkbox into a superpower.
The result? Better experiences for learners and real results for businesses. We’re shaping the future of learning with a team that isn’t afraid to challenge the status quo. If you're excited by the idea of using AI to make work-life better for real people–you’ll feel right at home here. And it’s not just what we build, it’s how we show up. At Docebo, our values aren’t just posters on the wall—they guide how we work every day.
We call it the Docebo Heart: trust by default, assume positive intent, and create space for different perspectives to thrive. So… what are you waiting for? Join 900+ Docebians around the world and help us reinvent the way people learn, because learning never stops. 🚀 THE ADVENTURE AHEAD As our Senior Site Reliability Engineer II, you’ll step up as the ultimate guardian of reliability across our critical product domains and services.
You will architect high-performing platforms, turn complex systems into resilient superpowers, and collaborate directly with engineering leaders to set bulletproof standards. By keeping our systems seamlessly scalable, you directly power the AI-driven learning experiences used by millions around the globe. ⚡ THE DAY-TO-DAY - Master the Domain: Partner with engineering teams to set crystal-clear reliability objectives (SLIs/SLOs) and build long-term reliability roadmaps for core services.
- Tackle High-Stakes Incidents: Lead cross-service incident responses during critical outages, guiding teams with confidence and translating chaos into clear, durable solutions. - Architect Resilient Futures: Drive the company-wide adoption of cutting-edge reliability patterns, deployment safety capabilities, and infrastructure scalability. - Elevate Observability: Build and refine advanced observability tools to catch issues before they reach production.