Docebo Logo

Docebo

Senior Site Reliability Engineer

Posted 6 Days Ago
Be an Early Applicant
Hybrid
Toronto, ON
Senior level
Hybrid
Toronto, ON
Senior level
Lead incident response and own shared infrastructure reliability. Perform capacity forecasting, performance tuning, failure testing, improve observability/alerting, run retrospectives and ensure remediation, mentor junior SREs, and drive reliability improvements across engineering.
The summary above was generated by AI

Artificial Intelligence. Actual Impact.

At Docebo, we’re using AI to change how people learn at work—and we mean actually change it. We’re an AI-powered learning platform that helps organizations create, deliver, and manage training all in one place. But our real mission goes deeper: we help teams move faster, work smarter, and focus on the work that truly matters. Our platform is built with intelligent, time-saving tools that personalize learning, eliminate busywork, and turn training from a checkbox into a superpower. The result? Better experiences for learners and real results for businesses.

We’re shaping the future of learning with a team that isn’t afraid to challenge the status quo. If you're excited by the idea of using AI to make work-life better for real people–you’ll feel right at home here. And it’s not just what we build, it’s how we show up. At Docebo, our values aren’t just posters on the wall—they guide how we work every day. We call it the Docebo Heart: trust by default, assume positive intent, and create space for different perspectives to thrive.

So… what are you waiting for? Join 900+ Docebians around the world and help us reinvent the way people learn, because learning never stops.

Role Overview

As a Senior Site Reliability Engineer, you'll take a hands-on lead role in high severity incident response while also shaping the underlying infrastructure that supports the platform's scale. The position blends deep technical ownership with cross-team influence, giving you room to fix root causes, not just symptoms, and to raise the reliability bar across the wider engineering org.


What You'll Be Doing
  • Take point on major incidents, coordinating the response across engineering teams and making the real-time calls needed to restore service quickly.

  • Hold ownership over key pieces of the shared infrastructure, working continuously to improve its resilience and ability to scale.

  • Run structured reliability work such as capacity forecasting, performance tuning, and controlled failure testing.

  • Improve how the org detects problems before customers do, by refining metrics, alerting, and overall observability.

  • Run thorough retrospectives after incidents and make sure the resulting action items actually get built, not just logged.

  • Support less experienced SREs through code and design review, pairing, and ongoing feedback, and help spread strong reliability practices across Product and Engineering more broadly.

What You Bring
  • 4 to 8 years working in SRE, DevOps, or production engineering roles within SaaS environments.

  • Solid, practical grounding in Linux, containerized systems, and cloud native infrastructure, including hands-on experience with a major cloud provider and its networking layer.

  • Comfort writing code or scripts to manage infrastructure, with a genuine preference for infrastructure as code over manual configuration.

  • Real experience building and running CI/CD pipelines, observability tooling, and version control workflows at scale.

  • A track record of staying level headed during live incidents, and the ability to explain technical issues clearly to people who aren't engineers.

Our Hybrid Work Philosophy 🤝

Great work can happen anywhere but coming together helps us go further. Our team spends three days a week in the office (Tuesday-Thursday) to collaborate, solve problems, and learn from each other. With flexibility the rest of the week, it’s a balance designed to help everyone do their best work and keep growing.

Our Total Rewards Philosophy 🎉

Our Total Rewards Philosophy centers around three core areas to reward and care for our People:

  • Rewarding Impact: We lead with competitive pay to reward the impact, skills and traits that fuel our success.

  • Fostering Holistic Wellbeing: We care deeply about and invest in the whole person with programs that support our people’s physical, mental, and financial well-being.

  • Empowering Our Talent Culture: We build a culture of trust and empowerment by designing our rewards and benefits with transparency, equity, and flexibility, enabling our people to do their best work and stay for the long haul.

Our Promise to You 😍
  • Financial Wellness: Own a piece of Docebo through our Employee Share Purchase Plan (ESPP) at a 15% discount, plus a competitive compensation package.

  • Your Well-Being, Covered: You’ll get access to health benefits, so you can get the care you need when you need it.

  • Rest, Relax, Repeat: Rest and recharge with paid vacation days, two company-wide Docebo Days, floating holidays for cultural celebrations, and your birthday off!

  • Family First: We provide coverage offering you time with your little one(s) so you can soak up all those precious moments. Fun fact: we had 30 Docebian babies join the family in 2025!

  • Connections That Count: Connect with global communities through our Employee Resource Groups (including PRIDE, DWA, BIDOC, and Green Ambassadors) and company-wide events that keep the fun rolling all year long.

About Docebo 💙

At Docebo, we create seamless, AI-powered learning experiences for over 3,000 customers worldwide. We have successfully achieved two IPOs (TSX: DCBO & NASDAQ: DCBO), been recognized as a top SaaS e-learning solution, and are growing exponentially in the process. We're a global company, with office across North America, EMEA, APAC, and beyond. Our team is guided by five core values—Grow Together Win Together, Build with Our Customer, Clear is Kind, Own Outcomes, Progress Over Perfection—that shape everything we do. If this resonates with you, now is the perfect time to join one of the fastest-growing learning technology companies in the world.

Docebo is an Equal Employment Opportunity employer. We are committed to diversity and inclusion in our workforce. All qualified applicants and employees will receive consideration for employment regardless of their race, colour, religion, sex (including pregnancy, gender identity, and sexual orientation), national origin, citizenship status, age, disability, genetic information, or any other category protected under applicable law.

As a federal contractor, Docebo is committed to the principles of affirmative action and equal employment opportunity for protected veterans and individuals with disabilities. Docebo does not discriminate because of protected veteran status or on the basis of disability, and Docebo takes affirmative action to employ and advance in employment qualified protected veterans and individuals with disabilities.

Any individuals requiring a reasonable accommodation or would like to voluntarily disclose a disability or protected veteran status to assist with their employment application should send an e-mail to [email protected]. The email should also include the position you’re interested in.

Similar Jobs

21 Days Ago
Hybrid
Senior level
Senior level
Artificial Intelligence • Big Data • Enterprise Web • Fintech • Software • Financial Services
Design and maintain CI/CD pipelines and AWS infrastructure using IaC; manage containerized deployments (Docker, ECS/EKS); provide on-call incident triage and post-incident reviews; lead reliability, disaster recovery, and security efforts; implement monitoring and automation (Splunk, CloudWatch, New Relic); write automation scripts in Python/Bash; document runbooks and collaborate with global teams.
Top Skills: Aws CloudwatchAws Ec2Aws EcsAws EksAws IamAws LambdaAws RdsAws Route 53Aws S3Aws SamAws VpcBashCdkClaude CodeCloudFormationDatadogDockerGithub ActionsGithub CopilotHarnessJenkinsLinux/UnixNew RelicPythonServerless FrameworkSplunkTerraform
2 Days Ago
Easy Apply
Remote or Hybrid
Easy Apply
Senior level
Senior level
Artificial Intelligence • Marketing Tech • Software
Hands-on Senior Site Reliability Engineer responsible for improving infrastructure tooling and automation, building and supporting core applications, monitoring capacity and performance, troubleshooting incidents, participating in on-call rotations, and collaborating across teams to scale a multi-cloud, multi-region content serving platform and advance observability and SLO frameworks.
Top Skills: AWSEksGCPGkeGoGrafana AlloyKubernetesLinuxLokiNode.jsPrometheusPythonRubyShell ScriptingTempoTerraformThanos
24 Days Ago
Hybrid
Senior level
Senior level
Healthtech • Software
Own reliability, observability, and security for AI/ML platforms (data processing, workspaces, labeling, model serving). Build IaC and automation, define SLOs/error budgets, run incident response and DR exercises, implement security controls, mentor engineers, and optimize cost, capacity, and operational standards across cloud environments.
Top Skills: Azure Ai (Azure Ml)Blue/Green DeploymentCanary DeploymentCi/CdContainer OrchestrationData Lineage ToolingDatabricksDistributed TracingEncryptionFinopsGitopsKey ManagementKubernetesLoggingObservability (MetricsSecrets ManagementSli/Slo FrameworksTerraformTraces)

What you need to know about the Ottawa Tech Scene

The capital city of Canada and the nation's fourth-largest urban area, Ottawa has proven a rapidly growing global tech hub. With over 1,800 tech companies, many of which are leaders in their sectors, the city's tech talent now makes up more than 13 percent of its total workforce. This growth is driven not only by the big players like UL Solutions and Dropbox, but also by a thriving startup ecosystem, as new businesses emerge to follow in the footsteps of those that came before them.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account