Ciklum Logo

Ciklum

Senior DevOps Engineer

Posted 4 Hours Ago
Remote
Hiring Remotely in Canada
Senior level
Remote
Hiring Remotely in Canada
Senior level
Design, build, and operate highly available Azure and AKS infrastructure using Terraform. Develop secure CI/CD pipelines, manage Kubernetes clusters, implement monitoring and observability, troubleshoot complex infrastructure issues, and maintain multi-region disaster recovery. The role also involves cloud security, secrets management, networking, database resilience, infrastructure automation, and collaboration with application engineering teams to establish operational standards.
The summary above was generated by AI

Ciklum is looking for a Senior DevOps Engineer to join our team in Canada.

We are a custom product engineering company that supports both multinational organizations and scaling startups to solve their most complex business challenges. With a global team of over 4,000 highly skilled developers, consultants, analysts and product owners, we engineer technology that redefines industries and shapes the way people live.

About the role:

As a Senior DevOps Engineer, you'll become a part of a cross-functional development team engineering experience of tomorrow. We are looking for a talented Senior DevOps Engineer to support our engineering work, with a strong focus on the underlying infrastructure that we need to get from source code to production, plus keep it happy once it's there. We need someone who understands software development and enjoys working with developers on all the things necessary to improving, deploying, monitoring, and operating production services.

Responsibilities:

  • Build complex, highly available, and cost-optimized cloud infrastructure solutions on Azure and in Kubernetes (AKS)
  • Design, build, and maintain production-grade AKS clusters, including private cluster networking, node pool sizing and autoscaling, workload identity, ingress, certificate management, and cluster upgrade lifecycle
  • Own infrastructure as code end to end in Terraform: author and refactor reusable modules, manage per-environment workspaces and state, and drive changes through pull request, plan review, and apply
  • Design and implement advanced CI/CD pipelines ensuring quality, security, and efficiency, including container image build/publish and progressive promotion from lower environments to production
  • Establish comprehensive monitoring and observability strategies across cluster, application, and Azure platform services
  • Proactively identify and mitigate risks, leading troubleshooting efforts for complex issues across compute, networking, identity, and data layers
  • Implement high-level security practices, integrating vulnerability scanning and threat modeling, and manage secrets through a centralized secrets store with identity-based (rather than credential-based) access
  • Maintain and exercise multi-region high availability and disaster recovery, including failover and failback procedures for Kubernetes workloads and managed databases
  • Work with application engineering to cultivate operational standards

Requirements:

We know that sometimes, you can't tick every box. We would still love to hear from you if you think you're a good fit!

  • Bachelor's degree in Computer Science, Software Engineering, or a related field, or equivalent work experience
  • 5+ years of experience in a DevOps engineering role with significant Azure experience
  • Proven track record of designing and implementing large-scale Azure solutions
  • Hands-on experience running AKS in production, not just in development or proof-of-concept environments — including private clusters, cluster upgrades, autoscaling, and incident response on live workloads
  • Advanced, production-level Terraform experience: module design, state and workspace management, and multi-environment/multi-subscription deployments
  • Demonstrated ability to solve complex technical problems and architect robust solutions
  • Excellent communication and collaboration skills
  • Azure certifications (e.g., Azure Solutions Architect Expert, Azure DevOps Engineer Expert, Certified Kubernetes Administrator)

Desirable:

  • Cloud: Deep knowledge of Azure services, architectures, and design patterns, including Entra ID, RBAC, managed identities and workload identity federation, Key Vault, and subscription/landing zone structure
  • Infrastructure as Code (IaC): Advanced skills in Terraform, or other IaC tools
  • CI/CD: Strong experience with pipeline-as-code (Azure Pipelines, GitHub Actions, or equivalent), service connections and federated pipeline authentication, and GitOps-style continuous delivery to Kubernetes (Argo CD, Flux, or similar)
  • Programming and Scripting: Strong scripting skills (PowerShell, Bash, Python) and proficiency with one or more programming languages (e.g., Go, C#/.NET, Node.js)
  • Containerization and Microservices: Extensive experience with Docker, Kubernetes (AKS), Helm, container registries, and microservices architectures
  • Complex Configuration Management: Expertise in configuration management tools (Ansible, Chef, Puppet, etc.)
  • Robust Networking: In-depth knowledge of networking concepts, Azure virtual network and subnet design, private endpoints and private DNS zones, hybrid connectivity (site-to-site VPN, network virtual appliances/firewalls), and network security
  • Data and Messaging Services: Familiarity with Azure managed data services such as SQL Managed Instance (including failover groups), Cosmos DB, Redis/managed cache, Storage, and Functions
  • Observability: Experience with Azure Monitor/Log Analytics and KQL, plus at least one third-party observability platform (Datadog, Grafana, or similar)
  • Resilience: Experience designing and testing multi-region redundancy, RTO/RPO-driven DR runbooks, and failover automation
  • Security Focus: Strong understanding of cloud security principles, penetration testing, and compliance standards (e.g., SOC 2, ISO 27001, PCI DSS)
  • Multi-cloud: ScriptCycle's footprint spans both Azure and Google Cloud, with an active workload migration track. Exposure to GCP and GKE (including Autopilot) and to cross-cloud networking and identity is a strong plus, though Azure is the primary focus of this role; experience supporting a GCP migration

What’s in it for you?

  • Strong community: Work alongside top professionals in a friendly, open-door environment
  • Growth focus: Take on large-scale projects with a global impact and expand your expertise
  • Tailored learning: Boost your skills with internal events (meetups, conferences, workshops), Udemy access, language courses, and company-paid certifications
  • Endless opportunities: Explore diverse domains through internal mobility, finding the best fit to gain hands-on experience with cutting-edge technologies
  • Care: Healthcare, Basic Life Insurance, Short and Long-term disability insurance according to the Company’s Benefit Plans

About us:

At Ciklum, we are always exploring innovations, empowering each other to achieve more, and engineering solutions that matter. With us, you’ll work with cutting-edge technologies, contribute to impactful projects, and be part of a One Team culture that values collaboration and progress. Now expanding across Canada, we’re looking for talented professionals to strengthen our North American footprint. Join us to innovate at scale and deliver world-class solutions to global clients.

Explore, empower, engineer with Ciklum!

Interested already? We would love to get to know you! Submit your application. We can’t wait to see you at Ciklum.

#LI-IK1

Similar Jobs

5 Days Ago
Remote
Canada
Senior level
Senior level
Artificial Intelligence • Information Technology • Professional Services • Software • Analytics • Generative AI • Big Data Analytics
Owns end-to-end DevOps, cloud infrastructure, CI/CD, Kubernetes, Infrastructure as Code, monitoring, security, reliability, and production operations. The role designs automated deployment processes, manages cloud environments, troubleshoots complex infrastructure and networking issues, implements observability and disaster recovery, optimizes cloud costs, and guides engineering teams on secure, scalable DevOps practices.
Top Skills: AnsibleApi GatewaysAWSAzureAzure DevopsBashCcpaDatadogDnsDockerElkFirewallsGCPGdprGitGithub ActionsGitlab CiGrafanaHelmJenkinsKubernetesLinuxLoad BalancersPowershellPrometheusPythonService MeshSplunkSsl/TlsTerraform
5 Hours Ago
Remote or Hybrid
Canada
Senior level
Senior level
Fintech • Financial Services • Cryptocurrency • NFT • Web3
Designs, develops, and maintains scalable cloud infrastructure and DevOps platforms. Responsibilities include monitoring, architecture, infrastructure as code, Terraform and Kubernetes management, CI/CD pipelines, disaster recovery, backup solutions, and operational automation. The role also requires troubleshooting, stakeholder communication, cross-functional collaboration, and maintaining highly available systems across global teams.
Top Skills: Argo CdAWSCi/CdFluent BitGitGitopsGrafanaKubernetesLinuxLokiPrometheusTerraform
4 Days Ago
Remote
Canada
Senior level
Senior level
Edtech • Software
Designs, operates, and improves AWS and Azure cloud infrastructure supporting customer-facing higher education products. Owns Kubernetes platforms, infrastructure as code, deployment automation, CI/CD, monitoring, security, reliability, disaster recovery, and cost optimization. Troubleshoots complex infrastructure and application issues, leads incidents, participates in on-call rotations, mentors engineers, contributes to technical standards, and partners across Development and Information Security on modernization and production readiness.
Top Skills: Amazon EcsAmazon EksAmazon MskAnsibleArgo CdAWSDnsFargateGithub ActionsGitlab Ci/CdGoHelmHttp/TlsJava/JvmKafkaKeycloakKubernetesLinuxAzureOpenid ConnectOpensearchPowershellPythonSAMLShell ScriptingSite-To-Site VpnSsh TunnelsTerraformWindows

What you need to know about the Ottawa Tech Scene

The capital city of Canada and the nation's fourth-largest urban area, Ottawa has proven a rapidly growing global tech hub. With over 1,800 tech companies, many of which are leaders in their sectors, the city's tech talent now makes up more than 13 percent of its total workforce. This growth is driven not only by the big players like UL Solutions and Dropbox, but also by a thriving startup ecosystem, as new businesses emerge to follow in the footsteps of those that came before them.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account