RVC
JobsFor Employers
  1. Jobs
  2. /
  3. DevOps Engineer

DevOps Engineer

SAP | Enterprise software and cloud services
1h 9m ago
Hybrid
Full Time
Canada
97.8k - 166.2k CAD/yr

We help the world run better
At SAP, we keep it simple: you bring your best to us, and we'll bring out the best in you. We're builders touching over 20 industries and 80% of global commerce, and we need your unique talents to help shape what's next. The work is challenging – but it matters. You'll find a place where you can be yourself, prioritize your wellbeing, and truly belong. What's in it for you? Constant learning, skill growth, great benefits, and a team that wants you to grow and succeed. 

Important information:
  • This is a hybrid role based out of SAP Montreal office, working in-office with the team 3 days per week.
  • Candidates must be legally entitled to work in Canada at the time of application. This position is not eligible for employer-sponsored work authorization (e.g., LMIA or other immigration support).
What you'll do We are looking for an engineer to join an already established SRE team for the SAP Business AI Platform. As a Site Reliability Engineer, you will have the opportunity to operate and support business critical Cloud services. As part of your daily job, you will proactively monitor the service behavior and identify areas for improvement. You will participate in the development of tools for monitoring and troubleshooting cloud services built on latest open source and SAP technologies, following SRE principles. Responsibilities:
  • Act as technical expert during Live site incidents (downtimes of supported services in scope), investigate and solve incidents on a deep technical level.
  • Drive root cause analysis and follow-up improvements to prevent issues from reoccurring.
  • Perform in-depth troubleshooting and log analysis to identify and solve complex issues in accordance with internal and external SLAs.
  • Build software-based solutions to address improvements in service reliability and stability.
  • Enhance infrastructure and platform monitoring by gathering system metrics (4 Golden Signals) and implementing tools for recovery.
  • Integrate and collaborate closely with development teams and work with them on outputs from Postmortems and product improvements.
  • Learn new technologies and keep up to date with latest development increments.
  • Create and maintain technical documentation.
  • Define, advocate, apply SRE best practices.
  • Participate in the on-call rotation (follow the sun approach) to react to major incidents. On-call has a special compensation package.
What you bring
  • Experience with Kubernetes and good understanding of container technologies.
  • Understanding of modern cloud architectures (experience with Cloud Platforms such as AWS, Azure, GCP are a plus).
  • Experience with Unix/Linux operating system
  • Scripting skills, CI/CD (ArgoCD, Concourse, Github Actions and are a plus) - enthusiasm for automation - make the computers do the work for you.
  • Experience using AI-assisted engineering tools (e.g., Claude Code CLI, GitHub Copilot, or similar) to improve troubleshooting, automation, documentation, root cause analysis, and operational efficiency.
  • 2+ years experience in SRE.
  • Working efficiently in emergency situations. Affinity to quickly analyze and solve problems in a global team setup.
  • Excellent team player, passionate about his/her work, self-motivated and driven.
  • Excellent communication skills - precise, based on facts.
  • Fluency in English.
  • Preferred Additional Skills and Competencies:
    • Coding experience with Python, GO, Bash
    • CKA/CKAD/CKS certifications
    • Experience with modern monitoring, logging, and alerting tools (Grafana, Prometheus, Kibana, Loki, Splunk On-Call, Dynatrace)
    • Security best practices for application development and operations in a public Cloud Environment
    • Contribution to open-source projects
Meet the team The Reliability Engineering organization provides multitude of products and services related to operations and continuity of business delivery. The Site Reliability Engineering teams make the SAP Business AI Platform run better by providing 24x7 deep technical coverage for Incident Management (Outages and other incidents with major customer impact) applying SRE principles. We share a "Live-Site" First culture and care for the business continuity of our customers running mission critical applications in the Cloud. #LI-GL1

AI Usage in the Recruitment Process

For information on the responsible use of AI in our recruitment process, please refer to our Guidelines for Ethical Usage of AI in the Recruiting Process.

Please note that any violation of these guidelines may result in disqualification from the hiring process.
Requisition ID: 458700  | Work Area: Software-Development Operations  | Expected Travel: 0 - 10%  | Career Status: Professional  | Employment Type: Regular Full Time   | Additional Locations:  #LI-Hybrid
​

Required Skills

kubernetesunix/linuxargocdgithubpythongolangbash

Required Languages

🇬🇧 English

Related searches

  • Jobs in Canada
  • Hybrid Jobs
  • Mid-Level Jobs
1 jobs
Sort by
1h 9m ago

DevOps Engineer

SAP·Enterprise software and cloud services
97.8k - 166.2k CAD/yr
kubernetesunix/linuxargocdgithubpython+2
🏢Hybrid
|Canada
General
No similar jobs match these filters. Try changing or clearing a filter.
Remote roles
PythonJavaReactGoNode.jsC# / .NET
Countries
United StatesUnited KingdomCanadaUS & EMEAGermanyPolandSpainNetherlandsPortugal
Experience
SeniorMid-levelJunior
R© 2026 RVC Globalbuild 4a005c7a
AboutMCPPricingContactPrivacyCookiesRefundsTerms & ConditionsFor Employers