DevOps Engineer
We help the world run better
At SAP, we keep it simple: you bring your best to us, and we'll bring out the best in you. We're builders touching over 20 industries and 80% of global commerce, and we need your unique talents to help shape what's next. The work is challenging – but it matters. You'll find a place where you can be yourself, prioritize your wellbeing, and truly belong. What's in it for you? Constant learning, skill growth, great benefits, and a team that wants you to grow and succeed.
- This is a hybrid role based out of SAP Montreal office, working in-office with the team 3 days per week.
- Candidates must be legally entitled to work in Canada at the time of application. This position is not eligible for employer-sponsored work authorization (e.g., LMIA or other immigration support).
- Act as technical expert during Live site incidents (downtimes of supported services in scope), investigate and solve incidents on a deep technical level.
- Drive root cause analysis and follow-up improvements to prevent issues from reoccurring.
- Perform in-depth troubleshooting and log analysis to identify and solve complex issues in accordance with internal and external SLAs.
- Build software-based solutions to address improvements in service reliability and stability.
- Enhance infrastructure and platform monitoring by gathering system metrics (4 Golden Signals) and implementing tools for recovery.
- Integrate and collaborate closely with development teams and work with them on outputs from Postmortems and product improvements.
- Learn new technologies and keep up to date with latest development increments.
- Create and maintain technical documentation.
- Define, advocate, apply SRE best practices.
- Participate in the on-call rotation (follow the sun approach) to react to major incidents. On-call has a special compensation package.
- Experience with Kubernetes and good understanding of container technologies.
- Understanding of modern cloud architectures (experience with Cloud Platforms such as AWS, Azure, GCP are a plus).
- Experience with Unix/Linux operating system
- Scripting skills, CI/CD (ArgoCD, Concourse, Github Actions and are a plus) - enthusiasm for automation - make the computers do the work for you.
- Experience using AI-assisted engineering tools (e.g., Claude Code CLI, GitHub Copilot, or similar) to improve troubleshooting, automation, documentation, root cause analysis, and operational efficiency.
- 2+ years experience in SRE.
- Working efficiently in emergency situations. Affinity to quickly analyze and solve problems in a global team setup.
- Excellent team player, passionate about his/her work, self-motivated and driven.
- Excellent communication skills - precise, based on facts.
- Fluency in English.
- Preferred Additional Skills and Competencies:
- Coding experience with Python, GO, Bash
- CKA/CKAD/CKS certifications
- Experience with modern monitoring, logging, and alerting tools (Grafana, Prometheus, Kibana, Loki, Splunk On-Call, Dynatrace)
- Security best practices for application development and operations in a public Cloud Environment
- Contribution to open-source projects
AI Usage in the Recruitment Process
For information on the responsible use of AI in our recruitment process, please refer to our Guidelines for Ethical Usage of AI in the Recruiting Process.
Please note that any violation of these guidelines may result in disqualification from the hiring process.
Requisition ID: 458700 | Work Area: Software-Development Operations | Expected Travel: 0 - 10% | Career Status: Professional | Employment Type: Regular Full Time | Additional Locations: #LI-Hybrid
Required Skills
Required Languages
🇬🇧 English