AI Platform / SRE Lead
Project description
Join Our Team: Innovating Health Care with Cutting-Edge Technology Combine two of the fastest-growing fields on the planet with a culture of performance, collaboration, and opportunity — and this is what you get. We are at the forefront of technology in an industry that is transforming the lives of millions. Here, innovation isn't just about creating another gadget; it's about making health care data accessible whenever and wherever people need it — safely and reliably. If you're passionate about driving change and looking for a place to make an impact, this is the place to be. It's an opportunity to do your life's best work. Project Description: We are looking for a senior AI Platform / SRE Lead to drive AI-enabled operational practices across distributed engineering teams. The role will focus on reliability, automation, observability, production support, and adoption of AI within DevOps and SRE processes.
Responsibilities
- Responsibilities: Lead AI-enabled operations, reliability, automation, observability, and production support initiatives. Identify and guide AI use cases for incident management, monitoring, troubleshooting, and operational automation. Define operational standards, governance, and guardrails. Drive adoption of AI-enabled DevOps and SRE practices. Improve operational efficiency and reliability through automation and AI. Serve as a technical advisor and escalation point across multiple delivery pods. Mandatory Skills: Site Reliability Engineering (SRE) DevOps Generative AI / AI Engineering Observability & Monitoring Incident Management Operational Automation CI/CD
SKILLS
Must have
- Strong experience in SRE, DevOps, production operations, automation, and Forward Deployed Engineering (FDE), with practical knowledge of applying AI/GenAI to operational and client-specific use cases. Expertise in observability, monitoring, incident management, troubleshooting, CI/CD, operational reliability, and deploying solutions directly into complex enterprise environments is required. Must be capable of working closely with client and engineering teams to translate business and operational needs into scalable technical solutions, establishing governance and guardrails, and serving as a senior technical escalation and advisory lead across multiple engineering pods. Experience implementing AI-assisted incident response, intelligent monitoring, automated troubleshooting, AI-enabled DevOps solutions, and FDE-led solution deployment and integration. Experience with cloud platforms, Infrastructure as Code, enterprise AI tools, customer-facing technical delivery, and large-scale healthcare or regulated environments is preferred.
Nice to have
Exceptional communication skills. Ability to deliver exceptional customer service with a positive attitude.
Required Skills
Required Languages
🇬🇧 English