Senior Cloud Platform Engineer
Job
Senior Cloud Platform EngineerDescription
POSITION TITLE: Senior Cloud Platform Engineer
LOCATION: Aachen, Germany – Hybrid
REPORTS TO: Director, AI Engineering & Platform
ABOUT THE ROLE:
Position Overview
Deluxe is seeking a Senior Cloud Platform Engineer to help design, build, automate, and operate cloud and platform infrastructure supporting AI services, automation initiatives, media workflows, internal platforms, and enterprise application environments.
This role will have a primary focus on infrastructure supporting AI platforms, cloud-based model services, high-performance workloads, and automation initiatives, while also contributing to broader Deluxe infrastructure needs across media workflows, internal platforms, and enterprise applications. The Senior Cloud Platform Engineer will collaborate closely with other infrastructure, security, DevOps, application engineering, and AI engineering teams to ensure cloud environments follow Deluxe standards for reliability, scalability, security, observability, and cost management.
The ideal candidate combines strong AWS infrastructure experience with practical DevOps and systems engineering skills. This person should be comfortable working across infrastructure-as-code, Linux systems, containers, CI/CD, monitoring, deployment automation, operational support, and production reliability.
RESPONSIBILITIES:
Design, build, automate, and maintain cloud infrastructure for AI services, media workflows, internal platforms, and enterprise application environments.
Develop and maintain infrastructure-as-code using tools such as Terraform, AWS CDK, CloudFormation, or similar technologies.
Support AWS networking, compute, storage, IAM, security controls, monitoring, environment automation, and cost management.
Build and support scalable infrastructure for production services, batch processing, high-throughput workloads, and GPU-enabled environments where needed.
Create and maintain CI/CD pipelines, deployment automation, configuration management, and operational tooling.
Support Linux systems, Dockerized applications, runtime environments, service deployments, and production release processes.
Partner with other infrastructure, security, application, DevOps, and AI engineering teams to align with Deluxe standards and shared platform practices.
Implement and improve standards for reliability, observability, security, backup/recovery, access management, incident response, and operational readiness.
Support containerized workloads, service deployments, cluster operations, and environment configuration.
Troubleshoot issues across infrastructure, systems, applications, networking, security, and deployment pipelines.
Help define reusable patterns for provisioning, configuration, monitoring, deployment, support, and environment management.
Support operational documentation, architecture diagrams, runbooks, and knowledge transfer materials.
Identify opportunities to reduce manual work through scripting, automation, platform improvements, and standardized tooling.
QUALIFICATIONS:
Strong professional experience in cloud infrastructure, DevOps, systems engineering, site reliability engineering, or platform engineering.
Hands-on experience designing and operating AWS infrastructure in production environments.
Experience with infrastructure-as-code tools such as Terraform, AWS CDK, CloudFormation, or similar technologies.
Strong knowledge of AWS networking, VPCs, IAM, security groups, load balancing, compute, storage, and monitoring services.
Strong Linux administration, scripting, automation, and troubleshooting skills.
Experience with Docker, containerized workloads, CI/CD pipelines, deployment automation, and production release practices.
Experience supporting production systems with monitoring, logging, alerting, incident response, and operational runbooks.
Understanding of cloud security, identity/access management, secrets management, and least-privilege access patterns.
Ability to troubleshoot complex issues across cloud, systems, network, application, and deployment layers.
Strong written and verbal communication skills.
PREFERRED EXPERIENCE:
Experience supporting GPU workloads, AI/ML environments, model-serving platforms, high-performance compute, or data-intensive systems.
Experience with Kubernetes, ECS, EKS, Slurm, HPC clusters, or distributed compute environments.
Experience with Ansible, Helm, GitHub Actions, GitLab CI, Jenkins, or similar tooling.
Experience with observability platforms such as CloudWatch, Prometheus, Grafana, Datadog, ELK/OpenSearch, or OpenTelemetry.
Experience supporting media, localization, content-processing, workflow automation, or high-throughput production platforms.
Familiarity with security hardening, vulnerability management, compliance-driven operations, and enterprise infrastructure standards.
AWS certifications are a plus.
Company
As the world’s leading multidisciplinary service provider, Deluxe underpins the media and entertainment industry, servicing content creators and distributors including Netflix, WarnerMedia, The Walt Disney Company, Amazon, Apple, Viacom, NBCU, Google, AT&T and many others, by providing Global Content Distribution, Localization, Accessibility and Mastering while leading end-to-end innovation with unparalleled scale and agility across the Streaming, Theatrical, Broadcast and Mobile landscapes. With headquarters in Los Angeles and offices around the globe, the company employs over 3,500 of the most talented and experienced industry individuals worldwide. For more information, please visit www.bydeluxe.com
Required Skills
Required Languages
🇬🇧 English