RVC
JobsFor Employers
9 jobsSort: Relevance
12d 22h ago

SRE Engineering Manager

Scaleway·General · Cloud Computing
🏢Hybrid
|Paris, France
grafanakubernetesprometheusproxmox
2d 21h ago

SRE DevOps Engineer

Lumnix·General · Cloud infrastructure
📡Fully Remote
ansiblebashbgpgrafana+9
3d 19h ago

Staff Site Reliability Engineer

Skydio·General · Aerospace
📡Fully Remote
|Relocation
$240k - $300k USD
awsci/cdgolangkubernetes+2
7d 3h ago

Senior Cloud Architect

LivePerson·General · Conversational AI
📡Fully Remote
$150k - $160k USD
bashgolanggcpinfrastructure_as_code+3
8d 13h ago

Site Reliability Engineer

Yuno·General · Financial Technology
📡Fully Remote
awsdockerevent-driven_architectureinfrastructure_as_code+4
10d 4h ago

Senior Site Reliability Engineer

Filevine·General · Legal Technology
📡Fully Remote
$175k - $195k USD
aiawsbashci/cd+7
11d 4h ago

Director of Engineering

Federato·General · Insurance Technology
📡Fully Remote
$250k - $275k USD
security
11d 10h ago

Senior Site Reliability Engineer

Ping Identity·General · Cybersecurity
📡Fully Remote
automationci/cdfluxgitops+3
13d 7h ago

Senior Site Reliability Engineer

Alpaca·General · Financial Services
📡Fully Remote
gitopsgolangkuberneteslinux+2

SRE Engineering Manager

Scaleway | General | Cloud Computing
Hybrid
Full Time
France, Paris, Lille, Toulouse, Rennes, Rouen, Bordeaux, Lyon
Position:
SRE Engineering Manager - GPU Cloud

Company:
Scaleway

Location:
France, Paris

Employment type:
Full-time (long term)

Work Arrangement:
Hybrid

Short Summary:
Join Scaleway and shape the sovereign cloud of tomorrow! Lead the Site Reliability Engineering (SRE) team to build, automate, and maintain a highly reliable GPU cluster infrastructure.

Responsibilities:
- Lead and manage a team of 6 Site Reliability Engineers, supporting their career growth and technical execution
- Design and implement automated solutions for server lifecycle management across GPU clusters
- Design and implement observability, logging, and monitoring solutions for large-scale GPU clusters
- Plan, prioritize, and manage the technical development roadmap for the SRE team
- Collaborate and coordinate closely with software engineering, product, and cross-functional teams
- Handle recruitment and career management for team members
- Maintain, scale, and optimize high-availability production systems under heavy load
- Participate in on-call rotations to ensure production reliability and fast incident resolution

Requirement:
- Strong experience managing engineering teams in high-constraint production environments
- Proven expertise with Kubernetes container orchestration
- Direct experience with cluster management and virtualization tools (Proxmox, Warewulf)
- Experience with monitoring, metrics, and observability stacks (Prometheus, Grafana)
- Exposure to modern GPU hardware ecosystems (Nvidia, AMD) and high-speed networking fabric (InfiniBand, Spectrum-X, Tomahawk)
- Knowledge of distributed and high-performance storage solutions (Lustre DDN, VAST)
- Strong engineering leadership and team management capabilities
- Technical rigor and high attention to detail in production-critical environments
- Ability to handle high-pressure operational situations and manage incident stress pragmatically
- Excellent communication skills with the ability to convey challenging messages effectively
- Collaborative mindset with a focus on empowering engineers rather than micromanaging

Benefits:
- Hybrid work: Up to 3 days of remote work per week
- Spacious, dynamic workspaces with outdoor spaces and bike parking
- Healthy meal service at headquarters and lunch support for regional sites
- Well-being commitments including gym access and daycare services
- International environment with a diverse workforce
- Opportunities for internal mobility within the Iliad Group

Required Skills

grafanakubernetesprometheusproxmox

Required Languages

🇬🇧 English, 🇫🇷 French

Key competency: General
Roles
PythonJavaReactTypeScriptNode.jsGoRustDevOpsData scienceProductDesign
Remote
United StatesUnited KingdomCanadaGermanyPolandSpainNetherlandsPortugal
Work type
Fully remoteRemote in-countryHybridSeniorMid-levelJuniorAll jobs
R© 2026 ReVacancybuild e3cb608f
AboutContactPrivacyCookiesRefundsTerms & ConditionsFor Employers