Site Reliability Engineer

Tradebyte · E-commerce

Job description

Position:
Site Reliability Engineer (SRE) - Tradebyte (All Genders)

Company:
Tradebyte

Location:
Germany - Ansbach

Employment type:
Full time

Short Summary:
The team serves as the bridge between Customer Operations, Software Development, and Infrastructure teams, ensuring stability, high-performance, and scalability across our global application landscape.

Responsibilities:
- Monitoring, Alerting & Observability: Implementing tools and dashboards to track system health and business processes in real-time.
- Incident & Problem Management (L3-Support): Providing expert-level troubleshooting to resolve complex technical issues.
- Performance Optimisation & Load Testing: Optimizing system performance and scalability.
- Security Management & Vulnerability Scanning: Identifying and mitigating security vulnerabilities.
- Compliance & Audit Readiness (ISO 27001, SOC2, GDPR): Ensuring operations meet legal and industry standards.

Requirement:
- Proficiency in implementing and managing dashboards and alerting frameworks (e.g., Prometheus, Grafana, or New Relic).
- Experience in Cloud-native environments (AWS) and maintaining CI/CD pipelines.
- Familiarity with Business Continuity, Disaster Recovery, and load/performance testing.
- A curious mindset and a collaborative approach.

Benefits:
- Employee shares program.
- 40% off fashion and beauty products sold and shipped by Zalando.
- Hybrid working model with occasional office attendance.
- 27 days of vacation a year for full-time employees.
- Health and wellbeing options, including mental health support.
- Relocation assistance available (subject to prior agreement).

Skills

  • aws
  • ci/cd
  • grafana
  • infrastructure_as_code
  • prometheus

Languages

EN, DE

Apply

Open this job in our interactive board to apply, save it, or sign up for matched alerts on similar roles.

View & apply
Site Reliability Engineer — Tradebyte | RVC