Senior Site Reliability Engineer

Alpaca · Financial Services

Job description

Position:
Senior Site Reliability Engineer

Company:
Alpaca

Location:
Remote - Americas

Employment type:
Not Specified

Short Summary:
Alpaca is a global leader in agent-first brokerage infrastructure, serving hundreds of financial institutions across 40 countries. We are looking for a Site Reliability Engineer to maintain our brokerage platform's reliability and observability, with a focus on PostgreSQL.

Responsibilities:
- Operate production day-to-day, including on-call, incident response, and postmortems.
- Own reliability practice by defining and refining SLIs/SLOs and error budgets.
- Strengthen observability across metrics, logs, traces, and alerting.
- Ship infrastructure through code in a GitOps workflow.
- Manage PostgreSQL performance tuning, schema and migration review, and online migrations.
- Mentor engineers on reliability and database fundamentals.

Requirement:
- 4+ years in SRE, DevOps, or backend engineering with production operations ownership.
- Hands-on experience with Kubernetes and GitOps workflows.
- Solid knowledge of PostgreSQL in production.
- Understanding of cloud networking fundamentals and debugging cross-service connectivity.
- Proficient with Linux and modern observability stacks.
- Experience in incident response and structured debugging.
- Proficiency in Go or Python, with strong communication skills.
- Genuine interest in databases and PostgreSQL expertise.

Benefits:
- Competitive Salary & Stock Options
- Health Benefits
- New Hire Home-Office Setup: One-time USD $500
- Monthly Stipend: USD $150 per month via a Brex Card

Skills

  • gitops
  • golang
  • kubernetes
  • linux
  • postgresql
  • python

Languages

EN

Apply

Open this job in our interactive board to apply, save it, or sign up for matched alerts on similar roles.

View & apply
Senior Site Reliability Engineer — Alpaca | RVC