Site Reliability Engineer

3d ago

$150k-$200k / yearest.- AI estimated, actual pay may differ

Remote|Senior|Full-time|Finance
📄Resume
✉️Cover Letter

🛠 Tech Stack

+1

💼 About This Role

You'll keep Alpaca's brokerage platform reliable and observable as it grows, working across cloud infrastructure, Kubernetes, observability stack, and data layer. You'll spend meaningful time leveling up PostgreSQL reliability on the trading-critical path while being a well-rounded SRE. This role offers deep ownership of database reliability in a fast-growing fintech company.

🎯 What You'll Do

  • Operate production day-to-day including oncall and incident response
  • Define and refine SLIs/SLOs and error budgets
  • Strengthen observability across metrics, logs, traces, and alerting
  • Ship infrastructure through code in a GitOps workflow

📋 Requirements

  • 4+ years in SRE, DevOps, or backend engineering with production operations ownership
  • Hands-on experience operating production services on Kubernetes
  • Solid working knowledge of PostgreSQL in production
  • Proficient with Linux at the operator level

✨ Nice to Have

  • Deeper PostgreSQL experience with OLTP clusters and HA/DR
  • Experience with typed SQL access layers in Go (pgx, gorm, sqlc)
  • Production experience with messaging systems at scale (Kafka, RabbitMQ)

🎁 Benefits & Perks

  • 💰 Competitive Salary & Stock Options
  • 🏥 Health Benefits
  • 🖥️ New Hire Home-Office Setup: One-time USD $500
  • 💳 Monthly Stipend: USD $150 per month via Brex Card

📨 Hiring Process

Estimated timeline: 2-4 weeks · AI estimate

  1. 1Recruiter Phone Screen· 30 min
  2. 2Technical Interview (SRE/Database)· 60 min
  3. 3Onsite (Virtual) Final Round· 3 hours

View original posting

Team & contacts

Loading...
0 0 0