Medal.tv · New York, United States
Join Medal, a fast-growing startup in the sports and entertainment industry. As a Senior/Staff Software Engineer focused on Site Reliability & Infrastructure, you will be responsible for ensuring the reliability and scalability of our infrastructure. You will own the on-call rotation, drive postmortems, and work directly with engineering teams to meet their infrastructure needs. The ideal candidate has experience in startups, strong fluency in Terraform, deep experience scaling and sharding relational databases, and a strong understanding of GCP and Elastic Search. Communication skills are crucial, as you will need to flag issues clearly during incidents and lead actionable postmortems.
- Great judgment: You know the difference between a durable, sustainable fix and a patch that buys you a week
- Experience at startups: You are comfortable in an environment of rapid growth where scaling up is a priority
- Infrastructure-as-code: Strong fluency in Terraform, with real experience owning infrastructure-as-code at scale
- Database scaling: Deep, hands-on experience scaling and sharding relational databases (MySQL, Postgres) in production
- GCP depth: Kubernetes, VPC, IAM, Cloud Logging, and the managed services ecosystem
- Elastic search depth: Hands-on experience running ES for user-facing features, not just as a log sink
- Communication (crucial!): You flag issues clearly and rapidly during incidents and lead/write actionable postmortems
- CI/CD: You've worked with GitHub Actions in a production environment
- Incident response instincts: You can work a P0 calmly, communicate clearly under pressure, and run a postmortem that prevents recurrence
Inicia sesión para generar una carta de presentación para esta vacante.
Iniciar sesiónInicia sesión para ver cómo encaja este empleo con tu perfil.
Iniciar sesión