Underdog Sports
7 months ago
Senior Site Reliability Engineer Infrastructure
Sign up to save this job, get alerts, and apply with an optimized CV.
Company information
- Company
- Underdog Sports
- Location
- United States United States
- Posted
- 7 months ago
Job description
At Underdog, we make sports more fun. Our thesis is simple: build the best products and we'll build the biggest company in the space, because there's so much more to be built for sports fans. We're just over five years in, and we're one of the fastest-growing sports companies ever, most recently valued at $1.3B. And it's still the early days. We've built and scaled multiple games and products across fantasy sports, sports betting, and prediction markets, all united in one seamless, simple, easy to use, intuitive and fun app. Underdog isn't for everyone. One of our core values is give a sh*t. The people who win here are the ones who care, push, and perform. If that's you, come join us. Winning as an Underdog is more fun. This is a rare opportunity to be a founding SRE at Underdog, helping define how reliability, scalability, and operational excellence work as the company continues to grow. You'll operate in exploration mode early on, identifying the highest-leverage reliability challenges and shaping our approach to incident response, observability, and SLOs. This is a high-impact role with real ownership from day one, partnering closely with platform, infrastructure, and product teams to ensure Underdog scales through peak traffic, game-day spikes, and rapid iteration while improving both system reliability and developer experience. About the role • Own and maintain the incident response process, including defining procedures, tools, and best practices • Guide teams in establishing and monitoring Service Level Objectives (SLOs), including setting up alerts and reporting systems • Lead capacity planning initiatives, focusing on scalability, reliability, and performance • Collaborate with platform, infrastructure, and product teams to ensure Underdog's systems are reliable, scalable, and performant • Develop and maintain a deep understanding of the company's technical landscape, including cloud providers, databases, and other critical systems • Work closely with the engineering organization to design, implement, and operate highly available and scalable systems • Participate in on-call rotations and be prepared to respond to production issues as needed • Stay up-to-date with industry trends, best practices, and emerging technologies to continuously improve Underdog's reliability and scalability. This is a rare opportunity to be a founding SRE at Underdog, helping define how reliability, scalability, and operational excellence work as the company continues to grow.
Required skills
Interested in this position?
Create your free account and tailor your CV to match this job.