About DraftKings
At DraftKings, AI is becoming an integral part of both our present and future, powering how work gets done today, guiding smarter decisions, and sparking bold ideas. It’s transforming how we enhance customer experiences, streamline operations, and unlock new possibilities. Our teams are energized by innovation and readily embrace emerging technology.
The Role
As a Senior Site Reliability Engineer, you'll build and scale the critical infrastructure behind every product. In this role, you'll take on complex challenges across global data centers, multiple cloud platforms, and on-premise systems—designing automation-first solutions that elevate performance and eliminate operational friction. You'll be trusted to drive stability at scale, influence architectural decisions, and build tools that empower our teams to move fast and deliver reliably.
Responsibilities
- Design, build, and maintain scalable cloud and on-premises infrastructure using Infrastructure as Code tools such as Terraform, Chef, and Ansible.
- Build observability into every layer of the platform by developing monitoring, alerting, and logging solutions that support service level objectives and long-term reliability.
- Develop deployment platforms and internal tooling that enable engineering teams to deliver software efficiently while maintaining governance and operational excellence.
- Define and standardize infrastructure patterns, deployment strategies, and configuration management practices across a rapidly growing engineering organization.
- Mentor engineers, share technical expertise, and help establish infrastructure standards that improve consistency, scalability, and engineering excellence.
- Participate in an on-call rotation, lead incident response efforts, perform root cause analysis, and implement long-term reliability improvements.
- Partner with Security, Networking, and Engineering teams to strengthen reliability, performance, and resilience across the full technology stack.