Senior Site Reliability Engineer
As a Senior Site Reliability Engineer, you'll build and scale the critical infrastructure behind every product. In this role, you'll take on complex challenges across global data centres, multiple cloud platforms, and on-premise systems-designing automation-first solutions that elevate performance and eliminate operational friction.
You'll be trusted to drive stability at scale, influence architectural decisions, and build tools that empower our teams to move fast and deliver reliably. This is where your impact won't just be felt, it'll be foundational.
- Design, build, and maintain scalable cloud and on-premises infrastructure using Infrastructure as Code tools such as Terraform, Chef, and Ansible.
- Build observability into every layer of the platform by developing monitoring, alerting, and logging solutions that support service level objectives and long-term reliability.
- Develop deployment platforms and internal tooling that enable engineering teams to deliver software efficiently while maintaining governance and operational excellence.
- Define and standardize infrastructure patterns, deployment strategies, and configuration management practices across a rapidly growing engineering organization.
- Mentor engineers, share technical expertise, and help establish infrastructure standards that improve consistency, scalability, and engineering excellence.
- Participate in an on-call rotation, lead incident response efforts, perform root cause analysis, and implement long-term reliability improvements.
- Partner with Security, Networking, and Engineering teams to strengthen reliability, performance, and resilience across the full technology stack.
- At least 3 years of experience designing, building, and operating production infrastructure across cloud and on-premises environments (required).
- Expertise with Infrastructure as Code and configuration management tools, including Terraform, Chef, and Ansible (required).
- Strong programming skills in Python, Go, Ruby, or a similar language, with experience supporting applications built using object-oriented programming languages (required).
- Experience managing distributed systems, containerized workloads, and Kubernetes platforms such as Amazon Elastic Kubernetes Service (EKS) or Rancher (required).
- Strong understanding of networking fundamentals and web technologies, including Domain Name System (DNS), Content Delivery Networks (CDNs), load balancing, Virtual Private Clouds (VPCs), hybrid networking, and packet-level troubleshooting (required).
- Hands-on experience implementing and supporting API gateways, reverse proxies, and modern infrastructure networking components (required).
- A collaborative, action-oriented mindset with a passion for building scalable, reliable platforms and continuously improving engineering practices (required).
- Comprehensive health benefits, including various medical plans, dental, and vision.
- Wellbeing program including free therapy sessions with Lyra Mental Health Solution, Lyra Employee Assistance Program, the Calm App, and Virtual Yoga Classes.
- 14 weeks of 100% paid parental leave to all global team members and workplace lactation support.
- Family planning support for adoption, surrogacy, fertility treatments, or other family planning needs.
- 25 vacation days per year.
- 200 BGN food vouchers per month.
- 50% of your monthly sports card covered in a wide network of sports facilities.
- Financial planning programmes through Origin Financial.
- Commuter benefits to help you save money on your daily commute.
DraftKings Inc. is a leading American digital sports entertainment and gaming company, headquartered in Boston, Massachusetts. Founded in 2012, it started in daily fantasy sports and has grown into one of the largest US online sportsbook and iGaming operators. The company offers mobile sports betting, online casino and daily fantasy contests across regulated North American markets. DraftKings is listed on the Nasdaq stock exchange and employs several thousand people.
