About this role
Internship Site Reliability Engineer - FanDuel (6 months), Hybrid
About Betfair Romania Development :
Betfair Romania Development is the largest technology hub of Flutter Entertainment, with over 2,000 people powering the world’s leading sports betting and iGaming brands. Exciting, immersive and safe experiences are delivered to over 18 million customers worldwide, from our office in Cluj-Napoca. Driven by relentless innovation and commitment to excellence, we operate our own unbeatable portfolio of diverse proprietary brands such as FanDuel, PokerStars, SportsBet, Betfair, Paddy Power, or Sky Betting & Gaming.
Our Values:
The values we share at Betfair Romania Development define what makes us unique as a team. They empower us by giving meaning to our contributions, and they ensure that we consistently strive for excellence in everything we do. We are looking for passionate individuals who align with our values and are committed to making a difference.
Win together | Raise the bar | Got your back | Own it | Positive impact
About FanDuel:
FanDuel is a leading force in the sports-tech entertainment industry, redefining how fans engage with their favourite sports, teams, and leagues. As the premier gaming destination in North America, FanDuel operates across multiple verticals, including sports betting, daily fantasy sports, online gaming, advance-deposit wagering, and media.
Role Overview:
As a Site Reliability Engineering Intern , you'll join the Reliability Engineering team and learn how FanDuel builds, operates, and continuously improves reliable services at scale. You'll work alongside experienced Reliability Engineers and application and platform teams, gaining hands-on exposure to production systems, observability, automation, and core Site Reliability Engineering practices.
This is a hands-on learning role. You'll develop an understanding of how modern distributed systems behave, how engineers use telemetry to understand service health, and how practices such as SLIs and SLOs, incident management, production readiness, and automation help improve reliability.
You'll contribute to real engineering initiatives while receiving guidance and mentoring from experienced engineers. As your knowledge grows, you'll take ownership of well-defined pieces of work, build automation and tooling, and contribute to improvements that help engineering teams operate their services more reliably.
A Sneak Peek Into Our Tech Stack:
- AWS, Kubernetes, Terraform, Helm, Ansible, Vault
- Datadog, OpenTelemetry, PagerDuty
- Buildkite, GitHub
- Bits AI SRE and AI-assisted investigation capabilities
- Locust, AWS Resilience Hub, AWS Fault Injection Service
Key Accountabilities & Responsibilities:
- Learning how FanDuel services and platforms operate and how Reliability Engineering supports application teams.
- Working with Reliability Engineers to understand service architectures, dependencies, and critical customer journeys.
- Using logs, metrics, traces, dashboards, and alerts to understand service behaviour and investigate reliability issues.
- Learning core SRE concepts including SLIs, SLOs, error budgets, incident management, production readiness, and operational toil.
- Supporting the creation and maintenance of dashboards, monitors, SLOs, runbooks, and reliability documentation.
- Contributing to production-readiness and reliability-maturity assessments under the guidance of experienced engineers.
- Supporting incident analysis and learning how production issues are investigated and turned into longer-term engineering improvements.
- Identifying repetitive operational activities and contributing to simple automation that reduces manual work.
- Building scripts and small tools that improve reliability and operational workflows.
- Learning how Kubernetes, AWS, infrastructure as code, and CI/CD are used to operate large-scale production services.
- Working with Observability Engineering to understand how logs, metrics, and traces provide visibility into production systems.
- Supporting Performance and Resilience Engineering activities such as performance tests, Gamedays, and resilience exercises.
- Exploring how AI-assisted investigation and automation can help engineers understand and resolve production problems.
- Contributing to team documentation, knowledge-sharing sessions, and reusable reliability patterns.
- Sharing what you learn and actively participating in technical discussions, pairing sessions, and engineering reviews.
Skills, Capabilities & Experience Required:
- Currently studying, or recently completed studies in, Computer Science or a related technical discipline.
- Basic understanding of software engineering and computer systems.
- Familiarity with at least one programming language such as Python, Java, Go, JavaScript.
- Basic understanding of Linux and networking concepts.
- Familiarity with Git or another source-control system.
- An interest in cloud technologies, distributed systems, infrastructure, DevOps, Site Reliability Engineering, or Platform Engineering.
- Basic understanding of APIs and modern application architectures.
- Advanced level of English Language skills.
- An interest in understanding how production systems behave and how engineers troubleshoot them.
- An interest in automation and reducing repetitive manual work.
- Good analytical and problem-solving skills.
- Willingness to ask questions, experiment, learn from feedback, and work alongside engineers from different technical areas.
- Good communication and collaboration skills.
- A continuous-learning mindset and curiosity about reliability, observability, automation, and large-scale systems.
- Previous experience with AWS, Kubernetes, Terraform, Datadog, CI/CD, or SRE concepts is helpful but not mandatory .
- Demonstrates a strong focus on achieving high-quality outcomes, consistently pushing for greater efficiency and performance
- Demonstrating curiosity and openness toward AI and emerging technologies, with a willingness to continuously learn, adapt, and share knowledge.
Benefits:
- €1,000 per year for self-development
- Company share scheme
- 25 days of annual leave per year
- 20 days per year to work abroad
- 5 personal days/year
- Flexible benefits: travel, sports, hobbies
- Extended health, dental and travel insurances
- Customized well-being programmes
- Career growth sessions
- Thousands of online courses through Udemy
- A variety of engaging office events
Disclaimer:
We are an inclusive employer. By embracing diverse experiences and perspectives, we create a lasting, positive impact for our employees, customers, and the communities we’re part of. You don't have to meet all the requirements listed to apply for this role. If you need any adjustments to make this role work for you, let us know, and we’ll see how we can accommodate them.
We thank all applicants for their interest; however, only the candidates who best meet the job requirements will be contacted for an interview.
By submitting your application online, you agree that your details will be used to progress your application for employment. If your application is successful, your details will be used to administer your personnel record. If your application is unsuccessful, we will retain your details for a period no longer than three years, to consider you for prospective roles within the company.