AdTechTalent
Engineering1 month agoOn-site

FreeWheel

Site Reliability Engineer, Streaming HUB - FreeWheel

AWSTerraformKubernetesAmazon EKSPythonGoAnsibleDockerInfrastructure as CodeCI/CDNetworkingSite Reliability EngineeringSRECloudAutomation

Key details

Salary

Not specified

Employment type

Full-time

Seniority

Mid-level

Years experience

3-5

Location

Reston, United States

Full job description

FreeWheel, a Comcast company, seeks a Site Reliability Engineer to manage and support the infrastructure of its Streaming Hub platform. Responsibilities include maintaining production infrastructure, designing and implementing infrastructure changes, managing Kubernetes environments (Amazon EKS), ensuring platform reliability and performance, troubleshooting production issues, developing Infrastructure as Code with Terraform, automating processes using Python, Go, Terraform, and Ansible, and collaborating with offshore teams for operational coverage. Required experience includes 2-5 years in cloud infrastructure, Kubernetes, Terraform, Ansible, CI/CD tools like Jenkins, and strong scripting skills in Python and Go. A Bachelor's degree or equivalent experience is preferred. The role is on-site in Reston, Virginia.

What you'll do

  • Manage and maintain business-critical production infrastructure supporting FreeWheel's Streaming Hub platform
  • Design, implement, and execute infrastructure changes for new product launches, client requirements, and engineering initiatives
  • Partner with product development and engineering teams to build and modify infrastructure
  • Build, manage, and optimize Kubernetes environments including Amazon EKS clusters and cloud-native infrastructure
  • Ensure platform availability, reliability, scalability, security, and performance in a 24x7 production environment
  • Participate in code-level troubleshooting, root cause analysis, and incident resolution for complex production issues
  • Develop and maintain Infrastructure as Code solutions using Terraform and automation frameworks
  • Automate operational processes using Python, Go, Terraform, and Ansible
  • Work with offshore operations teams to provide operational coverage and incident response across a 12-hour daily support model
  • Collaborate with engineering teams to improve system observability, monitoring, deployment processes, and production readiness
  • Support high-profile live events and traffic spikes ensuring platform stability during major streaming and advertising events

Requirements

  • Bachelor's Degree or equivalent combination of coursework and experience
  • 2-5 years relevant work experience
  • Strong experience with AWS and/or Oracle Cloud Infrastructure
  • Hands-on experience building and managing cloud infrastructure, networking, and production environments
  • Extensive experience operating and managing Kubernetes clusters in production
  • Strong Terraform experience
  • Experience using Ansible for infrastructure automation and configuration management
  • Experience with CI/CD tooling such as Jenkins
  • Strong scripting and automation experience using Python and Go
  • Solid understanding of network architecture and troubleshooting
  • Experience with load balancing, traffic management, network performance tuning, security and access controls, distributed systems networking

Tech stack

AWSOracle Cloud InfrastructureAmazon EKSKubernetesDockerTerraformAnsibleJenkinsPythonGoInfrastructure as CodeCI/CD

Benefits

Commission eligibility for sales positionsBonus eligibility for most non-sales positionsComprehensive benefits supporting physical, financial, and emotional well-beingPersonalized support options and expert guidance

Apply now

Ready to take the next step in your career? Click the button below to continue to the application process.

Similar jobs

More roles worth a look

Related opportunities based on specialty and working model so candidates can keep momentum.