AdTechTalent
Engineering4 days agoHybrid

Taboola

Senior Site Reliability Engineer (SRE)

site reliability engineeringSRELinuxKubernetesDockerTerraformAnsibleGoPythonRustPrometheusGrafanaELKIaCcloudinfrastructuremonitoringCDNnetworking

Key details

Salary

Not specified

Employment type

Full-time

Seniority

Senior

Years experience

5-10

Location

Tel Aviv, Israel

Full job description

Senior Site Reliability Engineer role in Tel Aviv office focused on building, scaling, and maintaining large-scale infrastructure across on-premise, public cloud, and AI/ML Kubernetes environments. Requires 7+ years experience with Linux system internals, network protocols, edge/CDN services, Infrastructure as Code tools, Kubernetes, Docker, and programming in Go, Python, or Rust. Responsibilities include maintaining hybrid infrastructure availability and performance, building automation tooling, troubleshooting full stack issues, designing monitoring and alerting systems, and participating in on-call rotations and incident management. Benefits include health coverage, stocked kitchen, location perks, and hybrid work schedule with flexibility.

What you'll do

  • Maintain hybrid infrastructure (on-prem, public cloud and AI/ML clusters) to be highly available, performant and cost-efficient
  • Build internal software tooling and manage IaC pipelines in Go, Python or Rust to eliminate repetitive operations
  • Perform deep-dive troubleshooting across the full stack from CDN edge configurations to Linux kernel tuning and network layer bottlenecks
  • Design and maintain monitoring and alerting setups to detect and address system health issues proactively
  • Participate in on-call rotations, lead incident resolution and conduct blameless post-mortems to ensure system resilience

Requirements

  • 7+ years of experience managing, scaling and troubleshooting large-scale distributed Linux environments in production
  • Deep understanding of Linux system internals and network protocols (TCP/IP, DNS, HTTP, gRPC)
  • Hands-on experience of edge/CDN services such as Fastly, Cloudflare, Akamai or CloudFront
  • Hands-on experience with Infrastructure as Code (IaC) and orchestration tools such as Terraform, Ansible, Puppet, ArgoCD or Jenkins
  • Production experience managing containerized environments using Kubernetes and Docker
  • Solid programming skills in at least one modern language (Go, Python or Rust)

Tech stack

LinuxTCP/IPDNSHTTPgRPCFastlyCloudflareAkamaiCloudFrontTerraformAnsiblePuppetArgoCDJenkinsKubernetesDockerGoPythonRustPrometheusGrafanaELK

Benefits

Comprehensive health benefitsFully stocked kitchenLocation-specific perks such as gym partnerships and parkingHybrid work schedule with 3 days in-office and flexibility

Apply now

Ready to take the next step in your career? Click the button below to continue to the application process.

Similar jobs

More roles worth a look

Related opportunities based on specialty and working model so candidates can keep momentum.