Full job description
Senior Site Reliability Engineer role in Platform Engineering & Developer Experience at Criteo. Responsibilities include designing, building, and operating developer platforms, ensuring reliability, scalability, security, and performance of critical systems, developing automation and platform services, improving integrations with CI/CD, evolving containerized and Kubernetes infrastructure, owning reliability lifecycle including monitoring and incident response, driving platform modernization, collaborating with engineering teams, contributing to technical strategy, mentoring peers, and documenting systems. Requirements include a master's degree or equivalent, 5+ years in SRE, platform engineering, DevOps, or software engineering, strong skills in Go, Python, C#, Ruby, understanding of distributed systems and cloud infrastructure, experience with Linux, containers, Kubernetes, CI/CD, infrastructure-as-code, source code management, and passion for automation and reliability. Nice to have experience with Gerrit, Gitlab, developer productivity platforms, and AI coding assistants. Soft skills include excellent communication, ownership mindset, problem-solving, and willingness for on-call rotation. Benefits include hybrid work model, career development, health and wellness support, inclusive team, competitive salary with performance rewards, and potential equity. Location: Bucharest, Romania.
What you'll do
- Design, build, and evolve the platform that powers software development across Criteo
- Operate and improve critical systems including source code management and remote development environments
- Ensure reliability, scalability, security, and performance for services used daily by engineering teams worldwide
- Support the adoption of modern developer workflows and AI-assisted development tools
- Develop platform services, APIs, and integrations that simplify software development and delivery
- Build automation and self-service capabilities that allow engineering teams to move faster while maintaining high operational standards
- Improve integrations with CI/CD and other internal platforms
- Contribute to the evolution of containerized and Kubernetes-based infrastructure
- Own the full reliability lifecycle of the platforms you support: monitoring, incident response, troubleshooting, root cause analysis, and long-term remediation
- Define and implement best practices around observability, automation, capacity planning, and platform security
- Continuously reduce operational toil through engineering and automation
- Drive platform modernization initiatives and technical improvements
- Partner closely with software engineers, ML engineers, data scientists, and platform teams to improve the overall developer experience
- Contribute to technical strategy and architectural decisions
- Share knowledge, mentor peers, and help define engineering best practices
- Document systems and processes to improve platform adoption and operational excellence
Requirements
- Master's degree in computer science or equivalent experience
- 5+ years of experience in SRE, Platform Engineering, DevOps, Software Engineering, or distributed systems
- Strong software development skills in Go, Python, C#, Ruby, or similar languages
- Solid understanding of distributed systems, cloud infrastructure, APIs, scalability, reliability, and software engineering best practices
- Experience building or operating large-scale engineering platforms or infrastructure services
- Power user on AI-assisted development practices
- Strong experience with Linux server environments
- Hands-on experience with containers and orchestration technologies such as Kubernetes
- Familiarity with CI/CD, infrastructure-as-code, and automation practices
- Experience operating developer platforms, source code management systems
- Passion for automation, observability, reliability, and reducing operational complexity
- Experience with Gerrit or Gitlab administration and operations (nice to have)
- Familiarity with developer productivity platforms, remote development environments, or AI coding assistants (nice to have)
- Experience supporting large engineering organizations (nice to have)
- Excellent communication and collaboration skills
- Strong ownership mindset and problem-solving abilities
- Passion for improving developer productivity and engineering efficiency
- Comfortable operating critical platforms at scale
- Willingness to participate in the team's on-call rotation
Tech stack
GoPythonC#RubyLinuxKubernetesContainersCI/CDInfrastructure-as-codeGerritGitlabAI-assisted development tools
Benefits
Hybrid work model blending home and in-office experiencesLearning, mentorship & career development programsHealth benefits, wellness perks & mental health supportDiverse, inclusive, and globally connected teamAttractive salary with performance-based rewards and family-friendly policiesPotential for equity depending on role and levelAdditional benefits vary depending on country and employment nature