AdTechTalent
Data Science9 days agoHybrid

Epsilon

Lead, Data Scientist - Decision Sciences Identity Team

lead research scientistmachine learningdata sciencepythonsqlscalasparkhadoopdatabricksawsllmsgenerative aiidentity resolutiondistributed computingcommunity detectionunsupervised learningnlpinformation retrievalmathematical optimizationtime-series analysiscausal inference

Key details

Salary

Not specified

Employment type

Full-time

Seniority

Lead

Years experience

3-5

Location

Boston, United States; Chicago, United States; Irving, United States; Wakefield, United States; Westminster, United States

Full job description

Epsilon seeks a PhD-level Lead Research Scientist to advance the Identity Resolution Platform for real-time digital marketing. Responsibilities include contributing to data science and machine learning R&D projects, conducting research, developing scalable identity solutions, implementing and optimizing algorithms in distributed environments, and collaborating with engineering and product teams. Qualifications include a Ph.D. in a quantitative field, research experience in machine learning, strong Python programming skills, 2+ years programming experience, and ability to deliver research projects. Preferred skills include experience with graphs, unsupervised clustering, large datasets, distributed computing, cloud platforms, SQL, Scala, Spark, LLMs, generative AI, and advanced domains such as NLP and causal inference. Benefits include flexible time off, paid holidays, sick time, parental leave, childcare and elder care assistance, adoption assistance, health coverage, 401(k), tuition assistance, commuter benefits, professional development, and more. Locations include Chicago, IL; Wakefield, MA; Irving, TX; Westminster, CO; and Boston, MA.

What you'll do

  • Individually contribute to data science and machine learning R&D projects
  • Use data science, machine learning, and/or computer science skills to conduct research and contribute to solutions to technology and business problems
  • Contribute to projects from early-stage research through development
  • Implement and optimize innovative algorithms in distributed environments
  • Develop an understanding of Epsilon personalization platform and proprietary datasets
  • Participate fully in collaborative research and applications projects

Requirements

  • Ph.D. in Computer Science, Electrical Engineering, Statistics, Mathematics, Economics, Physics, Chemistry, Operations Research, or related quantitative field
  • Research experience and coursework in Machine Learning
  • Strong understanding of statistical, analytical, and predictive modeling techniques
  • Strong Python programming skills and experience with data structures, algorithms, and production-quality code development
  • 2+ years of relevant programming experience through academic research, internships, or industry work
  • Ability to code independently without AI as well as effectively leverage AI-assisted development tools
  • Proven ability to design, execute, and deliver research projects
  • Excellent communication and interpersonal skills, including ability to translate complex technical concepts for non-technical audiences
  • Self-motivated, highly organized, and collaborative, with the ability to manage multiple priorities in highly collaborative, team-oriented environments
  • Experience with graphs/networks and unsupervised clustering algorithms (community detection on graphs) - preferred
  • Experience working with large, complex datasets in research or production environments - preferred
  • Experience with distributed computing and large-scale data processing (Spark, Hadoop, Databricks, cloud databases) - preferred
  • Experience with cloud platforms and data ecosystems (AWS, S3, Databricks) - preferred
  • Proficiency in SQL, Scala, Spark libraries, and advanced database querying - preferred
  • Familiarity with LLMs, Generative AI, and agentic frameworks - preferred
  • Experience in advanced domains such as Community Detection on Graphs, Unsupervised Learning, NLP, Information Retrieval, Mathematical Optimization, Control Theory, Time-Series Analysis, or Causal Inference - preferred
  • Ability to partner with business and technical collaborators to deploy algorithms into production platforms - preferred
  • Experience translating research innovations into scalable, customer-facing products - preferred

Tech stack

PythonSQLScalaSparkHadoopDatabricksAWSS3LLMsGenerative AIAgentic frameworks

Benefits

Flexible time off (FTO)15 paid holidaysPaid sick timeParental/new child leaveChildcare & elder care assistanceAdoption assistanceComprehensive health coverage401(k)Tuition assistanceCommuter benefitsProfessional developmentEmployee recognitionCharitable donation matchingHealth coaching and counseling

Apply now

Ready to take the next step in your career? Click the button below to continue to the application process.

Similar jobs

More roles worth a look

Related opportunities based on specialty and working model so candidates can keep momentum.