Senior Data Engineer at Simulmedia | AdTechTalent
Home / Jobs / Senior Data Engineer Data Science 1 month ago Remote
Simulmedia
Senior Data Engineer
python sql spark databricks delta lake redshift postgres airflow docker aws etl data pipelines data engineering rest api ci/cd big data data modeling workflow orchestration fastapi flask media advertising
Ready to apply?
Join Simulmedia and work on a role built for experienced AdTech operators.
Key details Location
Kyiv, Ukraine; Lviv, Ukraine
Full job description Simulmedia seeks a senior Data Engineer with 7+ years experience to design, build, and operate large-scale data pipelines and data models. The role involves working with Python, SQL, Spark/Databricks, Delta Lake, Redshift, and Airflow to ingest, validate, and transform multi-billion-row datasets. Responsibilities include migrating pipelines, building REST APIs, optimizing performance and cost, monitoring production, and collaborating cross-functionally. Candidates must have strong skills in Python, expert SQL, distributed data processing, data modeling, workflow orchestration, AWS, Docker, and engineering best practices. Must communicate effectively with U.S. teams and work 11:00 AM - 8:00 PM EEST. Location is mandatory in Ukraine, with offices in Kyiv and Lviv; work is primarily remote with occasional onsite meetings.
What you'll do Design and build batch data pipelines ingesting, validating, and transforming multi-billion-row datasets Model complex real-world data including dimensional, reference, and temporal data Similar jobs
More roles worth a look Related opportunities based on specialty and working model so candidates can keep momentum.
TripleLift
New York, US • 4 months ago
$90K – $120K
data science machine learning python
TripleLift
New York, US • 4 months ago
$170K – $200K
senior software engineer Java Python
Moloco
Quick snapshot
Kyiv, Ukraine; Lviv, Ukraine
Full-time
Develop and operate workloads on lakehouse platform (Databricks / Spark / Delta) and data warehouse (Redshift)
Migrate existing pipelines from warehouse to lakehouse
Orchestrate pipelines with Airflow including scheduling, dependencies, retries, backfills, and alerting
Design parity checks and reconciliation queries to prove correctness
Run large historical backfills and investigate data discrepancies down to row level
Build and maintain Python services and REST APIs serving internal products
Optimize performance and cost via query tuning, table design, workload management, and compute right-sizing
Monitor production pipelines, participate in incident triage and root-cause analysis, and harden systems
Collaborate cross-functionally with product managers, data scientists, and stakeholders
Work within an Agile team releasing new features regularly
Take ownership and experiment with new technologies to improve software Requirements Bachelor's degree in Computer Science, Computer Engineering, relevant technical field, or equivalent practical experience 7+ years of work experience as a data engineer Proficiency in Python as primary development language Expert-level SQL skills for complex analytical queries and debugging Hands-on experience with distributed data processing platforms (Spark/Databricks preferred; EMR, Snowflake, BigQuery also relevant) Experience with columnar data warehouses (Redshift, Snowflake, BigQuery, ClickHouse, etc.) Ability to design complex data models including normalized, dimensional, and temporal models Experience with workflow orchestration tools like Airflow Experience integrating third-party data feeds handling schema drift and data-quality issues Experience building REST services in Python (FastAPI, Flask, etc.) Experience developing, maintaining, and debugging large server-side codebases Working knowledge of AWS services and Docker Good knowledge of engineering best practices and testing (unit, integration, code review, CI/CD) High level of ownership and ability to learn quickly Must be able to communicate effectively with U.S.-based teams Ability to work 11:00 AM — 8:00 PM EEST Experience with Delta Lake / medallion lakehouse architectures is a plus Experience migrating legacy pipelines with strict parity requirements is a plus Experience with advertising, media or measurement industry data is a plus Tech stack Python SQL Spark Databricks Delta Lake Redshift Postgres Airflow Docker AWS (S3, ECS, EMR, RDS, IAM) GitHub Actions Jenkins Grafana Sentry OpenSearch FastAPI Flask
Apply now Ready to take the next step in your career? Click the button below to continue to the application process.
Company
Simulmedia No one knows how to drive growth and business performance through TV advertising better than Simulmedia, a marketer’s most trusted tech partner. In today’s complex, fast-moving TV industry, Simulmedia is your turnkey, one-stop shop to help you make the best decisions and solve your biggest and toughest challenges on TV. Powered by data science, our TV+® full-funnel TV advertising solution lets you intelligently plan and buy linear, streaming, and gaming. With TV+, you’ll have more certainty in achieving your business goals by knowing how to best find and engage your strategic audience on linear and CTV as cost-effectively as possible. In addition, we allow advertisers to extend their reach and connect with elusive younger audiences via PlayerWON™, the first engagement and monetization platform for free-to-play PC and console video games. With over 15 years of experience bringing data science into advanced TV media buying, Simulmedia is the first and only provider to create a patented predictive lookalike model based on years of viewership data to bring smarter media buying to clients like Mass Mutual, Monster, Electrolux, and Choice Hotels.
Industry
Software Development
Website
https://www.simulmedia.com
Posted
1 month ago
Category: Data Science
Seoul, South Korea • 23 months ago
machine learning deep learning GCP