AdTechTalent
Engineering4 days agoRemote

VideoAmp

Principal Database Infrastructure Engineer

rustc++godistributed systemsdatabase infrastructuredata platformquery engineapache datafusionapache icebergkubernetescloudperformance engineeringparquetarrowsql optimizerdistributed sqlgrpcprogrammingengineering

Key details

Salary

Not specified

Employment type

Full-time

Seniority

Senior

Years experience

5-10

Location

Remote, United States

Full job description

Principal Database Infrastructure Engineer role at VideoAmp. Remote position in the United States. Responsible for designing and executing scalable, production-critical distributed database systems. Key areas include distributed query execution, worker coordination, query optimizer internals, storage and caching, and performance engineering. Requires 8+ years software engineering experience in database infrastructure, distributed systems, or data platform engineering. Strong programming skills in Rust or C++/Go with async runtimes. Experience with distributed data systems, columnar formats (Arrow, Parquet), distributed systems fundamentals, and performance engineering. Preferred experience with Apache DataFusion, Apache Iceberg, Kubernetes, cloud infrastructure, and SQL query optimizers. Compensation includes $190,000 to $220,000 base salary, equity, comprehensive benefits, flexible PTO, and family leave.

What you'll do

  • Design and implement the physical plan distribution pass, including network shuffle, coalesce, and partition isolator insertion
  • Own the plan serialization codec that ships sub-plans to workers, and maintain the S3 and Flight result exchange paths
  • Own worker discovery and heartbeating, and lead development of the next generation of load balancing: a work-stealing protocol and a replication-aware hash ring
  • Work across cost-based join reordering, cross-stage bloom filter cascade, scan deduplication, and selectivity estimation
  • Upstream changes to DataFusion fork where appropriate and maintain delta where not
  • Own the NVMe LRU cache over S3, Parquet read strategies including full-file and range reads, and Iceberg partition pruning and snapshot handling
  • Close the performance gap on queries compared to Snowflake, focusing on multi-shuffle plans and redistribution after scalar-subquery extraction
  • Use TPC-H benchmarks as performance scorecard

Requirements

  • 8+ years of software engineering experience with significant depth in database infrastructure, distributed systems, or data platform engineering
  • Strong systems programming in Rust, or deep C++ or Go experience with a clear path to Rust, including async runtimes such as Tokio and concurrent data structures
  • Proven experience building or significantly modifying a distributed data system such as a query engine, stream processor, distributed database, or large-scale data pipeline, with a solid understanding of shuffles, partitioning, and network and memory bottlenecks
  • Fluency with columnar formats and vectorized execution, including Arrow, Parquet, and the mechanics behind their performance characteristics
  • Strong grounding in distributed systems fundamentals: consistent hashing, leader and heartbeat protocols, backpressure, partial failure, and graceful degradation
  • A performance engineering mindset: you profile before optimizing and defend changes with real benchmark numbers
  • Direct experience with Apache DataFusion or another SQL query planner or optimizer such as Spark Catalyst, Calcite, Trino, ClickHouse, or DuckDB (strongly preferred)
  • Experience with Apache Iceberg or a comparable open table format such as Delta Lake or Hudi (strongly preferred)
  • Familiarity with Kubernetes and cloud infrastructure including EKS, S3, and IRSA, particularly on ARM or Graviton (strongly preferred)
  • Query optimizer experience including join ordering, predicate pushdown, and cardinality or selectivity estimation (strongly preferred)
  • Experience with a distributed SQL store such as CockroachDB, Flight SQL or gRPC, or contributions to open source data infrastructure projects (nice to have)
  • Experience building or working with developer tooling in agentic or programmatic data access contexts (nice to have)

Tech stack

RustC++GoTokioApache DataFusionSpark CatalystCalciteTrinoClickHouseDuckDBApache IcebergDelta LakeHudiKubernetesEKSS3IRSAARMGravitonCockroachDBFlight SQLgRPCParquetArrow

Benefits

Base salary $190,000 to $220,000 (commensurate with experience)Equity participation includedDiscretionary & flexible PTO plus Spring, Summer & Winter company breaksInclusive and comprehensive medical, dental & vision401(k) with matchingHSA & FSAPaid Maternity & Parental Leave for all family additionsCell phone & wifi reimbursementCommuter benefits

Apply now

Ready to take the next step in your career? Click the button below to continue to the application process.

Similar jobs

More roles worth a look

Related opportunities based on specialty and working model so candidates can keep momentum.