Data Engineer, Analytics Data Products

The New York Times · New York, NY

Posted
11 days ago
Last confirmed live
2 days ago

What this role involves

The Data Engineer, Analytics Data Products role at The New York Times involves designing and implementing complex ELT/ETL pipelines and data models using dbt and PySpark within a medallion architecture. The position includes managing physical data storage across GCP and AWS, optimizing Spark compute resources, and owning components of the centralized analytics environment like Hex. Collaboration with cross-functional teams to translate requirements into scalable data models and ensuring data quality and observability are key responsibilities.

Skills this posting asks for

  • sql
  • dbt
  • pyspark
  • data modeling
  • dimensional modeling
  • kimball
  • obt
  • data vault
  • elt
  • etl
  • gcp
  • aws
  • spark
  • dataproc
  • emr
  • hex
  • data quality
  • metadata management
  • data lineage
  • rbac
  • python

Requirements

  • 2 years of experience
  • Level: mid

From the employer’s posting

<div id="labeledImage.LOCATION--uid38" class="WE-Y WMXY WBAB WF0Y" data-automation-id="responsiveMonikerInput" data-metadata-id="labeledImage.LOCATION" data-uxi-form-it…

Read the full description on The New York Times’s careers page

Apply without filling the form

Approve this role and the application is completed for you, including a résumé tailored to it. You get a confirmation when it lands, and a credit is only spent when a submission is confirmed.

Other roles at The New York Times

All 40 roles at The New York Times