Data Engineer, Analytics Data Products
The New York Times · New York, NY
- Posted
- 11 days ago
- Last confirmed live
- 2 days ago
What this role involves
The Data Engineer, Analytics Data Products role at The New York Times involves designing and implementing complex ELT/ETL pipelines and data models using dbt and PySpark within a medallion architecture. The position includes managing physical data storage across GCP and AWS, optimizing Spark compute resources, and owning components of the centralized analytics environment like Hex. Collaboration with cross-functional teams to translate requirements into scalable data models and ensuring data quality and observability are key responsibilities.
Skills this posting asks for
- sql
- dbt
- pyspark
- data modeling
- dimensional modeling
- kimball
- obt
- data vault
- elt
- etl
- gcp
- aws
- spark
- dataproc
- emr
- hex
- data quality
- metadata management
- data lineage
- rbac
- python
Requirements
- 2 years of experience
- Level: mid
From the employer’s posting
<div id="labeledImage.LOCATION--uid38" class="WE-Y WMXY WBAB WF0Y" data-automation-id="responsiveMonikerInput" data-metadata-id="labeledImage.LOCATION" data-uxi-form-it…
Read the full description on The New York Times’s careers pageApply without filling the form
Approve this role and the application is completed for you, including a résumé tailored to it. You get a confirmation when it lands, and a credit is only spent when a submission is confirmed.
Other roles at The New York Times
- Software Engineer, ReflectionsNew York, NY
- Senior Data Engineer, Customer-Facing Data ProductsNew York, NY
- Senior Product Designer, Platforms & Foundations (Temporary)New York, NY
- Staff Software Engineer, Application DeliveryRemote
- Associate Newsroom Software Engineer, Interactive NewsNew York, NY
- Staff Newsroom Security EngineerNew York, NY