Data Engineer
NYP Holdings, Inc · United States
About The Role
The New York Post is one of America’s most iconic and recognizable news brands — delivering bold headlines, unforgettable scoops, and sharp commentary since 1801. With a legacy built on fearless journalism and a distinctly New York attitude, The Post has evolved far beyond its print roots to become a modern media powerhouse.
Today, the New York Post Digital Network is one of the most influential voices in the industry, reaching over 80 million unique users each month across our expansive, ever-growing portfolio. Anchored by NYPost.com, PageSix.com, and Decider.com, our digital brands deliver must-read breaking news, sports, entertainment, pop culture, and lifestyle coverage with the same wit, edge, and energy that define The Post. With innovation at our core, we’ve built a true multi-platform ecosystem — spanning web, mobile apps, video, audio, social, print, TV, and commerce. From viral headlines and exclusive stories to original series, podcasts, and live events, we’re engaging audiences where they are — and always on the pulse of what’s next.
We’re looking for a Data Engineer to join our technology team and work on the systems that power The Post’s digital future. This is a hands-on engineering role with a focus on infrastructure services, data pipelines, and identity enrichment. Along the way you’ll also participate in the product development of data-dependent systems such as AI answer services, personalization systems, machine learning, and audience services.
You will build and maintain the foundational data services that power analytics and machine learning for our newsroom, and product teams to deliver cutting-edge thoughtful digital experiences at scale. You’ll work alongside senior engineers, product managers, and data teams, contributing to high-performance systems.
Key Responsibilities
- Design, build, and operate cloud-based data pipelines supporting specialized needs of our consumer websites, native mobile apps, and information services.
- Work daily with compute services, query services, storage services, and big data services with technologies like Vertex AI, Lambdas, DynamoDB, Cloud Functions, S3, Kubernetes, Glue, and BigQuery inside AWS and GCP.
- Develop and maintain data ETL pipelines for real-time and batch processing of large-scale content and event data.
- Write clean, testable, and maintainable code (Python preferred).
- Collaborate with product, editorial, and data science teams to deliver infrastructure and services that meet business needs.
- Monitor, optimize, and improve the reliability, security, and cost efficiency of cloud-based systems.
- Write and present ideas clearly. Document functional requirements for both technical and non-technical audiences.
- Participate in code reviews, architecture discussions, and agile sprint ceremonies.
Qualifications
Must-Haves
- MS or BS in computer science, or related field.
- 2+ years of professional software engineering experience
- Proficiency in Python, SQL, and PySpark
- Working knowledge of common AWS and/or GCP services
- Experience building and deploying data pipelines (e.g., Airflow, Dataflow, Glue, or similar tools)
- Knowledge of data warehouse technologies (BigQuery, Snowflake, Redshift, etc.)
- Firm grasp of the tools of the trade including AI tooling, IDEs, Git, JIRA, and agile methodologies.
- Familiarity with web standards and protocols.
- Comfort with code reviews, writing unit tests, QA and UAT processes, and exposure to end-to-end testing frameworks.
- Strong problem-solving skills, attention to detail, thoroughness, and an eagerness to learn.
- Proactive approach to problem-solving.
- Highly organized with the ability to manage multiple priorities in a fast-paced environment.
Great-to-Have Experience
- Familiarity with content-management systems, media, news, publishing, video technology, ad technology, or other high-traffic consumer-facing industries.
- Exposure to personalization and recommendation systems.
- Comfortable with the CLI, Linux administration, shell scripting, ssh, crontabs, etc.
- Hands-on experience with generative-AI models, LLM prompting, and/or machine learning
- Experience with monitoring, logging, dashboard, and alarming tools (e.g., CloudWatch, Datadog, Splunk, etc.)
- Familiarity with REST and GraphQL APIs, authentication/authorization, and API best practices.
- Scaling techniques in cloud environments.
- Troubleshooting and debugging experience with internet services.
- Understanding of CI/CD workflows and containerization (Docker, Kubernetes a plus)
Note: The New York Post follows a hybrid work model. This position would be expected to be in the office a minimum of 3 days per week (subject to change).
Equal Opportunity Employer
All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, age, national origin, protected veteran status, disability status or any other protected characteristic. EEO/Disabled/Vets
Reasonable Accommodation
We are committed to providing reasonable accommodation for qualified individuals with disabilities in our job application and/or interview process. If you need assistance or accommodation in completing your application or participating in an interview due to a disability, email us at humanresources@newscorp.com . Please put "Reasonable Accommodation" in the subject line and provide a brief description of the type of assistance you need. This inbox will not be monitored for application status updates.
This listing was posted by a verified recruiter at NYP Holdings, Inc. Report this listing
JobSpring