Responsibilities
- Own Pipeline Development End-to-End: Design, build, and maintain robust, scalable ETL/ELT pipelines that reliably deliver clean data to the platform.
- Add New Data Sources: Evaluate, scope, and integrate new data sources as they're identified.
- Partner Across Pillars: Work continuously with Network Priorities, Regional Enablement and AI Enablement to understand incoming data needs and translate them into pipeline work.
- Monitor & Troubleshoot: Proactively identify and resolve pipeline failures and data quality issues before they affect downstream users.
- Support Foundational Data Models: Work with the Senior Analytics Engineer to maintain the core data models the rest of the team depends on.
- Ensure Data Availability for Consumers: Make sure the data needed by Data Analysts and the Senior Machine Learning Engineer is reliably available.
- Follow Data Governance Standards: Apply the data standards, definitions, and sensitivity classifications set by the Data Governance Lead.
- Evaluate Pipeline Tooling: Explore and pilot new ETL/ELT tools or approaches that could improve onboarding speed or pipeline reliability.
- Quantify Impact: Track and articulate how pipeline reliability and onboarding speed affect downstream analytics and ML work.
Requirements
- Minimum of 2+ years of experience in data engineering
- Experience working in a DevOps-oriented culture that prioritizes continuous integration and continuous deployment
- Proficiency with Git and collaborative version control workflows (e.g., branching, pull requests, code review)
- Experience with Infrastructure as Code (e.g., Terraform, CloudFormation, or CDK) for provisioning and managing cloud infrastructure
- Proven experience in designing and deploying data solutions
- Experience designing, building, and onboarding new data sources into ETL/ELT pipelines
- Ability and desire to take product/project ownership
- Proficiency in SQL and experience with scripting languages such as Python, Java, or Scala
- Experience with data pipeline and workflow management tools
Nice to Have
- A BA/BS in Computer Science, Information Technology, or a related field strongly preferred
- Strong knowledge of big data tools and frameworks such as Hadoop, Spark, or Hive is a plus
- Experience using AI coding assistants and other AI tools to improve development speed and productivity
Compensation
An annual salary in the range of $62,900 - $83,000 based on experience and skills
Additional Information
- This is a Full-Time Role (40 hours per week, 5 days per week) with no option for part-time work.
- The candidate must hold U.S. work authorization without any time limitations or any other restrictions if applying for the New York role.
- The candidate must be a resident of New York State at the start of employment for the New York role.
- The candidate must be within commuting distance of our office in Philadelphia, New York City, or São Paulo.
- All applications must include a resume and complete responses to standard application questions.
- Cover letters should not be included.
- Interviews require cameras to be on via Google Meet or Zoom.
- First day of work must be in-person at one of our office locations to complete onboarding documents and meet with team members.
- Applications will be reviewed starting August 11th, 2026, and the job ad closes on August 30th, 2026.
- All candidates will receive a status update via email after application review, expected by mid-March.
- Reasonable accommodations available upon request.
- Feedback on equity or accessibility of recruitment welcome.