About this role
About the Role
Perpay is seeking a Data Engineer to join our cross-functional data team (Data Engineering, Data Science, and Strategic Analytics). You’ll own the warehouse, the orchestration layer, and the pipelines that move data from our operational systems to teams who depend on it, enabling trustworthy production analytics. This role supports scaling the credit portfolio with real-time ingestion, risk decisioning improvements, and AI-assisted data access on Redshift Serverless. You’ll partner with Engineering, Risk, Commerce, Accounting, and Compliance to solve business problems with data platforms.
What You'll Do
-
Own and evolve the data warehouse, orchestration, and end-to-end data pipelines, from ingestion to downstream consumers.
-
Design and implement real-time event ingestion, the risk decisioning service redesign, ERP/EDI standardization with Finance and Accounting, and AI-assisted workflows across the engineering lifecycle.
-
Provide data access and governance support using DataHub, enabling discoverability and lineage across teams.
-
Collaborate with Engineering, Risk, Commerce, Accounting, and Compliance; translate business needs into scalable data products and reliable pipelines.
-
Write production-grade code in Python and SQL; document decisions; advocate for data quality, observability, and reliability; participate in code reviews and architecture discussions.
-
Move across SQL, Python, orchestration (Airflow), and infrastructure-as-code (Terraform) without requiring one area to be your sole specialty; contribute to architecture decisions and end-to-end design.
-
Be the technical lead on meaningful projects within 3 months, guiding design, implementation, testing, and rollout, and engaging stakeholders as needed.
-
What We're Looking For
-
At least two years of production data engineering experience in a comparable environment.
-
Comfort moving across SQL, Python, orchestration, and infrastructure-as-code without needing one area to be your specialty.
-
Proficiency with AWS data stack: Redshift and Spectrum for warehousing; Glue and Fivetran for ingestion; Airflow for orchestration; ECS/ECR for services; DMS for replication; DataHub for cataloging and lineage; Terraform to hold it together.
-
Strong collaboration with cross-functional teams (Engineering, Risk, Commerce, Marketing/Ops, Finance, Compliance) and ability to communicate trade-offs clearly.
-
A passion for data quality, pipeline reliability, governance, and scalable data product thinking.
-
Nice to Have
-
Spark experience for larger-scale problems.
-
Experience with Redshift Serverless and AI/ML-enabled data access patterns.
-
Familiarity with DataHub, ERP/EDI standardization, and modern data governance practices.
-
Compensation & Benefits
-
Generous perks and benefits; center-city Philadelphia office with river views; flexible in-office culture with remote weeks around major holidays.
-
The team values meaningful mission, collaboration, equity, and growth opportunities.