Data Engineer
RemotePhiladelphia, Pennsylvania, United States
Job Summary
Design, build, and maintain serverless data pipelines and data models on AWS using Lambda, Glue, Athena, and S3 to support reporting, analytics, and AI use cases. Develop ETL/ELT workflows to ingest, transform, and load data from internal applications and third-party sources into the data lake, while implementing data quality controls, monitoring, and alerting to ensure accuracy and lineage. Partner with Product, Underwriting, Analytics, and AI stakeholders to translate data requirements into robust structures and SLAs, and optimize workloads for cost, performance, and scalability. Troubleshoot pipeline issues, resolve incidents, and document data contracts to enable self-service analytics. Own ingestion, transformation, and storage patterns to make high-quality, analytics-ready data available to downstream teams.
Required Qualifications
- Candidates must be authorized to work in the United States without sponsorship both now or in the future
- Strong experience with AWS data and serverless services (Lambda, Glue, Athena, S3, Step Functions or similar orchestration tools)
- Proficiency in Python and TypeScript or a similar language for data engineering, including building ETL/ELT jobs and reusable libraries
- Solid understanding of data modeling principles (dimensional, normalized, wide‐table, and event‐driven designs) for analytics and AI workloads
- Experience designing and operating data pipelines at scale, including batch and near‐real‐time ingestion
- Familiarity with SQL and query optimization in columnar data stores and engines such as Athena
- Knowledge of data quality, governance, and security practices, including handling sensitive healthcare and financial data
- Hands‐on experience with Infrastructure as Code, preferably AWS CDK, for provisioning and managing data infrastructure
- Strong communication skills and the ability to translate complex data concepts into clear, actionable language for business stakeholders
- Bachelor's degree in Computer Science, Engineering, Mathematics, or related field, or equivalent practical experience
- 3+ years of experience in data engineering or software engineering roles focused on data pipelines and analytics platforms
- 2+ years of hands‐on experience with AWS data services in a production environment (e.g., Lambda, Glue, Athena, S3)
- Experience building and maintaining data solutions that support analytics, BI, and/or AI/ML initiatives
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.