Data Engineer | Scrabble & Jigsaw
Job Description
- Design, build, and operate pipelines that ingest data from: - AWS RDS (PostgreSQL) - application DB for the lending platform, Salesforce CDC sync - DynamoDB - lead capture via GP service front-ends - SFTP drops - bureau data files from CRIF, Equifax, Experian (multi-hundred-million row tables) - Offline Excel / CSV MIS files - lender disbursement and agent invoicing sheets - Email attachments - auto-parsed MIS and ops reports landing in shared inboxes - Implement a Bronze-Silver-Gold Medallion architecture in Snowflake with clear ownership of schema design, DDL, and lineage documentation. - Ensure idempotent, observable, and failure-tolerant pipeline runs with proper alerting. 2. Reverse / Outbound Pipelines (Snowflake to App Databases) - Build reverse ETL flows to push aggregated or enriched datasets back to: - AWS RDS - feeding app-layer dashboards and internal ops tooling - MongoDB - intermediary store for product APIs and internal CRM features - Collaborate with backend engineers to agree on schemas and change-management processes so app-tier writes are non-breaking.
