Role: Senior Data Governance Consultant - Informatica CDGC
Location: Newark, NJ - Hybrid (3 Days onsite per week)
Duration: 12 months
About Role:
Senior consultant with expert-level Informatica CDGC skills to scan, catalog, govern, and publish a complex data estate spanning relational, NoSQL, API, and event-driven source systems. A significant portion of work involves non-SQL sources (MongoDB, AWS DocumentDB) where Informatica has no native connectors - requiring a resourceful engineer who can assess available tooling and implement a viable ingestion solution. The consultant will also stand-up Data Products in Informatica.
What you will do:
- Design and execute metadata scanning across relational, NoSQL, API, and event-driven source types into Informatica CDGC
- Implement viable scanning solutions for MongoDB and AWS DocumentDB using Informatica tooling, ingestion APIs, and scripted extraction
- Catalog REST APIs and microservices as governed assets; integrate event-driven sources (Kafka, AWS Kinesis/SQS) for lineage tracking
- Configure connections across Oracle, MS SQL Server, PostgreSQL, AWS S3, and NoSQL platforms
- Build and validate end-to-end data lineage from source systems through AWS S3, Microsoft Fabric, and Power BI
- Define and publish Data Products in Informatica Marketplace for governed, consumable asset delivery
- Support curation workflows: classification, glossary association, profiling, and stewardship alignment
- Troubleshoot scanning failures, connectivity issues, and metadata quality gaps
- Document all configurations, ingestion approaches, lineage patterns, and Data Product definitions for team knowledge transfer
Required Skills & Experience:
- Informatica CDGC - expert-level; hands-on production experience with scanning, cataloging, lineage, classification, and Data Products
- Informatica Profiling - hands-on experience running and interpreting data profiling in Informatica to assess data quality, completeness, and patterns across scanned assets
- Informatica Data Products & Marketplace - defining, publishing, and managing data products
- NoSQL scanning solutions - proven ability to implement metadata ingestion for sources without native Informatica connectors
- Non-SQL databases - MongoDB and AWS DocumentDB: schema structures, connection methods, metadata extraction
- API scanning - cataloging REST APIs and microservices as governed assets in Informatica CDGC
- Event processing - integrating Kafka, AWS Kinesis, or equivalent into a data catalog for lineage and governance
- Data Engineering - strong Python and/or Java for scripting and ingestion work
- Relational Databases - Oracle, MS SQL Server, PostgreSQL
- Cloud Platforms - AWS (S3, DocumentDB, Kinesis), Microsoft Azure
- Data Governance Concepts - catalog management, classification, glossary, stewardship workflows
Nice to Have:
- Microsoft Fabric or Power BI lineage integration
- Atacama RDM or Denodo experience
- Financial services or insurance industry background
- Informatica CDI experience
|