Use ↑ ↓ arrows or scroll
CoComply · Module 10

Data
Onboarding

Where the platform meets the estate — metadata-first, without moving underlying data.

10
01
Capability Overview

Sources in, governance at scale

Sources are registered, connectors configured, and assets flow into governance through guided registration or bulk import. Connecting a source begins harvesting structure and context immediately — no data movement required.

02
Pages in This Section

Three onboarding surfaces

PAGE 1

Source Registration

Guided registration of databases, warehouses, lakes, pipelines & BI platforms — with ownership and domain captured at entry.

PAGE 2

Connectors & Integrations

Configuration and health monitoring for every connector, covering auth models, harvest scheduling, and scope controls.

PAGE 3

Bulk Import

Template-driven mass registration of assets, CDEs, and mappings for programs migrating from spreadsheets or legacy tools.

03
Integrations Reference

One connector fabric, three interaction patterns

PATTERN 1

Read-Only Metadata APIs

Catalog, schema, and config harvesting through native metadata services. Incremental — never touches row-level data.

PATTERN 2

Query & Job Log Ingestion

Runtime execution traces from query history, audit logs, and pipeline runs — the richest as-executed lineage.

PATTERN 3

In-Environment Pushdown

Quality & profiling jobs run inside the client's own cloud. Only scores, outcomes, and metadata return.

04
Service Integration Summary

GCP, AWS & Azure — one fabric

FunctionGoogle CloudAWSAzure
Warehouse & queryBigQueryRedshift, AthenaSynapse, Fabric
Catalog harvestingDataplex, Data CatalogGlue Data CatalogMicrosoft Purview
Storage & lakeCloud StorageS3, Lake FormationADLS Gen2
Pipeline metadataDataflow, ComposerGlue ETL, Step FunctionsData Factory
Log ingestionCloud Audit LogsCloudTrail, query logsAzure Monitor
In-tenant AI (optional)Vertex AIBedrock, SageMakerAzure OpenAI, Azure ML
05
Data Residency by Design

The estate stays where it is

06
Key Takeaway

Onboarding starts the whole engine

Because connection begins metadata harvesting immediately, Data Lineage and Data Quality start working from the moment a source is registered — governance at scale, without moving the data.

07