DataLake Architect
Design a data-lake layout that supports discovery, retention, and recovery.
Map the data and control boundaries
Trace the reconciliation kit from source files to review output. Sketch the production sources, trust boundaries, integration path, storage, and downstream consumers before choosing a deployment.
Start with the browser demo or download the local kit. For the workspace exercise, you need approved access to the relevant Genedata capabilities and a reviewer for your output.
Practice tasks
0 / 4 completed
Notes and progress are saved on this device.
Make storage a usable data foundation.
Design a data-lake layout that supports discovery, retention, and recovery.
Inventory the data domains, artifact types, consumers, and lifecycle requirements.
Define a consistent storage layout, naming scheme, partition strategy, and ownership model.
Review retention, archival, access, and recovery procedures with governance and operations.
Publish the design and validate a representative ingestion and retrieval workflow before expanding it.
Expected handoff
A storage design with lifecycle rules, ownership, and recovery evidence.
Explore your platform capabilitiesMeasure your progress
- Time to locate the correct dataset or artifact
- Storage domains with reviewed lifecycle and recovery rules
Use AI to summarize layout alternatives; validate cost, access, and retrieval behavior against the real workload.