Microsoft purview and azure data governance
Skill vaquarkhan/data-engineering-agent-skills/skills/microsoft-purview-and-azure-data-governance
Production-grade Agent Skills for data engineering AI agents: 73 workflows, platform presets, safe backfill/replay, Kafka & Spark reliability, MCP observability, and VS Code/JetBrains installers.
npx -y skills add vaquarkhan/data-engineering-agent-skills --skill microsoft-purview-and-azure-data-governanceAssembled from the repository path, not quoted from the project. Check it against their README if it does not work.
One thing to look at
- 21 stars21 stars. Stars are a popularity signal and not a quality one, but at this level it is likely that nobody has read this closely except its author, and you would be relying on your own review.
What its author says it does
Copied from the file, not written here
Guides agents through Microsoft Purview and Azure-native data governance workflows. Use when designing collections, scans, classifications, lineage, policy boundaries, and governed publishing across ADLS, Synapse, Data Factory, Azure Databricks, Fabric, and Azure analytics estates.
SKILL.md
3.1 KB, 571 tokens by cl100k_base, as published. Nobody here has run it
Microsoft Purview And Azure Data Governance
Overview
Use this skill when Microsoft Purview is the governance control plane for Azure data platforms. It helps agents design metadata collections, classification strategy, scan scope, lineage expectations, and governed publish behavior across Microsoft analytics surfaces.
When to Use
- designing
Purviewcollection and ownership structure - defining scans, classifications, glossary, and lineage expectations
- governing datasets across
ADLS,Synapse,Data Factory,Azure Databricks, orFabric - improving trusted discovery and certification for shared data products
- aligning Azure-native governance with privacy, security, and release controls
Do not assume Purview is only a catalog tool. It often becomes the platform evidence and policy surface for governed analytics.
Workflow
-
Define governance scope. Clarify:
- in-scope platforms
- business domains
- critical data products
- stewardship and ownership model
-
Design the metadata operating model. Decide:
- collections
- glossary boundaries
- classifications and sensitivity labels
- scan cadence and ownership
-
Define trusted publish behavior. Require:
- certification or endorsement rules
- lineage completeness expectations
- ownership visibility
- ties to regulated-data controls where relevant
-
Align Azure services with governance. Check how
Purviewworks withADLS,Synapse,Data Factory,Databricks, andFabricrather than treating each system separately. -
Validate operational sustainability. Make sure scans, classifications, and lineage remain useful as assets, teams, and environments grow.
Common Rationalizations
| Rationalization | Reality |
|---|---|
| "Scanning everything is the same as governing it." | Governance also needs ownership, trust signals, and useful boundaries for consumers. |
| "Purview can be added after pipelines are done." | Late governance usually means weak lineage, poor certification, and inconsistent discovery. |
| "Each Azure service team can manage metadata separately." | Fragmented governance weakens platform-wide trust and policy evidence. |
Red Flags
- collections do not map to real ownership or domains
- scan scope is broad but lineage and certification are weak
- classifications are inconsistent across Azure services
Purviewis disconnected from publish or security decisions- stewardship expectations depend on tribal knowledge
Verification
- Governance scope and stewardship model are explicit
- Collections, scans, and classifications are intentionally designed
- Trusted publish behavior includes lineage and certification expectations
- Azure services align to one governance model
- The model stays sustainable as adoption grows
What ships with it
Read from the repository
Just SKILL.md. No reference files, no scripts.