Practise vocabulary for data lineage, metadata, business glossary, data stewardship, and PII tagging.
0 / 5 completed
1 / 5
A data catalog is described as the 'Google search for your data' because:
Data catalogs (Alation, Atlan, DataHub, OpenMetadata) solve data discoverability: without a catalog, analysts don't know what data exists, where it is, what it means, or whether it's trustworthy. A catalog indexes assets and makes them searchable with rich context.
2 / 5
A business glossary in a data catalog maps:
Business glossary example: 'Active User' may be defined as 'a user who has logged in at least once in the last 30 days and has completed at least one core action' — with a link to the fact table and the exact query implementing this definition. Without this, different teams use different definitions.
3 / 5
Data stewardship refers to:
Data stewards are typically business domain experts (not engineers) who ensure their domain's data is accurate and well-defined. They answer: 'What does this field mean?', 'Who is responsible for its quality?', 'When should this data be deleted?' — bridging business and technical teams.
4 / 5
PII tagging in a data catalog means:
PII tagging drives automated governance: tagged columns can have access restricted to authorised users, can trigger masking in non-production environments, and can be included in automated data deletion workflows for GDPR erasure requests — without engineers having to remember which columns contain personal data.
5 / 5
Data discoverability in the context of a data catalog means:
Poor discoverability creates shadow analytics: analysts build their own spreadsheets because they can't find the official data. A catalog with good metadata, descriptions, popularity signals, and verified owner contact information makes the official data the path of least resistance.
What does the "Data Catalog & Governance Vocabulary" exercise practise?How many questions are in this exercise?
This exercise has 5 questions, each multiple-choice with a full explanation shown after you answer.
What English level is this exercise for?
This exercise is tagged Intermediate. If the vocabulary feels difficult, browse the Data Engineering Language category page for an easier module to start with.
Is this exercise free to use?
Yes. Every exercise on CoderSlingo, including this one, is free with no account, sign-up, or paywall.
Do I get feedback if I answer incorrectly?
Yes — whichever option you choose, right or wrong, you'll immediately see an explanation clarifying the correct term and why the other options don't fit.
Can I retry this exercise?
Yes — once you finish all the questions, a "Try again" button on the results screen resets the exercise so you can practise as many times as you like.
Do I need an account to track my progress?
No account is required. Your progress bar and score for this session are tracked in the browser as you go, but nothing is saved once you leave the page.
Is "Data Catalog & Governance Vocabulary" part of a larger series?
Yes — it's one exercise in the Data Engineering Language category on CoderSlingo. See the category page for the full list of related exercises on similar terminology.
Can I link directly to this exercise?
Yes — this exercise has its own permanent URL, so you can bookmark it or share the link directly with a colleague or study partner.
Where can I find more exercises like this one?
See the Data Engineering Language category page for related exercises, or browse the main Exercises hub for other IT English topics.