<?xml version="1.0" encoding="UTF-8"?>
<rss xmlns:content="http://purl.org/rss/1.0/modules/content/" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:rdf="http://www.w3.org/1999/02/22-rdf-syntax-ns#" xmlns:taxo="http://purl.org/rss/1.0/modules/taxonomy/" version="2.0">
  <channel>
    <title>topic Re: Managed Disaster Recovery in Administration &amp; Architecture</title>
    <link>https://community.databricks.com/t5/administration-architecture/managed-disaster-recovery/m-p/169798#M5671</link>
    <description>&lt;P&gt;Hi Thomas,&lt;BR /&gt;I actually tried deleting the workspace and it works! once the workspaces are deleted, the workspace get de-attached from the failover group making failover group empty and then we can delete it successfully.&lt;/P&gt;</description>
    <pubDate>Fri, 25 Sep 2026 09:33:54 GMT</pubDate>
    <dc:creator>yxnz1e</dc:creator>
    <dc:date>2026-09-25T09:33:54Z</dc:date>
    <item>
      <title>Managed Disaster Recovery</title>
      <link>https://community.databricks.com/t5/administration-architecture/managed-disaster-recovery/m-p/169783#M5669</link>
      <description>&lt;P&gt;Hi,&amp;nbsp;&lt;/P&gt;&lt;P&gt;I have been doing testing on Managed DR, just few questions that i'm unable to find solution for:&lt;BR /&gt;1. I have observed that it takes more time in replication for empty catalogs rather than catalogs that have some data, does replication get stuck somewhere?&lt;BR /&gt;2. I had a metastore created for primary workspace which i eventually deleted and created new one, although the workspace picked up the new metastore, when i created failover group &amp;amp; tried to delete it, i got the error:&lt;/P&gt;&lt;P&gt;&lt;span class="lia-inline-image-display-wrapper lia-image-align-inline" image-alt="yxnz1e_0-1790318716914.png" style="width: 400px;"&gt;&lt;img src="https://community.databricks.com/t5/image/serverpage/image-id/31492i8A4FA8732A1483A0/image-size/medium?v=v2&amp;amp;px=400" role="button" title="yxnz1e_0-1790318716914.png" alt="yxnz1e_0-1790318716914.png" /&gt;&lt;/span&gt;&lt;/P&gt;&lt;P&gt;apparently this metastore ID belongs to the metastore i deleted. Then if i delete the workspace will the failover group be deleted too?&lt;BR /&gt;&lt;BR /&gt;&lt;/P&gt;&lt;P&gt;&lt;BR /&gt;&lt;BR /&gt;&lt;/P&gt;</description>
      <pubDate>Fri, 25 Sep 2026 06:46:17 GMT</pubDate>
      <guid>https://community.databricks.com/t5/administration-architecture/managed-disaster-recovery/m-p/169783#M5669</guid>
      <dc:creator>yxnz1e</dc:creator>
      <dc:date>2026-09-25T06:46:17Z</dc:date>
    </item>
    <item>
      <title>Re: Managed Disaster Recovery</title>
      <link>https://community.databricks.com/t5/administration-architecture/managed-disaster-recovery/m-p/169797#M5670</link>
      <description>&lt;P&gt;Hi,&lt;/P&gt;&lt;P&gt;I've tested managed DR as well, and the docs cover less of this than you'd hope, so here's what they do say.&lt;/P&gt;&lt;P&gt;&lt;STRONG&gt;1. Slow replication on empty catalogs.&lt;/STRONG&gt; The docs don't explain this directly, but two details help. The replication point is group-wide: it "shows the last time all in-scope resources were copied together." So an empty catalog only shows as replicated once the whole cycle finishes. Also, the system table doesn't track individual catalogs: it "does not list which individual objects replicated successfully," and a null lag means "at least one asset has never been replicated." To tell whether it's actually stuck or just waiting on the cycle, check the errors column:&lt;/P&gt;&lt;DIV&gt;&lt;DIV&gt;&lt;DIV&gt;&lt;DIV&gt;&amp;nbsp;&lt;/DIV&gt;&lt;/DIV&gt;&lt;DIV&gt;sql&lt;/DIV&gt;&lt;DIV&gt;&lt;PRE&gt;&lt;SPAN&gt;SELECT event_time, replication_state, replication_lag_ms, errors
FROM system.replication.states
WHERE failover_group_name LIKE '%&amp;lt;your-group&amp;gt;%'
ORDER BY event_time DESC;&lt;/SPAN&gt;&lt;/PRE&gt;&lt;/DIV&gt;&lt;/DIV&gt;&lt;/DIV&gt;&lt;P&gt;The table can take up to 3 hours to populate. If you see a rising lag with no errors, that's worth raising with your account team.&lt;BR /&gt;&lt;A href="https://docs.databricks.com/aws/en/admin/system-tables/replication" target="_blank" rel="noopener"&gt;link doc&lt;/A&gt;&lt;/P&gt;&lt;P&gt;&lt;STRONG&gt;2. The error about the deleted metastore.&lt;/STRONG&gt; The failover group records the metastores it manages (metastore_ids). Swapping the workspace to a new metastore doesn't update the group, so it's still pointing at the one you deleted. The docs don't describe any way to clean that up. There's no force-delete, and nothing on deleting a metastore or workspace that belongs to a failover group. They do say the group sets up a connection and a foreign catalog in each metastore, and that you shouldn't delete those yourself. It's likely that the missing old metastore is what blocks the teardown.&lt;BR /&gt;&lt;A href="https://docs.databricks.com/aws/en/admin/managed-disaster-recovery" target="_blank" rel="noopener"&gt;link doc&lt;/A&gt;&lt;/P&gt;&lt;P&gt;On deleting the workspace: I wouldn't. The docs never say it cascades to the failover group, so you could end up with an orphaned group, and deleting a workspace can't be undone. Managed DR is gated and enabled by the account team, so they (or a support ticket) are the right channel. Send them the failover group name, the old metastore ID from the error, and the group's current state (probably DELETION_FAILED). This needs a fix on their side.&lt;/P&gt;&lt;P&gt;For future tests, don't reassign or delete a metastore while a failover group references it. The documented teardown is to delete the group first and then turn off Mission Critical on each workspace.&lt;/P&gt;</description>
      <pubDate>Fri, 25 Sep 2026 09:26:29 GMT</pubDate>
      <guid>https://community.databricks.com/t5/administration-architecture/managed-disaster-recovery/m-p/169797#M5670</guid>
      <dc:creator>ThomazNeto</dc:creator>
      <dc:date>2026-09-25T09:26:29Z</dc:date>
    </item>
    <item>
      <title>Re: Managed Disaster Recovery</title>
      <link>https://community.databricks.com/t5/administration-architecture/managed-disaster-recovery/m-p/169798#M5671</link>
      <description>&lt;P&gt;Hi Thomas,&lt;BR /&gt;I actually tried deleting the workspace and it works! once the workspaces are deleted, the workspace get de-attached from the failover group making failover group empty and then we can delete it successfully.&lt;/P&gt;</description>
      <pubDate>Fri, 25 Sep 2026 09:33:54 GMT</pubDate>
      <guid>https://community.databricks.com/t5/administration-architecture/managed-disaster-recovery/m-p/169798#M5671</guid>
      <dc:creator>yxnz1e</dc:creator>
      <dc:date>2026-09-25T09:33:54Z</dc:date>
    </item>
  </channel>
</rss>

