Forum Discussion
Azure SQL Mirrored DB says "metadata tables are corrupted" when stopping then restarting replication
I created a Mirrored SQL DB in fabric that connects to a source Azure SQL DB via a service principal. I was able to get the replication up and running as expected. It began mirroring tables as expected and the SQL Analytics Endpoint worked as expected. When I made changes to the source DB, those changes were mirrored into Fabric as expected. However...
When I turned off my Fabric compute and turned it back on agin, the Mirroring DB (in Fabric) said it was still replicating but changes in the source DB were not actually being updated in Fabric. I waited hours and still nothing. The last update timestamp on my mirrored tables were all from before I first turned off the Fabric compute. Finally, I stopped the replication in Fabric and tried to restart it ("Start Replication"). When I did, the replication failed to restart with this error: Code: SqlChangeFeedError, Type: UserError, Message: Cannot enable fabric link on the database because the metadata tables are corrupted. ArtifactId: 352bb680-e0cc-4927-85d9-333b0b391c78. Since then, my mirroring will not restart.
Replication is disabled in my source Azure SQL DB because when I manually "enable" it and then try to start replication, Fabric throws a different error saying that my source DB's replication is already on. But, with replication off in my DB, I have no way to troubleshoot the actual problem.
The only thing that has worked so far is restoring the source DB from a restore point before I ever started replicating, creating a brand new Mirrored DB in Fabric, and rebuilding the mirroring from scratch. But that will be a non-starter in production. I have to be able to fix a DB mirror without reverting my DB to a backup. And, more fundamentally, I need my DB mirrors in Fabric to be resiliant to disruptions in Fabric compute or other intermittent network partitions.
Any advice?
I heard back from MS support. TL;DR: you need to grant VIEW PERFORMANCE DEFINITION to the managed identity fabric is using to connect to the db:
GRANT SELECT, ALTER ANY EXTERNAL MIRROR, VIEW PERFORMANCE DEFINITION TO [User];The official docs only mention the need for `SELECT, ALTER ANY EXTERNAL MIRROR`.
This does not fix a database that already has corrupted metadata, but it makes it so that the reproduction steps I mention in the first post do not generate a metadata corruption in the first place.
I believe MS is planning to push a fix for this before the end of the month, but I don't know whether that will be just an update to the documentation or a code fix that makes the "ANY EXTERNAL MIRROR" permission sufficient.
12 Replies
- nlucero
Advocate I
I heard back from MS support. TL;DR: you need to grant VIEW PERFORMANCE DEFINITION to the managed identity fabric is using to connect to the db:
GRANT SELECT, ALTER ANY EXTERNAL MIRROR, VIEW PERFORMANCE DEFINITION TO [User];The official docs only mention the need for `SELECT, ALTER ANY EXTERNAL MIRROR`.
This does not fix a database that already has corrupted metadata, but it makes it so that the reproduction steps I mention in the first post do not generate a metadata corruption in the first place.
I believe MS is planning to push a fix for this before the end of the month, but I don't know whether that will be just an update to the documentation or a code fix that makes the "ANY EXTERNAL MIRROR" permission sufficient.
- WLY168
Advocate I
I can confirm that this access right solves my problem when using a "Service Principal" identity as registered in Entra against the workspace. i.e. the identity when registering the connection for the Mirror must pre-exist in SQL with the relevant logon, user and grants. You cannot create logins using the preview brower based session for querying Azure SQL Database. (You must use SQL Management Studio).
Assuming workspace name XXX and using Azure SQL Database (NOT managed instance)1. Use management studio to execute
CREATE LOGIN [XXX] FROM EXTERNAL PROVIDER;
2. Change to target Azure SQL Database and create userCREATE USER [XXX] FOR LOGIN [XXX]The documentation makes mention of a server role [##MS_ServerStateReader##] to which the login is added, but this does not make sense in the context of Azure SQL database where you cannot access master.
3. Grant the rights as specified in the documentation.GRANT SELECT, ALTER ANY EXTERNAL MIRROR, VIEW PERFORMANCE DEFINITION TO [XXX]; GO
Within Entra you will have to find the app registration for the workspace and generate a secret. Use that secret, the Entra Tenant ID and the Client ID of XXX (from Entra) to register the Service Principal based authenticated connection while registering the mirror connection.
I managed to stop and start the replication with impunity and many variations without causing any corruption to the target Azure SQL Satabase META data.
My only caveat is that I am working in a Trial Capacity that I cannot pause as per nlucero !!!
- v-ssriganesh
Community Support
Hi nlucero,
Thank you for reaching out to the Microsoft Fabric Forum Community.After thoroughly reviewing the details you provided, here’s how you can troubleshoot this without needing to restore your source database:
- Go to the mirrored database in the Fabric portal and use the Monitor Replication section. Look for outdated timestamps or error alerts on your tables. This will confirm if replication is stalled.
- In the Azure Portal, ensure the System Assigned Managed Identity (SAMI) for your Azure SQL logical server is enabled (check under Identity settings). Then, in Fabric, go to the mirrored database item, select Manage Permissions, and confirm the SAMI has Read and Write access.
- Check your Azure SQL Database allows connections from Azure services (check Networking settings in the Azure Portal). If you’re using a private endpoint, you may need to set up a virtual network data gateway in Fabric to maintain a stable connection.
If the error persists, you might need to delete the mirrored database item in Fabric (this won’t affect your source database) and create a new one using the same connection details. This can often resolve metadata issues without touching the source database.
If this information is helpful, please “Accept as solution” and give a "kudos" to assist other community members in resolving similar issues more efficiently.
Thank you.- nlucero
Advocate I
Thanks for the reply v-ssriganesh - I confirmed all of the steps above that they're not the root problem. Details below. But, in short, I am able to see the same error in the DB itself when I run
EXEC sys.sp_change_feed_enable_db @destination_type = 2; GOThe full error message is:
Msg 22710, Level 16, State 1, Procedure sys.sp_synapse_link_enable_db_internal, Line 415 [Batch Start Line 0] Could not update the metadata. The failure occurred when executing the command 'SetTridentLink(Value = 1)'. The error/state returned was 22697/1: 'Cannot enable fabric link on the database because the metadata tables are corrupted.'. Use the action and error to determine the cause of the failure and resubmit the request.So, it appears that the corrupted metadata tables are in my Azure SQL DB, independent of Fabric, and I don't know how to uncorrupt or reset them. Further, it will be a problem if the metadata tables become corrupted any time I stop replication and/or turn off my Fabric compute.
More Detail:
This is the view of the mirrored db in my Fabric portal. Notice that even though I previously had tables succesfully mirroring, the portal won't even list them when this error is thrown.
I went to "manage permissions" and confirmed that my SAMI is configured with Read/Write permissions on the db. I then went into the connection definition itself, removed the credentials, and then re-added them to confirm that the connection parameters are valid and the connection is active (it is).
Also, the database does allow connections from Azure services. I confirmed this. But, additionally, Fabric is having no problem connecting to that Azure SQL DB in general. I opened a new Data Factory job and successfully confirmed a connection to the same DB using the same access credentials. So, the connection itself doesn't seem to be the problem. As summarized above, I'm getting the "currupt metadata" error in the Azure SQL DB itself, even when not accessing from Fabric.
- nlucero
Advocate I
I was able to isolate the issue and the specific scenario in which it occurs:If I have a working Fabric SQL Mirror connected to a source Azure SQL DB and then I pause and then resume the Fabric compute, the Fabric mirror will report that it is still replicating but it is not. If I "Stop Replication" and then "Start Replication" on the mirror in the Fabric UI after restarting the fabric compute, then the metadata will become corrupted.However, if after restarting the compute I instead go to "Configure Replication" in the Fabric Mirrored DB, change nothing, but then "Apply Changes", it will successfully resume the replication.So the problem occurs specifically in the following sequence of events:- Start with a Fabric Mirror that is successfully replicating.
- Stop the Fabric compute that the mirror depends on.
- Restart the Fabric compute
- "Stop Replication" on the mirror
- "Start Replication" on the mirror
- This always produces a metadata corruption in the Azure SQL DB.
The problem also intermittently occurs when:- Start with a Fabric Mirror that is successfully replicating.
- Stop the replication (keeping the fabric compute on)
- Start the replication again
- This intermittently produces a metadata corruption in the Azure SQL DB.
And to avoid the metadata corruption:- Start with a Fabric Mirror that is successfully replicating.
- Stop the Fabric compute that the mirror depends on.
- Restart the Fabric compute.
- In the Fabric Mirror UI, click "Configure replication". Change nothing, then click "Apply changes"
- The replication will resume as expected.
Even though I now know how to avoid this issue (namely, by reconfiguring the replication instead of ever starting or stopping it), I still want to emphasize that this is a significant issue for anyone putting a production workload on Fabric because if the metadata becomes corrupted for any reason, it seems to require an Azure SQL restore from backup to resolve it. Creating a new mirrored DB on top of an Azure SQL DB that has already experienced metadata corruption does not work; it will not replicate once corrupted. So, I still need to know if there is a way to recover or reset an Azure SQL DB with corrupted replication metadata without restoring the DB from backup, because restoring from a backup means downtime and lost data.- v-ssriganesh
Community Support
Hi nlucero,
We sincerely regret the inconvenience this issue has caused and appreciate your detailed investigation, especially your valuable workaround using the reconfiguration approach. Given that the metadata corruption resides in the Azure SQL Database and requires a reset beyond standard portal options, we recommend raising a Microsoft support ticket for further assistance. You can create a support ticket using the link below:
https://learn.microsoft.com/en-us/power-bi/support/create-support-ticketPlease include the following details in your ticket to help the support team:
- The full error message (Msg 22710... Cannot enable fabric link on the database because the metadata tables are corrupted).
- The ArtifactId (352bb680-e0cc-4927-85d9-333b0b391c78 from your original post).
- The specific sequence triggering the issue (e.g., compute pause > stop/start replication).
- Your finding that reconfiguration avoids corruption.
If this guidance helps, please “Accept as Solution” and drop a “kudos” to make it easier for other community members to find. We hope this resolves the issue for you.
Thank you.