priya9896
New Contributor

Hi everyone,

We're evaluating an architecture pattern and would appreciate any guidance or recommendations.

Current state:

Source data resides in Azure Databricks.
Consumers are in AWS .
The source team allows read-only access but does not allow data replication or ingestion into AWS.
Goal: We would like AWS consumers to access the data through Glue Catalog tables while keeping the data in Azure Databricks.

Has anyone implemented a similar cross-cloud pattern? Specifically:

Can AWS Glue be used to access Azure Databricks data without copying it?
Are there recommended approaches such as federation, Delta Sharing, custom connectors, JDBC/SQL endpoints, or other patterns?

Appreciate all your time and inputs regarding . Thanks in advance!