Hi everyone,
We're evaluating an architecture pattern and would appreciate any guidance or recommendations.
Current state:
Source data resides in Azure Databricks.
Consumers are in AWS .
The source team allows read-only access but does not allow data replication or ingestion into AWS.
Goal: We would like AWS consumers to access the data through Glue Catalog tables while keeping the data in Azure Databricks.
Has anyone implemented a similar cross-cloud pattern? Specifically:
Can AWS Glue be used to access Azure Databricks data without copying it?
Are there recommended approaches such as federation, Delta Sharing, custom connectors, JDBC/SQL endpoints, or other patterns?
Appreciate all your time and inputs regarding . Thanks in advance!