<?xml version="1.0" encoding="UTF-8"?>
<rss xmlns:content="http://purl.org/rss/1.0/modules/content/" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:rdf="http://www.w3.org/1999/02/22-rdf-syntax-ns#" xmlns:taxo="http://purl.org/rss/1.0/modules/taxonomy/" version="2.0">
  <channel>
    <title>topic Re: Data ingestion: Setting up a connector for Postgres with CDC enabled. in Data Engineering</title>
    <link>https://community.databricks.com/t5/data-engineering/data-ingestion-setting-up-a-connector-for-postgres-with-cdc/m-p/169238#M56063</link>
    <description>&lt;P&gt;Hi Wola,&lt;/P&gt;&lt;P&gt;Don't worry about your Postgres side, that's most likely not the problem. The greyed-out option is coming from Databricks. The PostgreSQL CDC connector in Lakeflow Connect is still in Public Preview, and this one isn't self-service: the docs say "Contact your Databricks account team to request access." There is nothing to flip on the Previews page. Until the workspace is enrolled, the wizard only lets you pick Query-based capture, which is exactly what your screenshot shows.&lt;BR /&gt;&lt;A href="https://docs.databricks.com/aws/en/ingestion/lakeflow-connect/postgresql" target="_blank"&gt;https://docs.databricks.com/aws/en/ingestion/lakeflow-connect/postgresql&lt;/A&gt;&lt;/P&gt;&lt;P&gt;So step one is a message to your account team asking to enable the PostgreSQL connector preview for your workspace.&lt;/P&gt;&lt;P&gt;While you wait, two things worth checking. The CDC pipeline needs serverless compute enabled in the workspace (the gateway runs on classic compute, the pipeline itself is serverless), and it only replicates from a primary instance, read replicas won't work.&lt;BR /&gt;&lt;A href="https://docs.databricks.com/aws/en/ingestion/lakeflow-connect/postgresql-pipeline" target="_blank"&gt;https://docs.databricks.com/aws/en/ingestion/lakeflow-connect/postgresql-pipeline&lt;/A&gt;&lt;/P&gt;&lt;P&gt;And the RDS checklist from the docs, just to compare with what you've done: rds.logical_replication = 1 in the parameter group, a replication user with the rds_replication role and the REPLICATION attribute, a publication, a replication slot created with pgoutput (the only plugin supported), and REPLICA IDENTITY set to DEFAULT or FULL on every table. The slot name and publication name are what the wizard asks for later, in the Database setup step.&lt;BR /&gt;&lt;A href="https://docs.databricks.com/aws/en/ingestion/lakeflow-connect/postgresql-source-setup" target="_blank"&gt;https://docs.databricks.com/aws/en/ingestion/lakeflow-connect/postgresql-source-setup&lt;/A&gt;&lt;BR /&gt;&lt;A href="https://docs.databricks.com/aws/en/ingestion/lakeflow-connect/postgresql-limits" target="_blank"&gt;https://docs.databricks.com/aws/en/ingestion/lakeflow-connect/postgresql-limits&lt;/A&gt;&lt;/P&gt;&lt;P&gt;If you can't wait, Query-based capture does work with Postgres today, no preview needed. Just know what you're getting: it doesn't read the WAL, it runs on a schedule and picks up changes through a cursor column (a timestamp or increasing id) per table. You won't see intermediate states of a row, and hard deletes are only tracked in Beta. Fine for a lot of use cases, but it isn't real CDC.&lt;BR /&gt;&lt;A href="https://docs.databricks.com/aws/en/ingestion/lakeflow-connect/query-based-overview" target="_blank"&gt;https://docs.databricks.com/aws/en/ingestion/lakeflow-connect/query-based-overview&lt;/A&gt;&lt;/P&gt;&lt;P&gt;Good luck with it.&lt;/P&gt;</description>
    <pubDate>Sun, 20 Sep 2026 12:00:29 GMT</pubDate>
    <dc:creator>ThomazNeto</dc:creator>
    <dc:date>2026-09-20T12:00:29Z</dc:date>
    <item>
      <title>Data ingestion: Setting up a connector for Postgres with CDC enabled.</title>
      <link>https://community.databricks.com/t5/data-engineering/data-ingestion-setting-up-a-connector-for-postgres-with-cdc/m-p/169117#M56042</link>
      <description>&lt;P&gt;Hello,&lt;BR /&gt;I'm trying to ingest data from my RDS instance and set up change data capture on Databricks. Everything on the Postgres side has been done so that replication is possible, but CDC is still greyed out. I have asked Claude, Gemini, and ChatGPT. I haven't got any concrete help. I read that it's an issue of Databricks needing to enable Lakeflow Connect for Postgres or something like.&lt;/P&gt;&lt;P&gt;&lt;span class="lia-inline-image-display-wrapper lia-image-align-inline" image-alt="Wola_0-1789753526825.png" style="width: 400px;"&gt;&lt;img src="https://community.databricks.com/t5/image/serverpage/image-id/31291i14A36CC3F04565B3/image-size/medium?v=v2&amp;amp;px=400" role="button" title="Wola_0-1789753526825.png" alt="Wola_0-1789753526825.png" /&gt;&lt;/span&gt;&lt;/P&gt;&lt;P&gt;P.S.: I might not have explained it well, but I really need help with it.&lt;/P&gt;</description>
      <pubDate>Fri, 18 Sep 2026 17:49:31 GMT</pubDate>
      <guid>https://community.databricks.com/t5/data-engineering/data-ingestion-setting-up-a-connector-for-postgres-with-cdc/m-p/169117#M56042</guid>
      <dc:creator>Wola</dc:creator>
      <dc:date>2026-09-18T17:49:31Z</dc:date>
    </item>
    <item>
      <title>Re: Data ingestion: Setting up a connector for Postgres with CDC enabled.</title>
      <link>https://community.databricks.com/t5/data-engineering/data-ingestion-setting-up-a-connector-for-postgres-with-cdc/m-p/169238#M56063</link>
      <description>&lt;P&gt;Hi Wola,&lt;/P&gt;&lt;P&gt;Don't worry about your Postgres side, that's most likely not the problem. The greyed-out option is coming from Databricks. The PostgreSQL CDC connector in Lakeflow Connect is still in Public Preview, and this one isn't self-service: the docs say "Contact your Databricks account team to request access." There is nothing to flip on the Previews page. Until the workspace is enrolled, the wizard only lets you pick Query-based capture, which is exactly what your screenshot shows.&lt;BR /&gt;&lt;A href="https://docs.databricks.com/aws/en/ingestion/lakeflow-connect/postgresql" target="_blank"&gt;https://docs.databricks.com/aws/en/ingestion/lakeflow-connect/postgresql&lt;/A&gt;&lt;/P&gt;&lt;P&gt;So step one is a message to your account team asking to enable the PostgreSQL connector preview for your workspace.&lt;/P&gt;&lt;P&gt;While you wait, two things worth checking. The CDC pipeline needs serverless compute enabled in the workspace (the gateway runs on classic compute, the pipeline itself is serverless), and it only replicates from a primary instance, read replicas won't work.&lt;BR /&gt;&lt;A href="https://docs.databricks.com/aws/en/ingestion/lakeflow-connect/postgresql-pipeline" target="_blank"&gt;https://docs.databricks.com/aws/en/ingestion/lakeflow-connect/postgresql-pipeline&lt;/A&gt;&lt;/P&gt;&lt;P&gt;And the RDS checklist from the docs, just to compare with what you've done: rds.logical_replication = 1 in the parameter group, a replication user with the rds_replication role and the REPLICATION attribute, a publication, a replication slot created with pgoutput (the only plugin supported), and REPLICA IDENTITY set to DEFAULT or FULL on every table. The slot name and publication name are what the wizard asks for later, in the Database setup step.&lt;BR /&gt;&lt;A href="https://docs.databricks.com/aws/en/ingestion/lakeflow-connect/postgresql-source-setup" target="_blank"&gt;https://docs.databricks.com/aws/en/ingestion/lakeflow-connect/postgresql-source-setup&lt;/A&gt;&lt;BR /&gt;&lt;A href="https://docs.databricks.com/aws/en/ingestion/lakeflow-connect/postgresql-limits" target="_blank"&gt;https://docs.databricks.com/aws/en/ingestion/lakeflow-connect/postgresql-limits&lt;/A&gt;&lt;/P&gt;&lt;P&gt;If you can't wait, Query-based capture does work with Postgres today, no preview needed. Just know what you're getting: it doesn't read the WAL, it runs on a schedule and picks up changes through a cursor column (a timestamp or increasing id) per table. You won't see intermediate states of a row, and hard deletes are only tracked in Beta. Fine for a lot of use cases, but it isn't real CDC.&lt;BR /&gt;&lt;A href="https://docs.databricks.com/aws/en/ingestion/lakeflow-connect/query-based-overview" target="_blank"&gt;https://docs.databricks.com/aws/en/ingestion/lakeflow-connect/query-based-overview&lt;/A&gt;&lt;/P&gt;&lt;P&gt;Good luck with it.&lt;/P&gt;</description>
      <pubDate>Sun, 20 Sep 2026 12:00:29 GMT</pubDate>
      <guid>https://community.databricks.com/t5/data-engineering/data-ingestion-setting-up-a-connector-for-postgres-with-cdc/m-p/169238#M56063</guid>
      <dc:creator>ThomazNeto</dc:creator>
      <dc:date>2026-09-20T12:00:29Z</dc:date>
    </item>
    <item>
      <title>Re: Data ingestion: Setting up a connector for Postgres with CDC enabled.</title>
      <link>https://community.databricks.com/t5/data-engineering/data-ingestion-setting-up-a-connector-for-postgres-with-cdc/m-p/169252#M56068</link>
      <description>&lt;P&gt;Thank you very much!&lt;/P&gt;&lt;P&gt;It's an individual account. So when I saw the message about reaching out to your Databricks account team so they can reach out to your Databricks rep, I was a bit confused, as I'm the account owner. I'm learning and didn't want to use a free account, so I set up a paid individual account.&lt;/P&gt;&lt;P&gt;Regarding the serverless compute for CDC, I read that serverless was enabled by default in most workspaces, but with the aid of Claude, I set a policy for the classic compute indicating the instance family and other things needed, if at all, classic compute was needed.&lt;/P&gt;&lt;P&gt;Once again, thank you very much.&lt;BR /&gt;I was already frustrated.&lt;/P&gt;</description>
      <pubDate>Sun, 20 Sep 2026 18:29:54 GMT</pubDate>
      <guid>https://community.databricks.com/t5/data-engineering/data-ingestion-setting-up-a-connector-for-postgres-with-cdc/m-p/169252#M56068</guid>
      <dc:creator>Wola</dc:creator>
      <dc:date>2026-09-20T18:29:54Z</dc:date>
    </item>
    <item>
      <title>Re: Data ingestion: Setting up a connector for Postgres with CDC enabled.</title>
      <link>https://community.databricks.com/t5/data-engineering/data-ingestion-setting-up-a-connector-for-postgres-with-cdc/m-p/169254#M56069</link>
      <description>&lt;P&gt;Quick question, being that my account is an individual account, how do I go about this&amp;nbsp;&lt;SPAN&gt;"Contact your Databricks account team to request access"&amp;nbsp;&lt;/SPAN&gt;&lt;/P&gt;</description>
      <pubDate>Sun, 20 Sep 2026 18:32:25 GMT</pubDate>
      <guid>https://community.databricks.com/t5/data-engineering/data-ingestion-setting-up-a-connector-for-postgres-with-cdc/m-p/169254#M56069</guid>
      <dc:creator>Wola</dc:creator>
      <dc:date>2026-09-20T18:32:25Z</dc:date>
    </item>
    <item>
      <title>Re: Data ingestion: Setting up a connector for Postgres with CDC enabled.</title>
      <link>https://community.databricks.com/t5/data-engineering/data-ingestion-setting-up-a-connector-for-postgres-with-cdc/m-p/172286#M56617</link>
      <description>&lt;P&gt;Greetings &lt;a href="https://community.databricks.com/t5/user/viewprofilepage/user-id/258220"&gt;@Wola&lt;/a&gt;, I did some digging and here is what I found.&lt;/P&gt;
&lt;P&gt;&lt;a href="https://community.databricks.com/t5/user/viewprofilepage/user-id/245135"&gt;@ThomazNeto&lt;/a&gt; already pointed you at the two things that matter most: the PostgreSQL connector for Lakeflow Connect is in Public Preview and needs workspace enrollment, and logical replication only works against a primary instance. I'll build on that and answer your account team question directly.&lt;/P&gt;
&lt;P&gt;First, the greyed-out CDC option. That's the preview gate, not your RDS configuration or your compute policy. The policy you built isn't wasted (the ingestion gateway runs on classic compute and will use it), but no change on your side flips that selector. Only enrollment does.&lt;/P&gt;
&lt;P&gt;On "contact your account team" when you are the account: with a pay-as-you-go individual account you don't have a named rep, and the docs don't offer a self-service path for this preview. Three things to try, in order:&lt;/P&gt;
&lt;OL&gt;
&lt;LI&gt;Check Settings &amp;gt; Previews in your workspace, and the account console since you're the admin. Most previews aren't toggleable there and I don't expect this one is, but it takes ten seconds to rule out.&lt;/LI&gt;
&lt;LI&gt;Open a support case at &lt;A href="https://help.databricks.com" target="_blank"&gt;https://help.databricks.com&lt;/A&gt; if your plan includes support. Ask specifically to have your workspace enrolled in the PostgreSQL Lakeflow Connect Public Preview, and include your workspace ID.&lt;/LI&gt;
&lt;LI&gt;If you don't have a support entitlement, use the Contact Us form at &lt;A href="https://www.databricks.com/company/contact" target="_blank"&gt;https://www.databricks.com/company/contact&lt;/A&gt;. It routes to sales, which sounds like the wrong door, but that's who can put a preview enablement request in front of the right people.&lt;/LI&gt;
&lt;/OL&gt;
&lt;P&gt;Once enrolled, the workspace side is mostly done already: Unity Catalog and serverless enabled (you're right that serverless is on by default in most workspaces, and the ingestion pipeline runs there while the gateway runs on classic). Confirm you also have &lt;CODE&gt;CREATE CONNECTION&lt;/CODE&gt; on the metastore and &lt;CODE&gt;USE CATALOG&lt;/CODE&gt;, &lt;CODE&gt;USE SCHEMA&lt;/CODE&gt;, &lt;CODE&gt;CREATE TABLE&lt;/CODE&gt;, and &lt;CODE&gt;CREATE VOLUME&lt;/CODE&gt; on the target.&lt;/P&gt;
&lt;P&gt;On the RDS side, you've likely covered most of this, but these are the ones that trip people up:&lt;/P&gt;
&lt;UL&gt;
&lt;LI&gt;&lt;CODE&gt;rds.logical_replication = 1&lt;/CODE&gt; in the parameter group, reboot, and &lt;CODE&gt;SHOW wal_level;&lt;/CODE&gt; returns &lt;CODE&gt;logical&lt;/CODE&gt;.&lt;/LI&gt;
&lt;LI&gt;PostgreSQL 13 or later, primary instance only.&lt;/LI&gt;
&lt;LI&gt;A dedicated replication user with &lt;CODE&gt;ALTER USER ... WITH REPLICATION&lt;/CODE&gt;, &lt;CODE&gt;GRANT rds_replication TO your_user;&lt;/CODE&gt; (RDS-specific and easy to miss), plus &lt;CODE&gt;CONNECT&lt;/CODE&gt;, &lt;CODE&gt;USAGE&lt;/CODE&gt;, and &lt;CODE&gt;SELECT&lt;/CODE&gt; on the tables.&lt;/LI&gt;
&lt;LI&gt;A publication created as a superuser or table owner, then a replication slot created with the &lt;CODE&gt;pgoutput&lt;/CODE&gt; plugin (the only one Databricks supports). Publication first, then slot, and create the slot while running as the replication user via &lt;CODE&gt;SET ROLE&lt;/CODE&gt;.&lt;/LI&gt;
&lt;LI&gt;&lt;CODE&gt;REPLICA IDENTITY FULL&lt;/CODE&gt; on tables without a primary key or with large TEXT/BYTEA columns; &lt;CODE&gt;DEFAULT&lt;/CODE&gt; otherwise.&lt;/LI&gt;
&lt;LI&gt;&lt;CODE&gt;max_slot_wal_keep_size&lt;/CODE&gt; set to a finite value so an idle slot can't bloat WAL.&lt;/LI&gt;
&lt;LI&gt;Security group allows inbound 5432 from the Databricks workspace.&lt;/LI&gt;
&lt;/UL&gt;
&lt;P&gt;When the button lights up, the wizard is Data Ingestion &amp;gt; Add data &amp;gt; PostgreSQL, and it asks for the slot and publication names on the Database setup page.&lt;/P&gt;
&lt;P&gt;If you need something while you wait, Lakeflow Connect's query-based connector can ingest from PostgreSQL with no WAL or CDC setup. It runs on a schedule using a cursor column, so it captures the latest state of changed rows rather than every intermediate change, and it isn't a substitute for log-based CDC. Hard-delete tracking is in Beta and needs API configuration.&lt;/P&gt;
&lt;P&gt;Hang in there. You did the hard part on the Postgres side; the piece that's blocking you is the one piece you can't do yourself.&lt;/P&gt;
&lt;P&gt;References:&lt;/P&gt;
&lt;UL&gt;
&lt;LI&gt;PostgreSQL ingestion connector: &lt;A href="https://docs.databricks.com/aws/en/ingestion/lakeflow-connect/postgresql" target="_blank"&gt;https://docs.databricks.com/aws/en/ingestion/lakeflow-connect/postgresql&lt;/A&gt;&lt;/LI&gt;
&lt;LI&gt;Ingest data from PostgreSQL (workspace requirements and UI steps): &lt;A href="https://docs.databricks.com/aws/en/ingestion/lakeflow-connect/postgresql-pipeline" target="_blank"&gt;https://docs.databricks.com/aws/en/ingestion/lakeflow-connect/postgresql-pipeline&lt;/A&gt;&lt;/LI&gt;
&lt;LI&gt;Configure PostgreSQL for ingestion (source setup, RDS notes): &lt;A href="https://docs.databricks.com/aws/en/ingestion/lakeflow-connect/postgresql-source-setup" target="_blank"&gt;https://docs.databricks.com/aws/en/ingestion/lakeflow-connect/postgresql-source-setup&lt;/A&gt;&lt;/LI&gt;
&lt;LI&gt;PostgreSQL connector limitations: &lt;A href="https://docs.databricks.com/aws/en/ingestion/lakeflow-connect/postgresql-limits" target="_blank"&gt;https://docs.databricks.com/aws/en/ingestion/lakeflow-connect/postgresql-limits&lt;/A&gt;&lt;/LI&gt;
&lt;LI&gt;Query-based connectors: &lt;A href="https://docs.databricks.com/aws/en/ingestion/lakeflow-connect/query-based-overview" target="_blank"&gt;https://docs.databricks.com/aws/en/ingestion/lakeflow-connect/query-based-overview&lt;/A&gt;&lt;/LI&gt;
&lt;/UL&gt;
&lt;P&gt;Regards, Louis.&lt;/P&gt;</description>
      <pubDate>Thu, 08 Oct 2026 12:41:42 GMT</pubDate>
      <guid>https://community.databricks.com/t5/data-engineering/data-ingestion-setting-up-a-connector-for-postgres-with-cdc/m-p/172286#M56617</guid>
      <dc:creator>Louis_Frolio</dc:creator>
      <dc:date>2026-10-08T12:41:42Z</dc:date>
    </item>
  </channel>
</rss>

