<?xml version="1.0" encoding="UTF-8"?>
<rss xmlns:content="http://purl.org/rss/1.0/modules/content/" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:rdf="http://www.w3.org/1999/02/22-rdf-syntax-ns#" xmlns:taxo="http://purl.org/rss/1.0/modules/taxonomy/" version="2.0">
  <channel>
    <title>topic Re: How to use the same job cluster in diferents job runs inside the one workflow in Data Engineering</title>
    <link>https://community.databricks.com/t5/data-engineering/how-to-use-the-same-job-cluster-in-diferents-job-runs-inside-the/m-p/75214#M34896</link>
    <description>&lt;P&gt;Hi,&lt;BR /&gt;&lt;BR /&gt;If I understand correctly, you are hoping to reduce overall job execution time by reducing the Cloud Service Provider instance provisioning time. Is that correct?&lt;BR /&gt;&lt;BR /&gt;If so, you may want to consider:&lt;/P&gt;
&lt;UL&gt;
&lt;LI&gt;Using a Pool of instances: &lt;A href="https://docs.databricks.com/en/compute/pool-index.html" target="_blank"&gt;https://docs.databricks.com/en/compute/pool-index.html&lt;/A&gt;&lt;/LI&gt;
&lt;LI&gt;Enable Serverless Compute for Workflows (currently in Public Preview at time of writing)&lt;A href="https://docs.databricks.com/en/workflows/jobs/run-serverless-jobs.html" target="_blank"&gt;&amp;nbsp;https://docs.databricks.com/en/workflows/jobs/run-serverless-jobs.html&lt;/A&gt;&amp;nbsp;&lt;/LI&gt;
&lt;LI&gt;If the above considerations do not work for you, given that Jobs Clusters are designed to start when a job begins and stop when a job completes, you may consider creating a persistent All-Purpose cluster&lt;/LI&gt;
&lt;/UL&gt;
&lt;P&gt;Hope this helps.&lt;/P&gt;</description>
    <pubDate>Thu, 20 Jun 2024 15:15:28 GMT</pubDate>
    <dc:creator>brockb</dc:creator>
    <dc:date>2024-06-20T15:15:28Z</dc:date>
    <item>
      <title>How to use the same job cluster in diferents job runs inside the one workflow</title>
      <link>https://community.databricks.com/t5/data-engineering/how-to-use-the-same-job-cluster-in-diferents-job-runs-inside-the/m-p/75211#M34894</link>
      <description>&lt;P&gt;I created a Workflow with notebooks and some job runs, but I would to use only one job cluster to run every job runs, without creating a new job cluster for each job run. Because I didn't want to increase the execution time with each new job cluster instantiated for each job run!&lt;/P&gt;&lt;P&gt;&amp;nbsp;&lt;/P&gt;&lt;P&gt;&amp;nbsp;&lt;/P&gt;</description>
      <pubDate>Thu, 20 Jun 2024 14:38:24 GMT</pubDate>
      <guid>https://community.databricks.com/t5/data-engineering/how-to-use-the-same-job-cluster-in-diferents-job-runs-inside-the/m-p/75211#M34894</guid>
      <dc:creator>Eiki</dc:creator>
      <dc:date>2024-06-20T14:38:24Z</dc:date>
    </item>
    <item>
      <title>Re: How to use the same job cluster in diferents job runs inside the one workflow</title>
      <link>https://community.databricks.com/t5/data-engineering/how-to-use-the-same-job-cluster-in-diferents-job-runs-inside-the/m-p/75214#M34896</link>
      <description>&lt;P&gt;Hi,&lt;BR /&gt;&lt;BR /&gt;If I understand correctly, you are hoping to reduce overall job execution time by reducing the Cloud Service Provider instance provisioning time. Is that correct?&lt;BR /&gt;&lt;BR /&gt;If so, you may want to consider:&lt;/P&gt;
&lt;UL&gt;
&lt;LI&gt;Using a Pool of instances: &lt;A href="https://docs.databricks.com/en/compute/pool-index.html" target="_blank"&gt;https://docs.databricks.com/en/compute/pool-index.html&lt;/A&gt;&lt;/LI&gt;
&lt;LI&gt;Enable Serverless Compute for Workflows (currently in Public Preview at time of writing)&lt;A href="https://docs.databricks.com/en/workflows/jobs/run-serverless-jobs.html" target="_blank"&gt;&amp;nbsp;https://docs.databricks.com/en/workflows/jobs/run-serverless-jobs.html&lt;/A&gt;&amp;nbsp;&lt;/LI&gt;
&lt;LI&gt;If the above considerations do not work for you, given that Jobs Clusters are designed to start when a job begins and stop when a job completes, you may consider creating a persistent All-Purpose cluster&lt;/LI&gt;
&lt;/UL&gt;
&lt;P&gt;Hope this helps.&lt;/P&gt;</description>
      <pubDate>Thu, 20 Jun 2024 15:15:28 GMT</pubDate>
      <guid>https://community.databricks.com/t5/data-engineering/how-to-use-the-same-job-cluster-in-diferents-job-runs-inside-the/m-p/75214#M34896</guid>
      <dc:creator>brockb</dc:creator>
      <dc:date>2024-06-20T15:15:28Z</dc:date>
    </item>
  </channel>
</rss>

