<?xml version="1.0" encoding="UTF-8"?>
<rss xmlns:content="http://purl.org/rss/1.0/modules/content/" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:rdf="http://www.w3.org/1999/02/22-rdf-syntax-ns#" xmlns:taxo="http://purl.org/rss/1.0/modules/taxonomy/" version="2.0">
  <channel>
    <title>article 💡 ML Training Tip Of The Week #4: Speed up your ML workload with one toggle in Technical Blog</title>
    <link>https://community.databricks.com/t5/technical-blog/ml-training-tip-of-the-week-4-speed-up-your-ml-workload-with-one/ba-p/112237</link>
    <description>&lt;P&gt;&lt;SPAN&gt;Machine learning workload can be time-consuming due to training data processing such as ETL and feature engineering and iterative model training steps.&lt;/SPAN&gt;&lt;/P&gt;
&lt;P&gt;&lt;SPAN&gt;Recent ML Runtime releases have made both faster by 2X with a single toggle during the cluster creation:&lt;/SPAN&gt;&lt;/P&gt;
&lt;UL&gt;
&lt;LI style="font-weight: 400;" aria-level="1"&gt;&lt;SPAN&gt;Photon (available in MLR 15.2 or above), which is the C++ query execution engine, can speed up Spark SQL and Spark DataFrame that are commonly used in ETL and feature engineering by&lt;/SPAN&gt;&lt;STRONG&gt; 1.5~2X&lt;/STRONG&gt;&lt;/LI&gt;
&lt;LI style="font-weight: 400;" aria-level="1"&gt;&lt;SPAN&gt;Graviton instances (available in MLR 15.4 LTS or above) powered by ARM-based CPUs designed by AWS can speed up XGBoost, LightGBM, etc. algorithms by up to &lt;/SPAN&gt;&lt;STRONG&gt;1.5X.&lt;/STRONG&gt;&lt;/LI&gt;
&lt;/UL&gt;
&lt;P&gt;&lt;SPAN&gt;To enable Photon, select “Use Photon Acceleration” when creating a cluster as shown below:&lt;/SPAN&gt;&lt;/P&gt;
&lt;P&gt;&lt;span class="lia-inline-image-display-wrapper lia-image-align-inline" image-alt="linyuan_0-1741673531511.png" style="width: 400px;"&gt;&lt;img src="https://community.databricks.com/t5/image/serverpage/image-id/15330i15B17267A0FFEEA5/image-size/medium?v=v2&amp;amp;px=400" role="button" title="linyuan_0-1741673531511.png" alt="linyuan_0-1741673531511.png" /&gt;&lt;/span&gt;&lt;/P&gt;
&lt;P&gt;&lt;SPAN&gt;To use ARM-based Graviton instance on AWS,&amp;nbsp; search for “&lt;/SPAN&gt;&lt;I&gt;&lt;SPAN&gt;7g&lt;/SPAN&gt;&lt;/I&gt;&lt;SPAN&gt;” in the instance types as shown below:&lt;/SPAN&gt;&lt;/P&gt;
&lt;P&gt;&lt;span class="lia-inline-image-display-wrapper lia-image-align-inline" image-alt="linyuan_1-1741673531505.png" style="width: 400px;"&gt;&lt;img src="https://community.databricks.com/t5/image/serverpage/image-id/15329i71A15761933FCDD3/image-size/medium?v=v2&amp;amp;px=400" role="button" title="linyuan_1-1741673531505.png" alt="linyuan_1-1741673531505.png" /&gt;&lt;/span&gt;&lt;/P&gt;
&lt;P&gt;&lt;SPAN&gt;Read more about the use case of Photon and Graviton in the Databricks blog posts:&lt;/SPAN&gt;&lt;/P&gt;
&lt;UL&gt;
&lt;LI style="font-weight: 400;" aria-level="1"&gt;&lt;A href="https://www.databricks.com/blog/accelerate-feature-engineering-photon" target="_blank" rel="noopener"&gt;&lt;SPAN&gt;Accelerate Feature Engineering With Photon | Databricks Blog&lt;/SPAN&gt;&lt;/A&gt;&lt;SPAN&gt;&amp;nbsp;&lt;/SPAN&gt;&lt;/LI&gt;
&lt;LI style="font-weight: 400;" aria-level="1"&gt;&lt;A href="https://www.databricks.com/blog/unlock-faster-machine-learning-graviton" target="_blank" rel="noopener"&gt;&lt;SPAN&gt;Unlock Faster Machine Learning with Graviton | Databricks Blog&lt;/SPAN&gt;&lt;/A&gt;&lt;SPAN&gt;&amp;nbsp;&lt;/SPAN&gt;&lt;/LI&gt;
&lt;/UL&gt;</description>
    <pubDate>Tue, 11 Mar 2025 06:16:50 GMT</pubDate>
    <dc:creator>lin-yuan</dc:creator>
    <dc:date>2025-03-11T06:16:50Z</dc:date>
    <item>
      <title>💡 ML Training Tip Of The Week #4: Speed up your ML workload with one toggle</title>
      <link>https://community.databricks.com/t5/technical-blog/ml-training-tip-of-the-week-4-speed-up-your-ml-workload-with-one/ba-p/112237</link>
      <description>&lt;P&gt;&lt;SPAN&gt;Machine learning workload can be time-consuming due to training data processing such as ETL and feature engineering and iterative model training steps.&lt;/SPAN&gt;&lt;/P&gt;
&lt;P&gt;&lt;SPAN&gt;Recent ML Runtime releases have made both faster by 2X with a single toggle during the cluster creation:&lt;/SPAN&gt;&lt;/P&gt;
&lt;UL&gt;
&lt;LI style="font-weight: 400;" aria-level="1"&gt;&lt;SPAN&gt;Photon (available in MLR 15.2 or above), which is the C++ query execution engine, can speed up Spark SQL and Spark DataFrame that are commonly used in ETL and feature engineering by&lt;/SPAN&gt;&lt;STRONG&gt; 1.5~2X&lt;/STRONG&gt;&lt;/LI&gt;
&lt;LI style="font-weight: 400;" aria-level="1"&gt;&lt;SPAN&gt;Graviton instances (available in MLR 15.4 LTS or above) powered by ARM-based CPUs designed by AWS can speed up XGBoost, LightGBM, etc. algorithms by up to &lt;/SPAN&gt;&lt;STRONG&gt;1.5X.&lt;/STRONG&gt;&lt;/LI&gt;
&lt;/UL&gt;
&lt;P&gt;&lt;SPAN&gt;To enable Photon, select “Use Photon Acceleration” when creating a cluster as shown below:&lt;/SPAN&gt;&lt;/P&gt;
&lt;P&gt;&lt;span class="lia-inline-image-display-wrapper lia-image-align-inline" image-alt="linyuan_0-1741673531511.png" style="width: 400px;"&gt;&lt;img src="https://community.databricks.com/t5/image/serverpage/image-id/15330i15B17267A0FFEEA5/image-size/medium?v=v2&amp;amp;px=400" role="button" title="linyuan_0-1741673531511.png" alt="linyuan_0-1741673531511.png" /&gt;&lt;/span&gt;&lt;/P&gt;
&lt;P&gt;&lt;SPAN&gt;To use ARM-based Graviton instance on AWS,&amp;nbsp; search for “&lt;/SPAN&gt;&lt;I&gt;&lt;SPAN&gt;7g&lt;/SPAN&gt;&lt;/I&gt;&lt;SPAN&gt;” in the instance types as shown below:&lt;/SPAN&gt;&lt;/P&gt;
&lt;P&gt;&lt;span class="lia-inline-image-display-wrapper lia-image-align-inline" image-alt="linyuan_1-1741673531505.png" style="width: 400px;"&gt;&lt;img src="https://community.databricks.com/t5/image/serverpage/image-id/15329i71A15761933FCDD3/image-size/medium?v=v2&amp;amp;px=400" role="button" title="linyuan_1-1741673531505.png" alt="linyuan_1-1741673531505.png" /&gt;&lt;/span&gt;&lt;/P&gt;
&lt;P&gt;&lt;SPAN&gt;Read more about the use case of Photon and Graviton in the Databricks blog posts:&lt;/SPAN&gt;&lt;/P&gt;
&lt;UL&gt;
&lt;LI style="font-weight: 400;" aria-level="1"&gt;&lt;A href="https://www.databricks.com/blog/accelerate-feature-engineering-photon" target="_blank" rel="noopener"&gt;&lt;SPAN&gt;Accelerate Feature Engineering With Photon | Databricks Blog&lt;/SPAN&gt;&lt;/A&gt;&lt;SPAN&gt;&amp;nbsp;&lt;/SPAN&gt;&lt;/LI&gt;
&lt;LI style="font-weight: 400;" aria-level="1"&gt;&lt;A href="https://www.databricks.com/blog/unlock-faster-machine-learning-graviton" target="_blank" rel="noopener"&gt;&lt;SPAN&gt;Unlock Faster Machine Learning with Graviton | Databricks Blog&lt;/SPAN&gt;&lt;/A&gt;&lt;SPAN&gt;&amp;nbsp;&lt;/SPAN&gt;&lt;/LI&gt;
&lt;/UL&gt;</description>
      <pubDate>Tue, 11 Mar 2025 06:16:50 GMT</pubDate>
      <guid>https://community.databricks.com/t5/technical-blog/ml-training-tip-of-the-week-4-speed-up-your-ml-workload-with-one/ba-p/112237</guid>
      <dc:creator>lin-yuan</dc:creator>
      <dc:date>2025-03-11T06:16:50Z</dc:date>
    </item>
  </channel>
</rss>

