Data Engineering

Forum Posts

Sorted by:

by Serhii • Contributor

11-01-2022 10:13:46 AM

2517 Views
0 replies
0 kudos

Horovod Databricks Job - custom module not found error

We have used the following example to successfully create a distributed deep learning training notebook https://www.databricks.com/blog/2022/09/07/accelerating-your-deep-learning-pytorch-lightning-databricks.html that works as expected.We now want to...

Data Engineering

2517 Views
0 replies
0 kudos

11-01-2022 10:13:46 AM

by User16857281869 • Databricks Employee

06-17-2021 1:34:22 AM

1976 Views
1 replies
0 kudos

How do I benefit from parallelisation when doing machine learning?

There are in principle four distinct ways of using parallelisation when doing machine learning. Any combination of these can speed up the whole pipeline significantly.1) Using spark distributed processing in feature engineering 2) When the data set...

Data Engineering

1976 Views
1 replies
0 kudos

06-17-2021 1:34:22 AM

View Replies

Latest Reply

sean_owen
Databricks Employee

06-17-2021 11:25:11 AM

0 kudos

Good summary! yes those are the main strategies I can think of.

0 kudos

06-17-2021 11:25:11 AM

by User16788317466 • Databricks Employee

06-07-2021 10:59:10 AM

1569 Views
1 replies
0 kudos

When can Horovod be used for an ML problem?

Data Engineering

1569 Views
1 replies
0 kudos

06-07-2021 10:59:10 AM

View Replies

Latest Reply

User16788317466
Databricks Employee

06-07-2021 11:02:24 AM

0 kudos

Only when you have a gradient-descent problem. Pytorch and Tensorflow are the only candidate frameworks to use here. When using Horovod, start with single node, multi-GPU and measure training performance. If this is not sufficient, look at a multi-no...

0 kudos

06-07-2021 11:02:24 AM

Databricks Community

Horovod Databricks Job - custom module not found error

How do I benefit from parallelisation when doing machine learning?

When can Horovod be used for an ML problem?