cancel
Showing results for 
Search instead for 
Did you mean: 
Community Articles
Dive into a collaborative space where members like YOU can exchange knowledge, tips, and best practices. Join the conversation today and unlock a wealth of collective wisdom to enhance your experience and drive success.
cancel
Showing results for 
Search instead for 
Did you mean: 

Forum Posts

Tushar_Parekar
by Databricks Employee
  • 941 Views
  • 0 replies
  • 0 kudos

Learning Series | Databricks Performance Optimization

Databricks Academy offers the free Databricks Performance Optimization course to help data engineers improve workload performance on the Databricks Data Intelligence Platform. As part of the Advanced Data Engineering with Databricks series, it focuse...

Learning Series (800 x 800 px) (6) (1).png
  • 941 Views
  • 0 replies
  • 0 kudos
Tushar_Parekar
by Databricks Employee
  • 558 Views
  • 0 replies
  • 2 kudos

Solution Accelerator Series | Product Quality Inspection

Advances in deep learning have made computer vision a practical approach for product quality inspection. The Product Quality Inspection Solution Accelerator shows how organizations can implement an end-to-end computer vision pipeline, build and train...

product-quality-inspection-inbody-graphic.jpg
  • 558 Views
  • 0 replies
  • 2 kudos
balajij8
by Esteemed Contributor II
  • 1517 Views
  • 0 replies
  • 2 kudos

Real Time Healthcare Analytics with Databricks Zerobus and Lakehouse RT via Reyden

Real time capabilities in healthcare is a critical factor in care outcomes and operational efficiency. From streaming continuous vitals from IoT-enabled care monitors to routing instantaneous telemetry from wearable medical devices, modern healthcare...

1_LnW43wHnfrJH1ugbMJWk6Q
  • 1517 Views
  • 0 replies
  • 2 kudos
AmitDECopilot
by Contributor
  • 1489 Views
  • 3 replies
  • 2 kudos

Resolved! How would you design a Spark pipeline to process billions of records efficiently?

Interview Question:Many people start with the row count.I would start with the architecture.Billions of records are not new in enterprise data engineering. The real challenge is designing a pipeline that runs predictably, efficiently, and within SLA....

  • 1489 Views
  • 3 replies
  • 2 kudos
Latest Reply
savlahanish27
Databricks Partner
  • 2 kudos

Fair point - if it's genuinely append-only with no corrections, you're right, you're already touching the minimum amount of data each run. So, the problem really is what you originally said: how do you process a genuinely huge incoming batch efficien...

  • 2 kudos
2 More Replies
Ashwin_DSA
by Databricks Employee
  • 1457 Views
  • 1 replies
  • 3 kudos

Managed vs External Storage in Unity Catalog on Azure: Where Should Your Data Live?

  I'm currently working with a banking customer migrating from Hive Metastore to Unity Catalog. While planning their catalog layout, one of their platform engineers asked the question that prompted this post: for their schemas and volumes, should th...

Ashwin_DSA_0-1782751472981.png Ashwin_DSA_1-1782751512839.png Ashwin_DSA_2-1782751535604.png Ashwin_DSA_3-1782751554657.png
  • 1457 Views
  • 1 replies
  • 3 kudos
Latest Reply
savlahanish27
Databricks Partner
  • 3 kudos

Hi Ashwin,Really clean write-up - the schema override demo with the actual DESCRIBE DETAIL output is what makes it click. Most posts on this topic stop at "here's the theory," this one actually shows it happening.Something similar came up on a SAP HA...

  • 3 kudos
sagarpgowda7777
by New Contributor II
  • 1095 Views
  • 0 replies
  • 1 kudos

MCP on Databricks

Here’s What Nobody Tells You.A hands-on look at Genie MCP and DBSQL MCP — what works, what doesn’t, and when to skip MCP entirely.Let me start with something most MCP content skips. MCP servers don’t just expose tools. They expose three things — tool...

  • 1095 Views
  • 0 replies
  • 1 kudos
AmitDECopilot
by Contributor
  • 3060 Views
  • 0 replies
  • 1 kudos

LTAP: What Databricks New Transactional-Analytical Architecture Means for Data Engineers

For years, enterprise data architecture has followed a familiar pattern.An application writes customer orders, account updates, inventory changes, or transactions into an operational database.Then data engineering takes over.We capture changes throug...

  • 3060 Views
  • 0 replies
  • 1 kudos
AmitDECopilot
by Contributor
  • 685 Views
  • 0 replies
  • 0 kudos

From Business Requirements to Lakeflow Pipelines: A Governed Metadata-Driven Delivery Pattern

IntroductionDifferent organizations use different names for these artifacts: business requirements, mapping specifications, source-to-target mappings, data contracts, transformation rules, or semantic definitions. The name matters less than the goal:...

  • 685 Views
  • 0 replies
  • 0 kudos
emma_s
by Databricks Employee
  • 2432 Views
  • 2 replies
  • 4 kudos

Create an MCP for Azure DevOps To Use With Genie Code

Overview Prompted by a customer question, I wanted to see what was possible in terms of MCP integration into Genie Code, in order to try this out I decided to look at Azure Dev Ops, as it's a common workflow to want to see your tickets alongside the ...

Screenshot 2026-03-25 at 15.55.10.png
Community Articles
azure devops
Genie Code
MCP
  • 2432 Views
  • 2 replies
  • 4 kudos
Latest Reply
HectorOviedoM
New Contributor II
  • 4 kudos

Hi Emma how are you? is there any solution to make this agent avaliable for multiple users? the main restriction is that the path will be linked to a single user/service princpal.thanks!

  • 4 kudos
1 More Replies
Ashwin_DSA
by Databricks Employee
  • 2201 Views
  • 1 replies
  • 3 kudos

Reading Spark UI: A Repeatable Guide to Finding Performance Bottlenecks

A question came up in the community recently that I thought deserved more than a short answer. The question was around how to build a reliable investigation sequence for slow Spark jobs, specifically when symptoms overlap. A long-running stage with ...

Ashwin_DSA_0-1782413970227.png Ashwin_DSA_1-1782413989929.png Ashwin_DSA_2-1782414031747.png Ashwin_DSA_3-1782414084869.png
  • 2201 Views
  • 1 replies
  • 3 kudos
Latest Reply
srinivasu_nalla
New Contributor II
  • 3 kudos

Thanks for this! Very insightful and detailed. Your sequence for diagnosing slow Spark jobs when symptoms overlap is exactly what I needed. Bookmarked this for my team.

  • 3 kudos
Brahmareddy
by Esteemed Contributor II
  • 2025 Views
  • 11 replies
  • 32 kudos

A small song for the Databricks Community

One Question Can Light the SparkMy Databricks journey started in 2022 with simple interest, curiosity, and a dream to learn more.At that time, I was just trying to understand the platform, follow the updates, learn from others, and slowly build my co...

  • 2025 Views
  • 11 replies
  • 32 kudos
Latest Reply
Advika
Community Manager
  • 32 kudos

My turn!  (couldn't let this disappear from the homepage that quickly ).This one is officially going down in Community history now, @Brahmareddy .

  • 32 kudos
10 More Replies
Ashwin_DSA
by Databricks Employee
  • 4560 Views
  • 5 replies
  • 12 kudos

Databricks Multi-Table Transactions - Part 1

If you've ever worked on an insurance data warehouse, or really any warehouse where data arrives from different systems at different times, you know the pain of keeping things in sync. I spent years building data warehouses for a property and casual...

claim-wrapup-flow.png before-after-transactions.png Part1 Cover Pic.png
  • 4560 Views
  • 5 replies
  • 12 kudos
Latest Reply
antoalphi_db
Databricks Partner
  • 12 kudos

I know I'm reading this a bit late, but it's absolutely worth the read. Excellent write-up!Please share link for part-2

  • 12 kudos
4 More Replies
AmitDECopilot
by Contributor
  • 497 Views
  • 0 replies
  • 2 kudos

Building a Metadata-Driven ETL Framework on Databricks

 Stop Writing ETL Code for Every New PipelineAs organizations modernize their data platforms, one challenge continues to appear repeatedly:Every new source requires another ETL job.Every new business rule requires another SQL update.Every schema chan...

  • 497 Views
  • 0 replies
  • 2 kudos
szymon_dybczak
by Esteemed Contributor III
  • 1570 Views
  • 0 replies
  • 1 kudos

DataFlint on Databricks - the Open Source Spark UI Upgrade Apache Spark Has Needed for Years

 IntroductionApache Spark has become one of the most widely adopted engines for large-scale data processing. Its appeal is easy to understand: it supports batch processing, streaming workloads, feature engineering, machine learning pipelines, and lar...

szymon_dybczak_0-1782288499357.png szymon_dybczak_1-1782288499410.png szymon_dybczak_2-1782288499481.png szymon_dybczak_3-1782288499412.png
  • 1570 Views
  • 0 replies
  • 1 kudos
Labels