- 9462 Views
- 0 replies
- 0 kudos
hey i have this error from a while : Cannot resolve "(needed_skill_id = needed_skill_id)" due to data type mismatch: the left and right operands of the binary operator have incompatible types ("STRING" and "ARRAY<STRING>"). SQLSTATE: 42K09;and these ...
- 9462 Views
- 0 replies
- 0 kudos
- 7641 Views
- 0 replies
- 0 kudos
Hi, Is there any way to save a workbook without losing the macros in databricks?
- 7641 Views
- 0 replies
- 0 kudos
by
essura
• New Contributor II
- 3727 Views
- 1 replies
- 1 kudos
Hi there,We are trying to setup up a docker image for our dbt execution, primarily to improve execution speed, but also to simplify deployment (we are using a private repos for both the dbt project and some of the dbt packages).It seems to work curre...
- 3727 Views
- 1 replies
- 1 kudos
- 3104 Views
- 1 replies
- 0 kudos
I am creating cluster using rest api call but every-time it is creating all purpose cluster. Is there a way to create job cluster and run notebook using python code?
- 3104 Views
- 1 replies
- 0 kudos
Latest Reply
job_cluster_key
string
[ 1 .. 100 ] characters
^[\w\-\_]+$
If job_cluster_key, this task is executed reusing the cluster specified in job.settings.job_clusters.Create a new job | Jobs API | REST API reference | Databricks on AWS
- 4135 Views
- 1 replies
- 1 kudos
Hi guys,You have any idea how can I do a groupBy without aggregation (Pyspark API)like: df.groupBy('field1', 'field2', 'field3') My target is make a group but in this case is not necessary count records or aggregationThank you
- 4135 Views
- 1 replies
- 1 kudos
Latest Reply
df.select("field1","field2","field3").distinct()do you mean get distinct rows for selected column?
by
Innov
• New Contributor
- 1428 Views
- 0 replies
- 0 kudos
Looking for some help. If anyone has worked with nested json file in Databricks notebook. I am trying to parse nested json file to get coordinates and use that to create polygon for footprint. Do I need to read it as txt? How can I use the Databricks...
- 1428 Views
- 0 replies
- 0 kudos
- 2189 Views
- 1 replies
- 0 kudos
So i have this nested data with more than 200+columns and i have extracted this data into json file when i use the below code to read the json files, if in data there are few columns which have no value at all it doest inclued those columns in schema...
- 2189 Views
- 1 replies
- 0 kudos
Latest Reply
replying to my above questionwe cannot use inferschema on streaming table we need to externally specify schema can anyone please suggest a way to write data in nested form to streaming table and if this is possible?
- 4666 Views
- 3 replies
- 1 kudos
I am trying to connect AS Cube with Databricks notebook but unfortunately didn't find any solution yet. is there any possible way to connect AS cube with databricks notebook? if yes can someone please guide me
- 4666 Views
- 3 replies
- 1 kudos
Latest Reply
I am able to connect Azure analysis services using Azure Analysis services rest api. is yours on-prem?
2 More Replies
- 6487 Views
- 4 replies
- 5 kudos
Hi, everyone. I just recently started using Databricks on Azure so my question is probably very basic but I am really stuck right now.I need to capture some streaming metrics (number of input rows and their time) so I tried using the Spark Rest Api ...
- 6487 Views
- 4 replies
- 5 kudos
Latest Reply
hi @Roberto Baldrez​ ,if you think that @Gaurav Rupnar​ solved your question, then please select it as best response to it can be moved to the top of the topic and it will help more users in the future.Thank you
3 More Replies
- 4685 Views
- 2 replies
- 2 kudos
I have created a DLT pipeline which reads data from json files which are stored in databricks volume and puts data into streaming table This was working fine.when i tried to read the data that is inserted into the table and compare the values with t...
- 4685 Views
- 2 replies
- 2 kudos
Latest Reply
Keep your DLT code separate from your comparison code, and run your comparison code once your DLT data has been ingested.
1 More Replies
- 2430 Views
- 1 replies
- 1 kudos
Hello,We are in the process of migrating to Unity Catalog. So, can I know how to automate the process of Refactoring the Notebooks to Unity Catalog.
- 2430 Views
- 1 replies
- 1 kudos
Latest Reply
Hi @Avinash_Narala There is no one-click solution to refactor all table names notebooks with UC's three level namespaces. At least, manual updating table names is required during the migration process.One option is you can you search feature. Search ...
by
valjas
• New Contributor III
- 10886 Views
- 3 replies
- 0 kudos
We are working on creating a new databricks workspace for external entities. We have disabled Cluster and Warehouse creation permission but the external users are still able to create Jobs and job clusters. Is there a way to revoke Job creation permi...
- 10886 Views
- 3 replies
- 0 kudos
Latest Reply
It permits cluster creation during Workflow/Job/DLT pipeline creation. However, when attempting to start any of these, it fails with a 'Not authorized to create compute' error. Please try it and inform me of the outcome
2 More Replies
- 1617 Views
- 1 replies
- 0 kudos
I am facing this issue from long time but so far there is no solution. I have delta table. My bronze layer is picking up the old files (mostly 8 days old file) randomly. My source of files is azure blob storage.
- 1617 Views
- 1 replies
- 0 kudos
Latest Reply
Hey @jaimeperry12345 I will need more information to direct you in the right direction: Confirm the behavior: Double-check that your Delta table is indeed reading 8-day-old files randomly. Provide any logs or error messages you have regarding this.Ex...
- 8688 Views
- 3 replies
- 1 kudos
Hello,I have a parquet file test.parquet in the volume volume_ext_test. Tried to create an external table as below, it failed and says it "is not a valid URI".create table catalog_managed.schema_test.tbl_vol asselect * from parquet.`/Volumes/catalog_...
- 8688 Views
- 3 replies
- 1 kudos
Latest Reply
Hi @PaulineX As per the documentation, you cannot use volume for storing table data. It's for loading, storing and accessing files. You cannot use volumes as a location for tables. Volumes are intended for path-based data access only. Use tables for ...
2 More Replies
by
Ru
• Databricks Partner
- 3133 Views
- 0 replies
- 0 kudos
Hi Databricks Community,We've encountered an issue with setting a column only on insert when using DLT's apply_changes CDC merge functionality. It's important to note that this capability is available when using the regular Delta merge operation, spe...
- 3133 Views
- 0 replies
- 0 kudos