cancel
Showing results for 
Search instead for 
Did you mean: 
cancel
Showing results for 
Search instead for 
Did you mean: 
DATA + AI World Tour 2026

REGISTRATION IS NOW OPEN The Data + AI World Tour is heading your way A free, one-day event on how organizations are building AI they can trust to deliver real outcomes - by grounding models in their own enterprise data and business context. THE...

  • 385 Views
  • 0 replies
  • 0 kudos
Tuesday
Databricks Advanced Learning Festival: September 16 - October 14 2026

Databricks Advanced Learning Festival September 16 – October 14, 2026 Join us for a four-week event dedicated to learning, upskilling, and advancing your career in data engineering, analytics, machine learning, and generative AI. Whether you are...

  • 23760 Views
  • 30 replies
  • 23 kudos
2 weeks ago
Virtual Event | Closing the AI Context Gap: How to Teach AI How Your Business Actually Runs

Virtual Event Time: AMER Sept 23 / 9 AM PTEMEA Sept 24 / 9 AM BST / 10 AM CESTAPJ Sept 24 / 9:30 AM IST / 12 PM SGT / 1 PM JST/KST / 2 PM AEST (Japanese and Korean captions available) Register for the Workshop Note: Marking RSVP on this Community ...

  • 274 Views
  • 0 replies
  • 1 kudos
Wednesday
Databricks EMEA Learning Festival | September 29-30

Databricks EMEA Learning Festival Hands On Virtual Training 29–30 September 2026 · 09:00 BST / 10:00 CEST · Fully virtual · Free to join Ready to accelerate your journey on Data & AI? Join us for the EMEA Learning Festival.  A two-day, hands-on vir...

  • 2296 Views
  • 0 replies
  • 2 kudos
4 weeks ago
Databricks Community Champion - July 2026 - Emma Stowell

Our Community Champion Program has always been about recognizing individuals who consistently show up to support others, share their expertise, and strengthen the Databricks Community. Alongside our customers, we’re equally proud to recognize the inc...

  • 3244 Views
  • 5 replies
  • 6 kudos
a month ago
Solution Accelerator Series | Subscriber Churn Prediction

Subscriber attrition can be difficult to anticipate. The Subscriber Churn Prediction Solution Accelerator shows how to analyze behavioral data, identify subscribers at increased risk of cancellation, and use machine learning to estimate churn likelih...

  • 247 Views
  • 0 replies
  • 3 kudos
Tuesday
🌟 Community Pulse: Your Weekly Roundup! August 24 – 30, 2026

The Weekly Digest • August 24 – 30 Community Pulse A new week, a fresh wave of articles, answers, and discussions. Here's your weekly catch-up. This Week's Community Stars   The members who showed up and made a difference this week : @balajij8...

  • 230 Views
  • 0 replies
  • 1 kudos
Thursday

Community Activity

foodpackagiing
by > New Contributor
  • 20 Views
  • 1 replies
  • 0 kudos

Food Packaging Direct

Food Packaging Direct is one of the top suppliers of biodegradable food packing UK. With our extensive experience in distributing high-quality packaging supplies to consumers, we know what works and what does not for food packaging materials. Whether...

  • 20 Views
  • 1 replies
  • 0 kudos
Latest Reply
yashikab
New Contributor III
  • 0 kudos

why have you posted this here? 

  • 0 kudos
youssefmrini
by Databricks Employee
  • 28 Views
  • 0 replies
  • 1 kudos

What’s new in Databricks - August 2026

  ️Data Engineering Databricks Spark Declarative Improvements You can refresh Materialized Views and Streaming Tables even when the pipeline owner or run as identity is subject to an ABAC. Materialized views and Streaming tables refreshes in Datab...

Copy of Databricks for Developers - 2 person Thumbnails (11).png youssefmrini_0-1788774135131.jpeg youssefmrini_1-1788774135192.png youssefmrini_2-1788774134870.png
  • 28 Views
  • 0 replies
  • 1 kudos
Marbricks
by > New Contributor
  • 132 Views
  • 3 replies
  • 0 kudos

Using Temporary Functions

Hi everyone,I'm currently doing my professional internship and working on a reporting project that integrates with Databricks through Unity Catalog functions.The reporting application is external to Databricks and uses Data Sources, Parameters, Resul...

  • 132 Views
  • 3 replies
  • 0 kudos
Latest Reply
data_pulse
New Contributor
  • 0 kudos

@Marbricks Follow this approach: define the function’s interface first, then connect the real data. Use a temporary function for testing within the same session, and a persistent Unity Catalog function in a development schema for testing with the ext...

  • 0 kudos
2 More Replies
VibinRoy_C
by > New Contributor III
  • 117 Views
  • 2 replies
  • 0 kudos

Databricks-Native AI Agent for Job Incident Detection, RCA & Safe Remediation

I’m exploring an architecture for a Databricks-native AI Agent for intelligent Spark job incident detection, root cause analysis (RCA), and safe remediation, and I would love your technical feedback.The ProblemA common operational challenge is the "S...

  • 117 Views
  • 2 replies
  • 0 kudos
Latest Reply
Satyasai
New Contributor
  • 0 kudos

Exactly Similar situtation , i have crated my own using Pythopn , VetcorDB and Knowledge graph, you can simulate same or Hire me , i can able to do it for you LOL..See this Video on my Linked In

  • 0 kudos
1 More Replies
TyreseR19
by > New Contributor II
  • 532 Views
  • 2 replies
  • 0 kudos

Follow-up Regarding Missing 50% Certification Voucher – Learning Festival

Dear Databricks Team,I hope you are doing well.I am writing to follow up regarding the 50% certification voucher associated with the Databricks Learning Festival held from 15 June to 6 July.I successfully completed the courses under the Data Analyst ...

  • 532 Views
  • 2 replies
  • 0 kudos
Latest Reply
Advika
Community Manager
  • 0 kudos

Hello @TyreseR19, Could you please ensure that both required courses for the Data Analyst learning path were fully completed during the Learning Festival event window? Please verify that every component of each course is marked as completed, includin...

  • 0 kudos
1 More Replies
anushnagesh
by > New Contributor
  • 44 Views
  • 1 replies
  • 0 kudos

Tech Companies & Lakeflow Connect

Hello fellow techies,For the past few days I have been doing some R&D wrt Lakeflow connect. I came to the understanding that, Lakeflow Ingestion Pipeline and SQL server connection have a 1:1 mapping. This would mean that as number of servers scales n...

  • 44 Views
  • 1 replies
  • 0 kudos
Latest Reply
Satyasai
New Contributor
  • 0 kudos

Most modern enterprise data platforms favor managed ingestion (Lakeflow Connect) for standard relational databases like SQL Server. The operational complexity of monitoring 50 managed pipelines is significantly lower than maintaining custom CDC code,...

  • 0 kudos
wsm777
by > New Contributor
  • 179 Views
  • 1 replies
  • 0 kudos

Where can I find the Databricks Fundamentals Notebook?

Hi! I'm currently doing the Databricks Fundamentals Learning Plan and therés a couple of topics (Catalogs and Notebooks) where to follow up with the training I need a notebook name Databricks Fundamentals. In both videos, the video makes reference to...

wsm777_0-1788118093330.png
  • 179 Views
  • 1 replies
  • 0 kudos
Latest Reply
pikimu
Visitor
  • 0 kudos

did you ever figure this out, looking for those specific notebooks myself? I'm not sure if this is helpful, but in the Databricks free edition home page, there is a Databricks fundamentals course, which covers similar things (it is not exactly the sa...

  • 0 kudos
Dolly0503
by > New Contributor III
  • 78 Views
  • 3 replies
  • 1 kudos

How to extract table-level execution time and resource allocation within a multi-table Job?

Hi everyone,I am working on calculating accurate compute costs and execution times for individual tables within our Databricks environment, but I am running into an issue with metric granularity.Currently, we are fetching execution data based on job_...

  • 78 Views
  • 3 replies
  • 1 kudos
Latest Reply
Satyasai
New Contributor
  • 1 kudos

Try this link https://community.databricks.com/t5/data-engineering/how-to-calculate-cost-of-each-table-for-the-specific-databricks/td-p/167086 

  • 1 kudos
2 More Replies
gowri_databrick
by > New Contributor
  • 72 Views
  • 2 replies
  • 0 kudos

Understanding Parquet File Storage for Large Datasets

Hi everyone,I’m learning about Parquet files and how they are used in Databricks for storing large datasets.I’m trying to understand how column-based storage works in a practical situation.For example, suppose an e-commerce company has 500 million or...

  • 72 Views
  • 2 replies
  • 0 kudos
Latest Reply
balajij8
Esteemed Contributor II
  • 0 kudos

@gowri_databrick Parquet's columnar storage organizes data by column rather than by row - all values for order_date are stored together, all order_amount values are stored together and so on. In your commerce scenario with 500 million records, when t...

  • 0 kudos
1 More Replies
SantiNath_Dey
by > Contributor II
  • 108 Views
  • 2 replies
  • 2 kudos

MetaData Framework for Multi-Level Silver Layer PK/FK Creation

Hi Team,Our source data originates from MongoDB as hierarchical JSON and is ingested into our Silver layer, where it is already normalized into tabular structures. However, primary and foreign key constraints are not enforced during ingestion. We mus...

SantiNath_Dey_0-1788703344844.png
  • 108 Views
  • 2 replies
  • 2 kudos
Latest Reply
Ashwin_DSA
Databricks Employee
  • 2 kudos

Hi @SantiNath_Dey, Databricks has solid building blocks for this, and there's even a purpose-built metadata framework you can lean on. At its core, your pipeline needs to do three things...generate surrogate keys at each level, resolve parent referen...

  • 2 kudos
1 More Replies
Ashwin_DSA
by Databricks Employee
  • 3128 Views
  • 1 replies
  • 3 kudos

Speed Up Data Warehouse Migration Validation

You've spent months planning the migration, the pipelines are built, and data is flowing into the new platform. Then comes the hard part: proving the data actually matches. Business validation and reconciliation is where data migrations fail. Not be...

Ashwin_DSA_9-1778960752402.png Ashwin_DSA_10-1778960814388.png Ashwin_DSA_11-1778960845693.png Ashwin_DSA_12-1778960874222.png
  • 3128 Views
  • 1 replies
  • 3 kudos
Latest Reply
SantiNath_Dey
Contributor II
  • 3 kudos

thank you for your response

  • 3 kudos
AmitDECopilot
by > Contributor
  • 237 Views
  • 1 replies
  • 0 kudos

I Stopped Sending Every Data Engineering Task to an LLM - A Cost Aware Routing Pattern on Databricks

AI is becoming part of almost every data engineering workflow.We are asking AI to generate SQL, PySpark pipelines, data-quality rules, unit tests, technical specifications and even complete project structures. The results can be impressive, but there...

  • 237 Views
  • 1 replies
  • 0 kudos
Latest Reply
suryaprayaga
Contributor
  • 0 kudos

Yes that's true. There are hundreds (if not 100s) of tasks that can be done without depending on the agents. Databricks indeed has DQX for governance and data quality. Instead of relying on smaller activities (especially when they are repeating in na...

  • 0 kudos
gowri_databrick
by > New Contributor
  • 103 Views
  • 2 replies
  • 0 kudos

What is a Checkpoint in Structured Streaming?

Hi everyone,I’m learning about Structured Streaming in Databricks and came across checkpoints.I understand that checkpoints are used to keep track of the progress of a streaming query, but I’d like to understand their purpose more clearly.For example...

  • 103 Views
  • 2 replies
  • 0 kudos
Latest Reply
balajij8
Esteemed Contributor II
  • 0 kudos

@gowri_databrick Checkpoints in Structured Streaming serve as the ledger and recovery mechanism for streaming queries. It tracks which data has been processed and successfully written, enabling exactly once processing guarantees. A checkpoint contain...

  • 0 kudos
1 More Replies
srikanthp24
by > New Contributor
  • 182 Views
  • 1 replies
  • 0 kudos

Lakeflow connect Ingestion pipeline notification for gateway pipeline

Hi Guys,As you guys know that when we are building lake flow connect ingestion pipeline in UI. The pipeline consist both gateway pipeline and ingestion pipeline together. We have notification for the ingestion pipeline but not for gateway pipeline. I...

srikanthp24_0-1788542943980.png
  • 182 Views
  • 1 replies
  • 0 kudos
Latest Reply
Ashwin_DSA
Databricks Employee
  • 0 kudos

Hi @srikanthp24, Yeah... this is something others have run into as well. When you create a Lakeflow Connect ingestion pipeline via the UI wizard, the "Schedules and notifications" step (step 6) only configures alerts for that pipeline. The gateway pi...

  • 0 kudos
IM_01
by > Valued Contributor
  • 175 Views
  • 1 replies
  • 2 kudos

Lakeflow SDP Append Flow

Hi All,I'm using  append_flow to ingest data into the target table. Before returning the dataframe, I compute few column trnasformations, but those values aren't being calculated correctly (or: aren't showing up at all)so the append flow should not i...

  • 175 Views
  • 1 replies
  • 2 kudos
Latest Reply
Ashwin_DSA
Databricks Employee
  • 2 kudos

Hi @IM_01,   append_flow is not limited to just reading a source table. You can (and should) include column transformations, filters, and any other DataFrame operations inside the function decorated with @dp.append_flow. The function simply needs to ...

  • 2 kudos
Welcome to the Databricks Community!

Once you are logged in, you will be ready to post content, ask questions, participate in discussions, earn badges, and more.

Spend a few minutes exploring Get Started Resources, Learning Paths, Certifications, and Platform Discussions.

Join Learning Events here in the Community.

Connect with peers through Databricks User Groups and learn more about Community Events happening near you. We’re excited to see you get involved.

Top Kudoed Authors

Latest from our Blog

GeoGenie 2: Ask your map a question in plain English

A year ago we put a natural-language layer on a 3D map. This time we rebuilt the whole thing on Lakebase and PostGIS, wired Genie to the same tables, and made the map answer back, drawn region and all...

Image
566Views 2kudos