---
title: "Workflows: Crawl Databricks"
url: "https://atlan.com/demos/crawl-databricks/"
description: "Ready to build your metadata estate? Discover how to effortlessly set up and manage your Databricks Crawler in this hands-on demo!"
format: "Arcade"
demo_url: "https://demo.arcade.software/dLLx0nBXJvWIImHEufYA"
content_purpose: ["Product Overview"]
target_persona: ["IT Administrator"]
journey_stage: ["S2 - Discovery", "C1 - Onboarding", "C2 - First Value"]
use_case_context: ["Training"]
product: ["Enterprise Data Graph - Integration"]
content_type: "interactive demo step summary"
transcript_source: "arcade-steps"
---

# Workflows: Crawl Databricks

Step-by-step summary of the interactive demo at https://atlan.com/demos/crawl-databricks/

<!-- Body: the Arcade flow's hotspot captions, in order, with navigation-only captions removed. -->

Ready to build your metadata estate? Discover how to effortlessly set up and manage your Databricks Crawler in this hands-on demo!

## What the demo walks through

1. Begin by adding a new workflow
2. The Databricks Assets package allows us to crawl metadata from Databricks
3. Click the button to create a new connection with Databricks
4. (Optional) Or change to offline to load manually extracted data Choose direct extraction
5. Authenticate with a personal access token... ... or use AWS service principal authentication ... or use Azure service principal authentication
6. Enter connection details and credentials The HTTP path determines whether you're using a SQL endpoint or an interactive cluster Ensure the connection works
7. Once successful, proceed to the next step
8. Enter connection details and credentials Ensure the connection works
9. Once successful, proceed to the next step
10. Enter connection details and credentials Ensure the connection works
11. Once successful, proceed to the next step
12. Enter details for the location of the manually extracted data in S3 Proceed to the next step
13. Enter a connection name (Optional) Add or remove connection admins (Optional) Disable data access (Optional) Set the maximum number of rows that can be returned by a query
14. (Optional) Change extraction method (Optional) Limit the metadata to include Run preflight checks before running the crawler
15. Run immediately or schedule to run periodically
16. (Optional) Limit the metadata to include (Optional) Import Databricks tags
17. If importing Databricks tags, select the SQL warehouse you've configured Run preflight checks before running the crawler
18. Run immediately or schedule to run periodically
19. (Optional) Limit the metadata to include (Optional) Import Databricks tags
20. If importing Databricks tags, select the SQL warehouse you've configured Run preflight checks before running the crawler
21. Run immediately or schedule to run periodically
22. Run preflight checks before running the crawler Run immediately or schedule to run periodically
23. When run immediately, you'll be redirected to the monitoring page
