Skip to content

Databricks

Connect Databricks to track cluster compute costs, job spend, and optimization opportunities in CloudVerse AI.

What you get

  • Cluster and job cost tracking
  • Pipeline and workload cost analysis
  • Optimization recommendations for idle or oversized clusters

Prerequisites

  • Databricks workspace with admin access
  • A personal access token or service principal credentials

Connect Databricks

  1. Go to Settings > Integrations.
  2. Select Databricks.
  3. Enter a connection name — for example, databricks-production.
  4. Enter your Databricks Workspace URL and Access Token.
  5. Finish the flow and wait for initial data ingestion.

Verify the connection

  1. Open the Databricks connection detail page in Settings > Integrations.
  2. Open Data & AI Platform > Data Platform and select Databricks to verify cluster and job data.

Data expected

Data Used for
Workspace and cluster metadata Inventory, ownership, and cost attribution.
Job and pipeline activity Workload-level cost analysis and optimization.
Query or warehouse facts, when configured Data Platform query, table, and warehouse insights.
Cost and usage data Spend trend, allocation, and optimization calculations.

Best practices

  • Use a service principal token rather than a personal access token for production connections.
  • Rotate the token on schedule and update it in CloudVerse promptly.
  • Confirm workspace IDs and billing export configuration during acceptance.

Troubleshooting

See Integration setup issues if data does not appear after setup.