Databricks¶
Connect Databricks to track cluster compute costs, job spend, and optimization opportunities in CloudVerse AI.
What you get¶
- Cluster and job cost tracking
- Pipeline and workload cost analysis
- Optimization recommendations for idle or oversized clusters
Prerequisites¶
- Databricks workspace with admin access
- A personal access token or service principal credentials
Connect Databricks¶
- Go to Settings > Integrations.
- Select Databricks.
- Enter a connection name — for example,
databricks-production. - Enter your Databricks Workspace URL and Access Token.
- Finish the flow and wait for initial data ingestion.
Verify the connection¶
- Open the Databricks connection detail page in Settings > Integrations.
- Open Data & AI Platform > Data Platform and select Databricks to verify cluster and job data.
Data expected¶
| Data | Used for |
|---|---|
| Workspace and cluster metadata | Inventory, ownership, and cost attribution. |
| Job and pipeline activity | Workload-level cost analysis and optimization. |
| Query or warehouse facts, when configured | Data Platform query, table, and warehouse insights. |
| Cost and usage data | Spend trend, allocation, and optimization calculations. |
Best practices¶
- Use a service principal token rather than a personal access token for production connections.
- Rotate the token on schedule and update it in CloudVerse promptly.
- Confirm workspace IDs and billing export configuration during acceptance.
Troubleshooting¶
See Integration setup issues if data does not appear after setup.