Skip to main content
Databricks logo

Is Databricks Down?

No — Databricks is up

Reachable from all 8 checked regions

Average response time: 247ms

Last checked · checks run every 6 hours

Official status page: https://status.databricks.com

Databricks uptime

100%
Last 7 days
100%
Last 30 days
100%
Last 90 days
138ms
Avg response, 30 days

Measured from multiple regions every 6 hours. Percentages count only checks that returned an availability answer — 55 days measured so far. A dash means that window does not yet hold enough measured days to publish a figure.

30-day history

55-day clean streak
Aug 22: 100.00% uptime, 32 checks
Aug 23: 100.00% uptime, 32 checks
Aug 24: 100.00% uptime, 32 checks
Aug 25: 100.00% uptime, 32 checks
Aug 26: 100.00% uptime, 32 checks
Aug 27: 100.00% uptime, 32 checks
Aug 28: 100.00% uptime, 32 checks
Aug 29: 100.00% uptime, 32 checks
Aug 30: 100.00% uptime, 32 checks
Aug 31: 100.00% uptime, 32 checks
Sep 1: 100.00% uptime, 32 checks
Sep 2: 100.00% uptime, 30 checks
Sep 3: 100.00% uptime, 32 checks
Sep 4: 100.00% uptime, 32 checks
Sep 5: 100.00% uptime, 30 checks
Sep 6: 100.00% uptime, 21 checks
Sep 7: 100.00% uptime, 30 checks
Sep 8: 100.00% uptime, 16 checks
Sep 9: 100.00% uptime, 34 checks
Sep 10: 100.00% uptime, 32 checks
Sep 11: 100.00% uptime, 32 checks
Sep 12: 100.00% uptime, 31 checks
Sep 13: 100.00% uptime, 24 checks
Sep 14: 100.00% uptime, 32 checks
Sep 15: 100.00% uptime, 32 checks
Sep 16: 100.00% uptime, 31 checks
Sep 17: 100.00% uptime, 32 checks
Sep 18: 100.00% uptime, 32 checks
Sep 19: 100.00% uptime, 32 checks
Sep 20: 100.00% uptime, 16 checks
Aug 22 Today
No downtime Partial Downtime Not measurable No data

Reachability by region

Each region runs its own request from a different part of the world. A service can be up for one continent and down for another, which is usually the first sign of a routing or CDN problem.

ams
201ms
DNS 60ms TCP 2ms TLS 14ms TTFB 189ms
arn
144ms
DNS 44ms TCP 1ms TLS 14ms TTFB 126ms
nrt
340ms
DNS 251ms TCP 1ms TLS 13ms TTFB 332ms
ord
246ms
DNS 102ms TCP 2ms TLS 13ms TTFB 235ms
sin
302ms
DNS 167ms TCP 6ms TLS 24ms TTFB 283ms
sjc
279ms
DNS 159ms TCP 5ms TLS 19ms TTFB 254ms
syd
296ms
DNS 197ms TCP 0ms TLS 10ms TTFB 285ms
yyz
172ms
DNS 69ms TCP 1ms TLS 13ms TTFB 157ms

What Databricks does

Databricks is a data and AI platform built around the lakehouse model, combining notebooks, SQL warehouses, pipelines and machine learning on top of cloud storage. Data teams schedule production jobs on it, so an incident usually shows up first as overnight pipelines that failed rather than as people unable to log in.

What an outage looks like

Clusters fail to start or hang in a pending state, so notebooks attach to nothing. Scheduled jobs fail at launch and the failure is only noticed the next morning. SQL warehouses will not resume, leaving dashboards timing out. Unity Catalog problems block table access while compute is healthy, and notebooks may open while every command queues.

What to do about it

Check status.databricks.com, which reports Compute, Databricks SQL, Notebooks, Unity Catalog and Lakeflow separately across many cloud regions, and also tracks dependencies including AWS EC2, S3 and Route 53. Confirm your workspace region first. Let failed jobs be rerun deliberately rather than by an automatic retry storm against a platform that is still recovering.

Is it down for everyone, or just you?

If this page says Databricks is up but it is not loading for you, the problem is between you and them. Run a check against any URL from up to 18 cities across three regions to find out where it breaks.

Test it yourself

Related services

Databricks outage FAQ

Is Databricks down or is my cluster just slow to start?
Cluster startup normally takes several minutes while cloud instances are provisioned, which is easy to mistake for a hang. Check status.databricks.com for your region and the Compute component. If the region is healthy, look at instance availability and quota in your cloud account, which cause more startup failures than Databricks incidents do.
Which region should I check on the Databricks status page?
The one your workspace runs in, visible in the workspace URL. Databricks lists many regions across cloud providers and incidents are usually confined to one. The status page also tracks upstream dependencies such as AWS EC2 and S3, which is useful when the underlying cause is the cloud provider rather than Databricks itself.
Will failed jobs rerun automatically after an outage?
Only if you configured retries, and aggressive retries against a recovering platform can make things worse. Jobs that failed at launch usually need a deliberate rerun once the status page is clear. Check for partial writes first: a job that failed midway may have committed some output, and rerunning it blindly can duplicate data.
Why can I open a notebook but not run anything?
The workspace interface and the compute layer are separate. Serving a notebook is lightweight, while running a command needs a cluster, so the editor loads during incidents that stop execution entirely. Check the Compute component rather than concluding from a working notebook that the platform is healthy and the problem is your code.
Does a Databricks outage put my data at risk?
Data sits in your own cloud storage rather than inside Databricks, so an incident affects the ability to process it, not its durability. The exposure is to jobs interrupted mid-write, which can leave partial output. Checking the state of tables written by any job that failed is worth doing before rerunning anything.

How we measure this

  • We request Databricks's public endpoint every 6 hours from Fly.io regions across six continents — the most recent check ran from 8 of them.
  • A region counts as down only when it gets no usable HTTP response. A 403 or 429 means the origin answered and refused us, which we report as blocked, never as an outage.
  • A single failing region is treated as probe noise. We only change the verdict when two consecutive cycles agree.
  • Response times average only the regions that actually served the page, so a timeout never inflates the number.
  • Where Databricks publishes an official status feed we read it too. An all-clear from the vendor can soften an unconfirmed degradation; a vendor-declared outage only worsens our verdict when our own checks corroborate it.

Get alerted when Databricks goes down

This page refreshes every 6 hours. Your own monitors run as often as every 30 seconds, from up to 18 cities across three regions, and tell you the moment something breaks.

Start Free Monitoring
Free plan available No credit card required