Skip to main content
Version: 0.0.42

Apache Druid

Connect an Apache Druid cluster to Lakehousecat using a connection URI. Druid is a real-time OLAP database designed for high-performance analytics on large event-driven and time-series datasets. It exposes a SQL endpoint that Lakehousecat uses for schema discovery and query execution.

Connection Fields​

FieldRequiredDescription
Connection URIYesFull connection string pointing to the Druid SQL endpoint. Example: druid+pydruid://host:8082/druid/v2/sql/
Use SSHNoEnable if the Druid Broker is behind an SSH bastion host. When enabled, SSH settings fields appear: Host, Port, Username, Password, Private Key, and Private Key Passphrase. See SSH Tunneling.

URI Format​

druid+pydruid://host:port/druid/v2/sql/

Example (local development):

druid+pydruid://localhost:8082/druid/v2/sql/

Prerequisites​

  • Admin or Builder role in Lakehousecat
  • A running Apache Druid cluster with the SQL API enabled (available since Druid 0.14)
  • Network connectivity from the Lakehousecat backend to the Druid Broker (port 8082 by default)

Notes​

  • The connection targets the Druid Broker's SQL endpoint at /druid/v2/sql/.
  • Druid exposes all ingested datasources as SQL tables in the druid schema.
  • Authentication (Basic Auth, TLS) can be configured via additional URI parameters if your cluster has security enabled.
  • SSH tunnel support depends on network configuration — enable the Use SSH toggle if access requires a bastion host.

Next Steps​

After creating the datasource, go to the Operations tab to trigger semantic extraction.