IBM DB2
Connect an IBM DB2 database to Lakehousecat using a connection URI.
Connection Fields
| Field | Required | Description |
|---|---|---|
| Connection URI | Yes | Full connection string. Example: db2+ibm_db://user:password@host:50000/database |
| Use SSH | No | Enable if the database is behind an SSH bastion host. See SSH Tunneling. |
Prerequisites
- Admin or Builder role in Lakehousecat
- A Db2 user with
SELECTaccess to the target schemas and tables - Default port: 50000
- The IBM Db2 driver, provisioned once through a job definition. It is not preinstalled — see Provision the IBM driver below.
- Your Lakehousecat instance must be deployed with
architecture: "amd64". IBM does not ship a Linux/ARM64 driver for Db2, so this datasource type is unavailable on anarm64instance. See thearchitecturefield in your deployment configuration. Every other datasource type is unaffected by this — it is specific to Db2.
Notes
- Use a read-only database user to follow the principle of least privilege.
- Incremental Load is supported; use a timestamp column (e.g.,
UPDATED_AT) as the merge key. - Db2 schema names are typically uppercase — ensure filter settings use the correct casing.
Provision the IBM driver (one-time)
IBM Db2 is fully supported, but IBM's Db2 CLI driver is IBM's software under IBM's terms, so
Lakehousecat does not ship it. You fetch it once through a job definition, and you accept IBM's
licence terms while doing so. IBM's licence files are stored unchanged in the downloaded archive
under clidriver/license/.
- Import the
db2_driver_provisioningjob definition, setLHC_ACCEPT_VENDOR_EULAto"yes"and run it. The job downloads IBM'sibm_dbwheel from PyPI, verifies its SHA-256 checksum and stores the driver in File Storage. It prints aLHC_VENDOR_DRIVER_FILE_ID. - Import
db2_loadanddb2_schema_validate. In each, setDATASOURCE_ID,LHC_VENDOR_DRIVER_FILE_ID(from step 1) andLHC_ACCEPT_VENDOR_EULAto"yes". - Load through
db2_load, also on a schedule. Validate the connection and discover the schema throughdb2_schema_validate.
Requirements:
- Outbound HTTPS from the provisioning job to
pypi.organdfiles.pythonhosted.org. Without internet access, build the archive yourself (clidriver/andibm_db.libs/side by side) and upload it withlhc files upload. - An
amd64instance. Onarm64the provisioning job stops with an explanation.
Validate and load calls made directly against the platform API don't work for Db2 and return a message that points to these job definitions. After a first successful load, the Filters tab reads the schema from your warehouse and needs no driver.
For the general mechanism, see Vendor Drivers.
Next Steps
After provisioning the driver, load the datasource through the db2_load job definition, then
open the Operations tab to trigger semantic extraction.