Skip to main content
Version: 0.0.36

Backup

The backup process is a critical component of the Lakehousecat framework, ensuring the automated safeguarding of all storage units. These backups are managed and executed through job definitions within the Lakehousecat framework.

What is Backed Up?​

The backup processes include the regular safeguarding of the following components:

  • Relational Database: Ensures the protection of structured data essential for system operations.
  • VectorStore: Secures vector data used for semantic and analytical queries.
  • Analytical Backend: Protects analytical data and processes to maintain data integrity.
  • Object Storage Data: Safeguards unstructured data stored in the object storage.
  • EBS Storage Data: Backs up Elastic Block Store data accessed by various services within the Lakehousecat framework.

How Does the Backup Process Work?​

The backups are consolidated and stored in a central archive within an S3 bucket. This approach ensures secure and centralized data storage. The backup process is implemented using the AWS Backup System, providing a reliable and scalable solution for data protection.

Benefits of the Backup System​

  • Automation: Minimizes manual intervention through automated job definitions.
  • Reliability: Leverages the AWS Backup System for robust and scalable backups.
  • Data Integrity: Protects all critical data sources within the Lakehousecat framework.
  • Centralization: Stores all backups in a central S3 bucket for easy management and recovery.

This backup system is an integral part of the administration within the Lakehousecat framework, ensuring that your data is always protected and recoverable.