ScyllaDB University LIVE, FREE Virtual Training Event | March 21
Register for Free
ScyllaDB Documentation Logo Documentation
  • Deployments
    • Cloud
    • Server
  • Tools
    • ScyllaDB Manager
    • ScyllaDB Monitoring Stack
    • ScyllaDB Operator
  • Drivers
    • CQL Drivers
    • DynamoDB Drivers
    • Supported Driver Versions
  • Resources
    • ScyllaDB University
    • Community Forum
    • Tutorials
Install
Search Ask AI
ScyllaDB Docs ScyllaDB Manual ScyllaDB for Administrators Procedures Backup and Restore Procedures Restore from a Backup and Incremental Backup
For AI agents: a documentation index is available at https://docs.scylladb.com/manual/branch-2026.3/llms.txt. A Markdown version of this page is at https://docs.scylladb.com/manual/branch-2026.3/operating-scylla/procedures/backup-restore/restore.md.

Restore from a Backup and Incremental Backup¶

Restoring a keyspace from a backup requires all snapshot files of the tables, and (if available) incremental backup files taken after the snapshot. Before restoring from backup, the table data must be truncated, making sure that the existing data does not overwrite the restored data.

Note

For cluster-wide backup and restore, see the ScyllaDB Manager documentation.

Choosing a restore method¶

ScyllaDB supports several restore methods. Choose the one that matches where your backup files are and whether the cluster topology changed since the backup was taken:

  • Restore from object storage - restore SSTables backed up to S3-compatible object storage with nodetool backup. Runs on a running cluster and works regardless of cluster topology changes since the backup.

  • Restore with load and stream - upload backed-up SSTable files to a running cluster; the data is streamed to the nodes owning it. Works regardless of cluster topology changes since the backup.

  • Restore to an identical cluster - copy snapshot files back in place and restart the nodes. Requires a cluster with the same number of nodes and the same token distribution as at the time of the backup, and each node must be restored from the backup of the same node. Suitable for vnode-based keyspaces only.

For cluster-wide backup and restore, use ScyllaDB Manager, which orchestrates the process across the cluster.

Prerequisites¶

The following steps and notes apply to all restore methods.

From one of the nodes, recreate the schema.

cqlsh -e "SOURCE '/path_to_schema/<schema_name.cql>'"

For example:

cqlsh -e "SOURCE 'centos/db_schema.cql'"

Only a superuser should perform it.
If the tables you are restoring already exist and contain data, truncate each of them, so that the existing data does not overwrite the restored data. Truncating a base table also truncates its materialized views and secondary indexes, no extra action is needed for them.

cqlsh -e "TRUNCATE <keyspace_name>.<table_name>"

For example:

cqlsh -e "TRUNCATE mykeyspace.team_players"

Note

If you are restoring encrypted backup files, make sure ScyllaDB is configured with the same keys that were used to encrypt the data before starting the restore process.

Restore from object storage¶

Use this method to restore SSTables backed up to S3-compatible object storage with nodetool backup. The node downloads the SSTables from the bucket and streams their contents to the nodes owning the data (load and stream), so the restore works regardless of cluster topology changes since the backup, and the cluster stays online.

The object storage endpoint must be configured on the nodes, as described in Configuring Object Storage.

Note

If the table has any Materialized Views (MV) or Secondary Indexes (SI), view updates are generated automatically as the base table data is streamed. Restore the base table SSTables only; restoring MV or SI SSTables is not supported and will fail.

Procedure

  1. Complete the prerequisites.

  2. List the backed-up SSTables in the bucket under the prefix used during the backup. The restore command takes the paths of the TOC.txt components of the SSTables to restore, relative to the prefix – the remainder of each object key after the prefix. Note that listing tools print full object keys, from the bucket root, so the prefix needs to be stripped. For example:

    aws s3 ls --recursive s3://bucket-foo/ks/cf/24601/ | awk '/-TOC.txt$/ { print $4 }' | sed 's|^ks/cf/24601/||'
    
  3. Run nodetool restore, passing the endpoint, bucket, prefix, target keyspace and table, and the list of prefix-relative TOC paths:

    nodetool restore --endpoint s3.us-east-2.amazonaws.com --bucket bucket-foo --prefix ks/cf/24601 \
      --keyspace ks --table cf \
      me-3gdq_0bki_2dy4w2gqj6hoso4mw1-big-TOC.txt \
      me-3gdq_0bki_2dipc1ysb2x2a3btgh-big-TOC.txt
    

    Alternatively, put the same prefix-relative TOC paths (newline-separated) in a file and pass it with the --sstables-file-list option:

    cat > sstables.list <<EOF
    me-3gdq_0bki_2dy4w2gqj6hoso4mw1-big-TOC.txt
    me-3gdq_0bki_2dipc1ysb2x2a3btgh-big-TOC.txt
    EOF
    
    nodetool restore --endpoint s3.us-east-2.amazonaws.com --bucket bucket-foo --prefix ks/cf/24601 \
      --keyspace ks --table cf --sstables-file-list sstables.list
    
  4. Monitor the restore. By default, the command waits for the restore to finish and reports its final status. With the --nowait option, it returns a task ID immediately; use the nodetool tasks commands to track progress or cancel the operation.

Speeding up the restore

A single nodetool restore invocation runs on one node, which downloads and streams all the listed SSTables. To parallelize the work, split the list of SSTables between the nodes and run nodetool restore on each of them. The --scope option (node, rack, dc, or all) constrains where each node streams the data, so that concurrent restores don’t stream the same partition to a replica more than once. See nodetool restore for details on combining --scope with per-node SSTable lists.

With the --primary-replica-only option, each partition is streamed only to its primary replica. This reduces the amount of streamed data, but you must run a full cluster repair after the restore completes to replicate the data to the remaining replicas: for vnode-based keyspaces, run nodetool repair -pr on every node; for tablet-based keyspaces, run nodetool cluster repair on any single node.

Restore with load and stream¶

Use this method when the backed-up SSTable files are available on disk (for example, snapshot files copied back from external storage). The SSTables are read and their contents are streamed to the nodes owning the data, so the method works regardless of cluster topology changes since the backup. Each SSTable needs to be uploaded to only one node, any node, and the cluster stays online.

Note

If the table has any Materialized Views (MV) or Secondary Indexes (SI), view updates are generated automatically as the base table data is streamed. Upload the base table SSTables only; uploading MV or SI SSTables is not supported and will fail.

Procedure

  1. Complete the prerequisites.

  2. Copy the backed-up SSTable files of a table to that table’s upload directory on one of the nodes, and make sure the files are owned by the scylla user and group:

    sudo cp /path/to/backup/sstables/* /var/lib/scylla/data/mykeyspace/team_players-6e856600017f11e790f4000000000000/upload/
    sudo chown -R scylla:scylla /var/lib/scylla/data/mykeyspace/team_players-6e856600017f11e790f4000000000000/upload/
    

    You can distribute the backup files between several nodes to parallelize the restore; make sure each SSTable is uploaded to only one node.

  3. Run nodetool refresh with the --load-and-stream option on each node holding uploaded files:

    nodetool refresh mykeyspace team_players --load-and-stream
    

    See Load and Stream for the --scope and --primary-replica-only options that constrain the set of target replicas. If --primary-replica-only is used, run a full cluster repair after the restore completes to replicate the data to the remaining replicas: for vnode-based keyspaces, run nodetool repair -pr on every node; for tablet-based keyspaces, run nodetool cluster repair on any single node.

Restore to an identical cluster¶

This method places the snapshot files directly back into the table directories and restarts the nodes.

Note

The following procedure assumes data is restored to the same cluster that was backed-up:

  • same number of nodes

  • same token range per node

The procedure restores each node using the backup file of the same node. If this is not the case, use the load and stream method described above instead. It works regardless of topology changes, but is slower than restoring to an identical cluster.

This method is suitable for vnode-based keyspaces only. For tablet-based keyspaces, use the object storage or load and stream method instead.

Complete the prerequisites first.

Note

Best practise is not to restore Materialized Views (MV) and Secondary Indexes (SI) SSTables. It is recommended to:

  • Drop the MV and SI using DROP MATERIALIZED VIEW or DROP INDEX

  • Restore the base table only (see below)

  • Recreate the MV or SI, using the original description from the CQL backup, using CREATE MATERIALIZED VIEW or CREATE INDEX

Repeat the following steps for each node in the cluster:¶

  1. Run the nodetool drain command to ensure the data is flushed to the SSTables

  2. Shut down the node

    sudo systemctl stop scylla-server
    
    docker exec -it some-scylla supervisorctl stop scylla
    

    (without stopping some-scylla container)

  3. Delete all the files in the commitlog. Deleting the commitlog will prevent the newer insert from overriding the restored data.

    sudo rm -rf /var/lib/scylla/commitlog/*

  4. Delete all the files in the keyspace_name_table. Note that by default the snapshots are created under ScyllaDB data directory /var/lib/scylla/data/keyspace_name/table_name-UUID/.

    Make sure NOT to delete the existing snapshots in the process.

    For example:

    sudo ll /var/lib/scylla/data/mykeyspace/team_players-6e856600017f11e790f4000000000000
    
    -rw-r--r-- 1 scylla   scylla     66 Mar  5 09:19 nba-team_players-ka-1-CompressionInfo.db
    -rw-r--r-- 1 scylla   scylla    669 Mar  5 09:19 nba-team_players-ka-1-Data.db
    -rw-r--r-- 4 scylla   scylla     10 Mar  5 08:46 nba-team_players-ka-1-Digest.sha1
    -rw-r--r-- 1 scylla   scylla     24 Mar  5 09:19 nba-team_players-ka-1-Filter.db
    -rw-r--r-- 1 scylla   scylla    218 Mar  5 09:19 nba-team_players-ka-1-Index.db
    -rw-r--r-- 1 scylla   scylla     38 Mar  5 09:19 nba-team_players-ka-1-ScyllaDB.db
    -rw-r--r-- 1 scylla   scylla   4446 Mar  5 09:19 nba-team_players-ka-1-Statistics.db
    -rw-r--r-- 1 scylla   scylla     89 Mar  5 09:19 nba-team_players-ka-1-Summary.db
    -rw-r--r-- 4 scylla   scylla    101 Mar  5 08:46 nba-team_players-ka-1-TOC.txt
    drwx------ 5 scylla   scylla     69 Mar  6 08:14 snapshots
    drwx------ 2 scylla   scylla      6 Mar  5 08:40 upload
    
    sudo rm -f  /var/lib/scylla/data/mykeyspace/team_players-6e856600017f11e790f4000000000000/*
    
    rm: cannot remove ‘/var/lib/scylla/data/nba/team_roster-c019f8108fda11e8b16a000000000001/snapshots’: Is a directory
    rm: cannot remove ‘/var/lib/scylla/data/nba/team_roster-c019f8108fda11e8b16a000000000001/upload’: Is a directory
    
    sudo ll /var/lib/scylla/data/mykeyspace/team_players-6e856600017f11e790f4000000000000/
    
    drwx------ 5 scylla   scylla     69 Mar  6 08:14 snapshots
    drwx------ 2 scylla   scylla      6 Mar  5 08:40 upload
    
  5. Select the snapshot you want to restore (usually the most recent one)

    /var/lib/scylla/data/keyspace_name/table_name-UUID/snapshots/<snapshot_name>
    

    For example:

    cd /var/lib/scylla/data/mykeyspace/team_players-6e856600017f11e790f4000000000000/snapshots/1487847672222
    
  6. Copy the snapshots directory content to the /var/lib/scylla/data/keyspace_name/table_name-UUID/ directory

    For example:

    sudo cp -r * /var/lib/scylla/data/mykeyspace/team_players-6e856600017f11e790f4000000000000
    

    Warning

    Copying files into the table’s data directory is only allowed while the ScyllaDB service is stopped. To load SSTables into a running node, place them in the table’s upload directory and use nodetool refresh instead.

  7. If you have incremental backup files, copy them from the backups folder /var/lib/scylla/data/keyspace_name/table_name-UUID/backups to the /var/lib/scylla/data/keyspace_name/table_name-UUID/ directory

    For example:

    sudo cp -r /var/lib/scylla/data/mykeyspace/team_players-6e856600017f11e790f4000000000000/backups/* /var/lib/scylla/data/mykeyspace/team_players-6e856600017f11e790f4000000000000
    
  8. Make sure that all files are owned by the scylla user and group:

    sudo chown -R scylla:scylla /var/lib/scylla/data/mykeyspace/team_players-6e856600017f11e790f4000000000000
    
  9. Start the node

    sudo systemctl start scylla-server
    
    docker exec -it some-scylla supervisorctl start scylla
    

    (with some-scylla container already running)

After performing the above on all nodes, run a full cluster repair: for vnode-based keyspaces, run nodetool repair -pr on every node; for tablet-based keyspaces, run nodetool cluster repair on any single node; run both if you restored keyspaces of both kinds. This makes sure that the data is consistent on all nodes and between each node.

Was this page helpful?

PREVIOUS
Backup your Data
NEXT
Delete Backups
  • Create an issue
  • Edit this page

On this page

  • Restore from a Backup and Incremental Backup
    • Choosing a restore method
    • Prerequisites
    • Restore from object storage
    • Restore with load and stream
    • Restore to an identical cluster
      • Repeat the following steps for each node in the cluster:
ScyllaDB Manual
Search Ask AI
  • 2026.3
    • master
    • 2026.3
    • 2026.2
    • 2026.1
    • 2025.4
    • 2025.3
    • 2025.2
    • 2025.1
  • Getting Started
    • Install ScyllaDB 2026.3
      • Launch ScyllaDB 2026.3 on AWS
      • Launch ScyllaDB 2026.3 on GCP
      • Launch ScyllaDB 2026.3 on Azure
      • Launch ScyllaDB 2026.3 on Oracle Cloud Infrastructure (OCI)
      • ScyllaDB Web Installer for Linux
      • Install ScyllaDB 2026.3 Linux Packages
      • Run ScyllaDB in Docker
      • Install ScyllaDB Without root Privileges
      • Air-gapped Server Installation
      • ScyllaDB Housekeeping and how to disable it
      • ScyllaDB Developer Mode
    • Configure ScyllaDB
    • ScyllaDB Configuration Reference
    • ScyllaDB Requirements
      • System Requirements
      • OS Support
      • Cloud Instance Recommendations
      • ScyllaDB in a Shared Environment
    • Migrate to ScyllaDB
      • Migration Process from Cassandra to ScyllaDB
      • ScyllaDB and Apache Cassandra Compatibility
      • Migration Tools Overview
    • Integration Solutions
      • Integrate ScyllaDB with Spark
      • Integrate ScyllaDB with KairosDB
      • Integrate ScyllaDB with Presto
      • Integrate ScyllaDB with Elasticsearch
      • Integrate ScyllaDB with Kubernetes
      • Integrate ScyllaDB with the JanusGraph Graph Data System
      • Integrate ScyllaDB with DataDog
      • Integrate ScyllaDB with Kafka
      • Integrate ScyllaDB with IOTA Chronicle
      • Integrate ScyllaDB with Spring
      • Shard-Aware Kafka Connector for ScyllaDB
      • Install ScyllaDB with Ansible
      • Integrate ScyllaDB with Databricks
      • Integrate ScyllaDB with Jaeger Server
      • Integrate ScyllaDB with MindsDB
  • ScyllaDB for Administrators
    • Administration Guide
    • Procedures
      • Cluster Management
      • Backup & Restore
      • Change Configuration
      • Maintenance
      • Best Practices
      • Benchmarking ScyllaDB
      • Migrate from Cassandra to ScyllaDB
      • Disable Housekeeping
    • Security
      • ScyllaDB Security Checklist
      • Enable Authentication
      • Enable and Disable Authentication Without Downtime
      • Creating a Superuser
      • Generate a cqlshrc File
      • Reset Authenticator Password
      • Enable Authorization
      • Grant Authorization CQL Reference
      • Certificate-based Authentication
      • Role Based Access Control (RBAC)
      • ScyllaDB Auditing Guide
      • Encryption: Data in Transit Client to Node
      • Encryption: Data in Transit Node to Node
      • Generating a self-signed Certificate Chain Using openssl
      • Configure SaslauthdAuthenticator
      • Encryption at Rest
      • LDAP Authentication
      • LDAP Authorization (Role Management)
      • Software Bill Of Materials (SBOM)
    • Admin Tools
      • Nodetool Reference
      • CQLSh
      • Admin REST API
      • Tracing
      • ScyllaDB SStable
      • ScyllaDB SStable Script API
      • ScyllaDB Types
      • SSTableLoader
      • cassandra-stress
      • ScyllaDB Logs
      • Seastar Perftune
      • Virtual Tables
      • Reading mutation fragments
      • Maintenance socket
      • Maintenance mode
      • Task manager
    • ScyllaDB Monitoring Stack
    • ScyllaDB Operator
    • ScyllaDB Manager
    • Upgrade Procedures
      • Upgrade Guides
    • System Configuration
      • System Configuration Guide
      • scylla.yaml
      • ScyllaDB Snitches
      • Configuration Parameters
    • Benchmarking ScyllaDB
    • ScyllaDB Diagnostic Tools
  • ScyllaDB for Developers
    • Develop with ScyllaDB
    • Tutorials and Example Projects
    • Learn to Use ScyllaDB
    • ScyllaDB Alternator
    • ScyllaDB Drivers
  • CQL Reference
    • CQLSh: the CQL shell
    • Reserved CQL Keywords and Types (Appendices)
    • Compaction
    • Consistency Levels
    • Consistency Level Calculator
    • Data Definition
    • Data Manipulation
      • SELECT
      • INSERT
      • UPDATE
      • DELETE
      • BATCH
    • Data Types
    • Definitions
    • Global Secondary Indexes
    • Expiring Data with Time to Live (TTL)
    • Functions
    • CQL Guardrails
    • Wasm support for user-defined functions
    • JSON Support
    • Materialized Views
    • DESCRIBE SCHEMA
    • Service Levels
    • ScyllaDB CQL Extensions
  • Alternator: DynamoDB API in ScyllaDB
    • Getting Started With ScyllaDB Alternator
    • ScyllaDB Alternator for DynamoDB users
    • Alternator-specific APIs
    • Reducing network costs in Alternator
    • Alternator Vector Search
  • Features
    • Lightweight Transactions
    • Global Secondary Indexes
    • Local Secondary Indexes
    • Materialized Views
    • Counters
    • Change Data Capture
      • CDC Overview
      • The CDC Log Table
      • Basic operations in CDC
      • CDC Streams
      • CDC Stream Changes
      • Querying CDC Streams
      • Advanced column types
      • Preimages and postimages
      • Data Consistency in CDC
    • Workload Attributes
    • Workload Prioritization
    • Backup and Restore
    • Incremental Repair
    • Automatic Repair
    • Vector Search
    • Full-Text Search
  • ScyllaDB Architecture
    • Data Distribution with Tablets
    • ScyllaDB Ring Architecture
    • ScyllaDB Fault Tolerance
    • Consistency Level Console Demo
    • ScyllaDB Anti-Entropy
      • ScyllaDB Hinted Handoff
      • ScyllaDB Read Repair
      • ScyllaDB Repair
    • SSTable
      • ScyllaDB SSTable - 2.x
      • ScyllaDB SSTable - 3.x
    • Compaction Strategies
    • Raft Consensus Algorithm in ScyllaDB
    • Zero-token Nodes
  • Troubleshooting ScyllaDB
    • Errors and Support
      • Report a ScyllaDB problem
      • Error Messages
      • Change Log Level
    • ScyllaDB Startup
      • Ownership Problems
      • ScyllaDB will not Start
    • Cluster and Node
      • Handling Node Failures
      • Failure to Add, Remove, or Replace a Node
      • Failed Decommission Problem
      • Cluster Timeouts
      • Node Joined With No Data
      • NullPointerException
      • Failed Schema Sync
    • Data Modeling
      • ScyllaDB Large Partitions Table
      • ScyllaDB Large Rows and Cells Table
      • Large Partitions Hunting
      • Failure to Update the Schema
    • Data Storage and SSTables
      • Space Utilization Increasing
      • Disk Space is not Reclaimed
      • SSTable Corruption Problem
      • Pointless Compactions
      • Limiting Compaction
    • CQL
      • Time Range Query Fails
      • COPY FROM Fails
      • CQL Connection Table
    • ScyllaDB Monitor and Manager
      • Manager and Monitoring integration
      • Manager lists healthy nodes as down
    • Installation and Removal
      • Removing ScyllaDB on Ubuntu breaks system packages
  • Knowledge Base
    • Upgrading from experimental CDC
    • Compaction
    • Consistency in ScyllaDB
    • Counting all rows in a table is slow
    • CQL Query Does Not Display Entire Result Set
    • When CQLSh query returns partial results with followed by “More”
    • Run ScyllaDB and supporting services as a custom user:group
    • Customizing CPUSET
    • Decoding Stack Traces
    • Snapshots and Disk Utilization
    • DPDK mode
    • Debug your database with Flame Graphs
    • Efficient Tombstone Garbage Collection in ICS
    • How to Change gc_grace_seconds for a Table
    • Gossip in ScyllaDB
    • How does ScyllaDB LWT Differ from Apache Cassandra ?
    • Map CPUs to ScyllaDB Shards
    • ScyllaDB Memory Usage
    • NTP Configuration for ScyllaDB
    • POSIX networking for ScyllaDB
    • ScyllaDB consistency quiz for administrators
    • Recreate RAID devices
    • How to Safely Increase the Replication Factor
    • ScyllaDB and Spark integration
    • Increase ScyllaDB resource limits over systemd
    • ScyllaDB Seed Nodes
    • How to Set up a Swap Space
    • ScyllaDB Snapshots
    • ScyllaDB payload sent duplicated static columns
    • Stopping a local repair
    • System Limits
    • How to flush old tombstones from a table
    • Time to Live (TTL) and Compaction
    • ScyllaDB Nodes are Unresponsive
    • Update a Primary Key
    • Using the perf utility with ScyllaDB
    • Configure ScyllaDB Networking with Multiple NIC/IP Combinations
  • Reference
    • AWS Images
    • Azure Images
    • GCP Images
    • Configuration Parameters
    • Glossary
    • Limits
    • API Reference
      • Authorization Cache
      • Cache Service
      • Collectd
      • Column Family
      • Commit Log
      • Compaction Manager
      • Endpoint Snitch Info
      • Error Injection
      • Failure Detector
      • Gossiper
      • Hinted Handoff
      • LSA
      • Messaging Service
      • Raft
      • Storage Proxy
      • Storage Service
      • Stream Manager
      • System
      • Task Manager Test
      • Task Manager
      • Tasks
    • Metrics
  • ScyllaDB FAQ
  • 2024.2 and earlier documentation
Docs Tutorials University Contact Us About Us
© 2026, ScyllaDB. All rights reserved. | Terms of Service | Privacy Policy | ScyllaDB, and ScyllaDB Cloud, are registered trademarks of ScyllaDB, Inc.
Last updated on 20 Sep 2026.
Powered by Sphinx 9.1.0 & ScyllaDB Theme 1.9.3