How to find and export your Amazon Redshift data.

Amazon Redshift reports table size in the SVV_TABLE_INFO system view, counted in 1 MB blocks. Superusers can query it, and a superuser can grant SELECT on it to other users. Sum the size column and report the total in GB.

At a glance

Report in
GB
Category
Databases and warehouses
Checked
September 2026
Check your data

Where Amazon Redshift shows how much you have.

  1. Connect as a superuser in Query Editor v2 or any SQL client. To let another user run the query, a superuser runs GRANT SELECT ON svv_table_info TO username;
  2. In each database, run: SELECT "schema", SUM(size) AS size_mb, SUM(estimated_visible_rows) AS visible_rows, MIN(create_time) AS first_table_created FROM svv_table_info GROUP BY "schema" ORDER BY size_mb DESC;
  3. Divide size_mb by 1024 for GB. The view only covers the database you are connected to and skips empty tables.
  4. For the cluster as a whole, check the PercentageDiskSpaceUsed metric in CloudWatch.

Plan and role. No special plan is needed. SVV_TABLE_INFO is visible only to superusers unless a superuser grants SELECT on it.

What to report on the Polyshares intake.

Report Amazon Redshift in GB. The intake also asks how many years the company has used it and how many seats it has.

On the Polyshares intake, report the summed size in GB. It counts 1 MB data blocks as stored, after column compression, and it includes rows marked for deletion but not yet vacuumed, so it will not match an uncompressed export. Use estimated_visible_rows rather than tbl_rows for live row counts. For years in use, take the earliest date in your largest fact tables; for seats, count the customers or users whose activity the tables record.

Export the full history.

  1. Give the cluster an IAM role that can write to your S3 bucket, and associate it with the cluster or set it as the default role.
  2. Unload each table to Parquet: UNLOAD ('SELECT * FROM public.orders') TO 's3://my-bucket/export/orders/' IAM_ROLE default FORMAT AS PARQUET MANIFEST VERBOSE; Add REGION 'us-west-2' if the bucket is in a different Region from the cluster.
  3. Check the output. UNLOAD writes files in parallel, one or more per slice, each up to 6.2 GB by default, and the verbose manifest lists every file with its row count.

UNLOAD writes pipe-delimited text by default and also supports CSV, JSON and Parquet. Parquet files compress each row group with Snappy, and TIMESTAMPTZ columns lose their time zone. The outer SELECT cannot use LIMIT, and GEOMETRY, HLLSKETCH and VARBYTE columns can only be unloaded to text or CSV. Output files are encrypted with SSE-S3 by default.

The intake only needs the size, so you do not have to run an export to fill it in. You also do not need to clean or scrub an export. Polyshares handles anonymization before anything moves.

Why Amazon Redshift data has value for AI training.

Redshift warehouses usually hold a business's cleaned and joined history: orders, billing, product usage and marketing touchpoints modeled into fact and dimension tables. Years of that history at row level show how customer behavior connects to revenue and retention, which is useful ground truth for training and evaluating models.

What stays private.

UNLOAD accepts any SELECT statement, so names, emails or card fields can be left out and only the agreed columns exported.

We anonymize identifying details before anything moves. Polyshares does that work, so your team does not have to scrub records first.

Your company keeps ownership of its data. The license is bounded to defined material and a defined use.

Polyshares is the buyer and pays your company directly. There is no fee or commission, and Polyshares reviews the record before it makes any offer.

Common questions about Amazon Redshift data.

What unit is the size column in SVV_TABLE_INFO?

It is a count of 1 MB data blocks. Divide the total by 1024 to get GB.

Why does SVV_TABLE_INFO return no rows for me?

The view is visible only to superusers by default, and it leaves out empty tables. A superuser can grant SELECT on svv_table_info to your user.

How large are the files UNLOAD creates?

Each file is at most 6.2 GB by default, and UNLOAD writes in parallel across slices. Set MAXFILESIZE to any value from 5 MB to 6.2 GB to change that.

Source: Amazon Redshift: SVV_TABLE_INFO, Amazon Redshift: UNLOAD, Amazon Redshift: Performance data (CloudWatch metrics). Checked September 2026.

Put your numbers on the intake.

List your systems, their sizes and your headcount in one intake. Polyshares reviews the record and prices it. There is no fee or commission to your company.

Check your data