Video | Tableau | Data prep | Data visualisation

How to create and use Tableau Extracts

If you want the best performance in Tableau, work with an extract by default — unless you're on a published data source or a database built for those speeds.

Watch on YouTube
  • Extracts optimise large files, let you take a snapshot of data, can speed up some computations, and let you share a portable subset of your data with others.
  • Since Tableau 2020.3, opening a .tde extract automatically upgrades it to the .hyper format, and you cannot downgrade a hyper workbook back to TDE.
  • When you connect to an extract, Tableau shows it as a single cylinder (live) connection to the hyper file, which can be confusing — you can toggle 'Use Extract' to switch between the snapshot and the live source.
  • The physical-tables extract option only appears when your logical layer contains joined physical tables; otherwise extracts work at the logical-table level.
  • Filters, aggregation to visible dimensions, row sampling (top N or random), and incremental refresh let you shrink and control exactly what data the extract captures.
  • Incremental refresh on a logical model requires you to pick a table and a unique field (typically a date) so Tableau can identify new rows, and the extract history records full versus incremental refreshes.

Extracts snapshot, shrink and speed up your data, and shaping them properly gives you control and performance you don't get from a live connection.

Tim connects to a standard data source and walks through creating, opening and shaping an extract, including the shift from the old .tde format to today's .hyper format.

The Breakdown
  1. What an extract is and why it helps 0:34

    An extract is a saved snapshot of your data (.hyper, formerly .tde) rather than a live connection. It optimises large or slow sources, lets you keep working while a database is unstable or distant, can speed up some computations, and lets you share a portable subset of data.

  2. Default to extracts for performance 1:17

    Use an extract by default for best performance, unless you're on a published data source or a database already built for speed.

  3. Creating your first extract 3:18

    Right-click your data source and choose to extract the data; accepting the defaults and saving creates the .hyper file. The source then shows a double-cylinder icon and previously greyed-out options (history, properties, use-extract toggle) become active.

  4. Opening hyper files and the single-cylinder confusion 5:34

    Opening a .hyper file directly gives a ready-made data source, but dragging it into a workbook shows it as a single-cylinder live connection to the extract — easy to mistake for a genuine live database link. Check Properties to confirm what you're actually pointed at.

  5. TDE to hyper: a one-way upgrade 8:03

    Since Tableau 2020.3, opening a .tde auto-upgrades it to .hyper with no way back. Functionality is mostly the same, but hyper is faster and preferable on recent versions — take care sharing upgraded files with people still on .tde-only Tableau.

  6. Toggling between live and extract 10:09

    Untick 'Use Extract' to switch back to the live source, then re-enable it later — handy for confirming a live connection still works (e.g. dev vs prod) without losing your extract.

  7. Logical vs physical table extracts 11:50

    A 'physical tables' extract option only appears when your logical layer contains joined physical tables; otherwise extraction is at logical-table level. This matters for cases like row-level security, where a join should run live rather than be pre-flattened.

  8. Shaping the extract: filters, aggregation, sampling, incremental refresh 15:02

    Filters, aggregation to visible dimensions, and top-N or random sampling are independent options you can mix to shrink an extract. Incremental refresh is different: it needs a table plus a unique, unchanging field (typically a date) to identify new rows, and the extract history then logs full versus incremental refreshes.

Worth Knowing
  • Extract data can behave oddly around dates and formats depending on when the source data was captured.
  • You can't downgrade a hyper workbook back to .tde, which can break sharing with collaborators on older Tableau versions.
  • Filtering an extract permanently excludes that data until you recreate or refresh with different settings.
  • Incremental refresh needs a field that can't change after creation (like order date) — a mutable field risks missed or duplicate rows.
Use It When

Reach for this when you need faster performance on a large or slow source, a stable snapshot while a database is in flux, or a filtered subset of data to hand off to someone else.

How this Rollup was made provenance & method

A Rollup is drafted by AI from the video's transcript, then reviewed and edited by Tim. Everything used to produce this one is listed below — the model, the exact prompt, and the source video — so the process is transparent and reproducible.

Transcription
On-device — NVIDIA Parakeet v3 for recent videos, OpenAI Whisper large-v3 for earlier ones. The transcript never leaves the machine or gets published.
Drafting
Claude Sonnet 5 in the cloud, from that transcript.
Prompt
The exact Rollup prompt (v2) — the full system prompt, unedited.
Source video
Watch on YouTube
Drafted
4 July 2026 at 21:56
Reviewed & edited
5 July 2026 at 09:41 · by Tim Ngwena

Model + prompt + video is everything you'd need to recreate a Rollup like this yourself. The one thing we don't share is the transcript.

Rights. The video and its transcript are the property of TN Media Ltd. Unauthorised use or download is prohibited. © TN Media Ltd.