Skip to content

fix: use full table schema when producing the prune map - #380

Open
gruuya wants to merge 1 commit into
JanKaul:mainfrom
splitgraph:pruning-fixes
Open

fix: use full table schema when producing the prune map#380
gruuya wants to merge 1 commit into
JanKaul:mainfrom
splitgraph:pruning-fixes

Conversation

@gruuya

@gruuya gruuya commented Jul 31, 2026

Copy link
Copy Markdown
Contributor

Instead of the partition schema, which contains only a subset of columns for partitioned tables in general, or no columns at all for unpartitioned tables, compute the prune map with the full table arrow schema instead.

Note that this didn't impact correctness, since row filtering was still applied correctly and moreover the files were pruned dynamically after inspecting their stats during execution. it's just that unpartitioned tables got zero plan-time file pruning and partitioned tables got none on non-partition columns.

Instead, now we just skip processing those files at plan time altogether, which can be especially useful for tables with many files (previously we payed a footer fetch during execution for each one just to discard it based on the stats).

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant