Commit 41d698ec authored 8 years ago by Eric Liang Committed by Wenchen Fan 8 years ago

[SPARK-18661][SQL] Creating a partitioned datasource table should not scan all files for table


## What changes were proposed in this pull request?

Even though in 2.1 creating a partitioned datasource table will not populate the partition data by default (until the user issues MSCK REPAIR TABLE), it seems we still scan the filesystem for no good reason.

We should avoid doing this when the user specifies a schema.

## How was this patch tested?

Perf stat tests.

Author: Eric Liang <ekl@databricks.com>

Closes #16090 from ericl/spark-18661.

(cherry picked from commit d9eb4c72)
Signed-off-by: Wenchen Fan <wenchen@databricks.com>

parent 8145c82b

No related branches found

No related tags found

Hide whitespace changes

Inline Side-by-side

Showing with 66 additions and 8 deletions

Please register or to comment