Core: Ignore split offsets when the last split offset is past the file length #8860

amogh-jahagirdar · 2023-10-17T16:18:19Z

Follow up to #8834

This change ignores split offsets when the last split offset is past the file length. This is done as a defensive check to make sure that readers don't attempt to use corrupted split offset stemming from the issue described in #8834. Note: It is possible that split offsets were corrupted and then the last split offset was not past the file length. In that case, it's still acceptable to use the split offsets because it won't break reading logic since the split offsets would be guaranteed to be within the file boundaries.

amogh-jahagirdar · 2023-10-17T16:18:53Z

cc @bryanck

core/src/main/java/org/apache/iceberg/BaseFile.java

core/src/test/java/org/apache/iceberg/TableTestBase.java

singhpk234 · 2023-10-17T17:10:50Z

core/src/main/java/org/apache/iceberg/BaseFile.java

+      return null;
+    }
+
+    // If the last split offset is past the file size this means the split offsets are corrupted and


[doubt] wondering if throwing an exception or having a pre-condition would be helpful to identify the buggy writer ?

Do you mean throwing at read time (i.e. instead of returning null, we throw)? If so, I don't think we want to do that because that's essentially unnecessarily breaking future readers and split offsets are optional anyways; this approach takes the stance of detecting the corruption and making sure read logic doesn't leverage the corrupted metadata.

If you mean throwing at the time of writing the manifest entry (a precondition check in the constructor of BaseFile), I went back and forth on this but I think the problem there is let's say someone upgrades. When some process is performed which rewrites a set of files (including some corrupted entries) it would fail due to the precondition. The benefit is it would prevent spreading the previous corruption which is nice, but at the cost of failing operations. Considering again the corrupted split offsets will be ignored at read time anyways, failing at write time seems needless.

To prevent spreading previous corrupted state, at the time of writing the manifest if the corruption is detected the split offsets could be recomputed (a sort of "fix-up" process). This requires more investigation though, not sure how feasible it is and the perf implications (e.g. for Parquet we'd need to go through the block metadata again)

let me know what you think!

Agree with you !

I was mostly coming from the point of view of a buggy writer (that didn't use core-lib as we expose split offsets via ParquetMetadata or purpose fully passed wrong offsets) which has already committed this metadata. Such writers will never be caught because we will be silently skipping the malformed offsets, was wondering if having a warning log then, during reads so that we could let the readers know of the corruptions and reads would be a bit un-optimized, so that it could help in backtracking the buggy writer, thoughts ?

looks like Split Offsets being ordered in ascending order is not being enforced during reads either, we just swallow, we should be fine in this case as well then :P

if (splitOffsets != null && ArrayUtil.isStrictlyAscending(splitOffsets)) { return () -> new OffsetsAwareSplitScanTaskIterator<>( self(), length(), splitOffsets, this::newSplitTask); } else { return () -> new FixedSizeSplitScanTaskIterator<>( self(), length(), targetSplitSize, this::newSplitTask); } }

…e length

rdblue · 2023-10-17T19:15:02Z

Merged. Thanks for getting this read, @amogh-jahagirdar!

For context, this catches bad metadata written by 1.4.0 and ignores it. This is needed if tables have bad split offsets in order to be read by Spark, Hive, and Flink.

…e length (apache#8860)

…e length (#8860) (#8861)

amogh-jahagirdar requested review from rdblue and danielcweeks October 17, 2023 16:18

github-actions bot added the core label Oct 17, 2023

bryanck reviewed Oct 17, 2023

View reviewed changes

core/src/main/java/org/apache/iceberg/BaseFile.java Show resolved Hide resolved

amogh-jahagirdar requested review from nastra and RussellSpitzer October 17, 2023 16:20

nastra added this to the Iceberg 1.4.1 milestone Oct 17, 2023

amogh-jahagirdar force-pushed the defensive-splitoffset-read branch from cc03e1e to f407974 Compare October 17, 2023 16:21

bryanck reviewed Oct 17, 2023

View reviewed changes

core/src/main/java/org/apache/iceberg/BaseFile.java Outdated Show resolved Hide resolved

amogh-jahagirdar force-pushed the defensive-splitoffset-read branch 3 times, most recently from a0adedb to 32e94b4 Compare October 17, 2023 16:51

amogh-jahagirdar commented Oct 17, 2023

View reviewed changes

core/src/test/java/org/apache/iceberg/TableTestBase.java Outdated Show resolved Hide resolved

singhpk234 reviewed Oct 17, 2023

View reviewed changes

Core: Ignore split offsets when the last split offset is past the fil…

8f83bdc

…e length

amogh-jahagirdar force-pushed the defensive-splitoffset-read branch from 32e94b4 to 8f83bdc Compare October 17, 2023 17:50

amogh-jahagirdar requested review from singhpk234 and bryanck October 17, 2023 18:15

rdblue approved these changes Oct 17, 2023

View reviewed changes

rdblue merged commit ad602a3 into apache:main Oct 17, 2023
45 checks passed

amogh-jahagirdar added a commit to amogh-jahagirdar/iceberg that referenced this pull request Oct 17, 2023

Core: Ignore split offsets when the last split offset is past the fil…

29351d3

…e length (apache#8860)

amogh-jahagirdar mentioned this pull request Oct 17, 2023

[1.4.x] Core: Ignore split offsets when the last split offset is past the file length #8861

Merged

nastra pushed a commit that referenced this pull request Oct 18, 2023

Core: Ignore split offsets when the last split offset is past the fil…

1e1c67e

…e length (#8860) (#8861)

amogh-jahagirdar mentioned this pull request Oct 26, 2023

Core: Ignore split offsets array when split offset is past file length #8925

Merged

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Core: Ignore split offsets when the last split offset is past the file length #8860

Core: Ignore split offsets when the last split offset is past the file length #8860

amogh-jahagirdar commented Oct 17, 2023 •

edited

Loading

amogh-jahagirdar commented Oct 17, 2023

singhpk234 Oct 17, 2023

amogh-jahagirdar Oct 17, 2023 •

edited

Loading

singhpk234 Oct 17, 2023 •

edited

Loading

singhpk234 Oct 17, 2023

rdblue commented Oct 17, 2023

Core: Ignore split offsets when the last split offset is past the file length #8860

Core: Ignore split offsets when the last split offset is past the file length #8860

Conversation

amogh-jahagirdar commented Oct 17, 2023 • edited Loading

amogh-jahagirdar commented Oct 17, 2023

singhpk234 Oct 17, 2023

Choose a reason for hiding this comment

amogh-jahagirdar Oct 17, 2023 • edited Loading

Choose a reason for hiding this comment

singhpk234 Oct 17, 2023 • edited Loading

Choose a reason for hiding this comment

singhpk234 Oct 17, 2023

Choose a reason for hiding this comment

rdblue commented Oct 17, 2023

amogh-jahagirdar commented Oct 17, 2023 •

edited

Loading

amogh-jahagirdar Oct 17, 2023 •

edited

Loading

singhpk234 Oct 17, 2023 •

edited

Loading