zhang-arvin commented on issue #17301:
URL: https://github.com/apache/iceberg/issues/17301#issuecomment-5835495313
@zhang-arvin here — I'd like to take this one forward.
Note the earlier attempt is closed only by the stale bot, not for technical
reasons: PR #17302 was **approved by @uros-b on 2026-07-23** ("Nice, thank you
@mattfaltyn!") and then auto-closed on 2026-08-30 for lack of activity. The
root cause in the issue is still valid on `main`.
Plan:
1. Revive the approved approach: in
`ParquetDictionaryRowGroupFilter#notNaN`, check for absent column metadata
before consulting nullability, so row groups stay readable when a float/double
column was added after an older file was written.
2. Keep the minimal repro that `mattfaltyn` gave: `notNaN("not_in_file")`
added to the `exprs` array in
`TestDictionaryRowGroupFilter#testColumnNotInFile`.
3. Rebase on current `main` and re-request review, crediting the original
author's approved patch.
Please assign this to me if convenient.
--
This is an automated message from the Apache Git Service.
To respond to the message, please log on to GitHub and use the
URL above to go to the specific comment.
To unsubscribe, e-mail: [email protected]
For queries about this service, please contact Infrastructure at:
[email protected]
---------------------------------------------------------------------
To unsubscribe, e-mail: [email protected]
For additional commands, e-mail: [email protected]