zhang-arvin commented on issue #17301:
URL: https://github.com/apache/iceberg/issues/17301#issuecomment-5835495313

   @zhang-arvin here — I'd like to take this one forward.
   
   Note the earlier attempt is closed only by the stale bot, not for technical 
reasons: PR #17302 was **approved by @uros-b on 2026-07-23** ("Nice, thank you 
@mattfaltyn!") and then auto-closed on 2026-08-30 for lack of activity. The 
root cause in the issue is still valid on `main`.
   
   Plan:
   1. Revive the approved approach: in 
`ParquetDictionaryRowGroupFilter#notNaN`, check for absent column metadata 
before consulting nullability, so row groups stay readable when a float/double 
column was added after an older file was written.
   2. Keep the minimal repro that `mattfaltyn` gave: `notNaN("not_in_file")` 
added to the `exprs` array in 
`TestDictionaryRowGroupFilter#testColumnNotInFile`.
   3. Rebase on current `main` and re-request review, crediting the original 
author's approved patch.
   
   Please assign this to me if convenient.


-- 
This is an automated message from the Apache Git Service.
To respond to the message, please log on to GitHub and use the
URL above to go to the specific comment.

To unsubscribe, e-mail: [email protected]

For queries about this service, please contact Infrastructure at:
[email protected]


---------------------------------------------------------------------
To unsubscribe, e-mail: [email protected]
For additional commands, e-mail: [email protected]

Reply via email to