rambleraptor opened a new issue, #4080:
URL: https://github.com/apache/iceberg-python/issues/4080

   ### Feature Request / Improvement
   
   We want to be able to read the row-lineage values on a table. This is a 
blocker for writing v3 tables.
   
   ```python
   table.scan(selected_fields=("id", "_row_id", 
"_last_updated_sequence_number"))
   ```
   
   - [ ] Data files with a null `first_row_id` inherit it from the manifest's 
`first_row_id` plus a running record count: #3972
   - [ ] Define the reserved columns `_row_id` (2147483540) and 
`_last_updated_sequence_number` (2147483539)
   - [ ] Let scans project them
   - [ ] Use the values stored in the file if there are any. Fill nulls with 
`first_row_id + pos` and the file's data sequence number.
   - [ ] Compute `pos` before applying deletes
   - [ ] Show `first_row_id` in `inspect.files()`
   - [ ] REST scan planning drops `first_row_id`
   - [ ] Integration test: each file's `first_row_id` matches Spark's 
`MIN(_row_id)`: #3972
   - [ ] Integration test: Spark `MERGE` / `UPDATE` / `rewrite_data_files`, 
then compare with Spark's `SELECT _row_id, _last_updated_sequence_number`


-- 
This is an automated message from the Apache Git Service.
To respond to the message, please log on to GitHub and use the
URL above to go to the specific comment.

To unsubscribe, e-mail: [email protected]

For queries about this service, please contact Infrastructure at:
[email protected]


---------------------------------------------------------------------
To unsubscribe, e-mail: [email protected]
For additional commands, e-mail: [email protected]

Reply via email to