Messages by Thread
-
-
[PR] refactor(proto): migrate ParquetSource serde [datafusion]
via GitHub
-
Re: [I] [EPIC] first class support for struct field / Variant access in Parquet [datafusion]
via GitHub
-
Re: [PR] Iceberg-Ballista Crate [datafusion-ballista]
via GitHub
-
Re: [PR] chore: move CometCollationSuite to spark-4.1+ test shim [datafusion-comet]
via GitHub
-
[I] Native Parquet writes report row counts but not row contents to WriteTaskStatsTrackers [datafusion-comet]
via GitHub
-
[I] Reconcile the two config namespaces for native writes [datafusion-comet]
via GitHub
-
[I] Native Parquet writer derives schema nullability and field IDs from Arrow rather than Catalyst [datafusion-comet]
via GitHub
-
[I] Native Parquet write of a zero-partition RDD produces no output file [datafusion-comet]
via GitHub
-
[I] Native Parquet writer ignores Spark's Parquet writer properties (block size, page size, dictionary, writer version, statistics) [datafusion-comet]
via GitHub
-
[PR] test: pin down CometCast fallback for non-default collated strings [datafusion-comet]
via GitHub
-
[I] Add a C++ UDF example and a published C header for the Comet UDF ABI [datafusion-comet]
via GitHub
-
Re: [I] [EPIC] Port ExecutionPlan serialization to try_to_proto / try_from_proto hooks [datafusion]
via GitHub
-
[PR] perf: build spark_size LargeList lengths from i64 offsets [datafusion-comet]
via GitHub
-
[PR] chore(proto): deprecate accessors that only existed for proto serialization [datafusion]
via GitHub
-
[I] perf: explore bulk copies for no-null fixed-width Arrow writes [datafusion-comet]
via GitHub
-
[PR] [PyLimit] adding skip and fetch with datafusion expr support [datafusion-python]
via GitHub
-
[PR] refactor(proto): destructure plan and proto structs in core plan serde hooks [datafusion]
via GitHub
-
[PR] refactor(proto): destructure plan and proto structs in aggregate and window serde hooks [datafusion]
via GitHub
-
[PR] refactor(proto): destructure plan and proto structs in join serde hooks [datafusion]
via GitHub
-
[PR] fix(proto): preserve HashJoinExec fetch across serialization [datafusion]
via GitHub
-
[I] PyLimit (LogicalPlan Limit node) has no way to read skip/fetch — regressed in #905 [datafusion-python]
via GitHub
-
Re: [I] Short circuit expensive comparissions in GroupColumn trait [datafusion]
via GitHub
-
[PR] feat: detect Iceberg V2 writes and emit fall-back reasons [datafusion-comet]
via GitHub
-
Re: [I] Remove the per-UDF Mutex serializing Rust UDF batches [datafusion-comet]
via GitHub
-
[I] Rust UDF library cache holds its write lock across dlopen [datafusion-comet]
via GitHub
-
[I] Rust UDF adapter rebuilds the kernel impl and re-resolves the return type on every batch [datafusion-comet]
via GitHub
-
[I] A Rust UDF silently answers calls to an ordinary Scala UDF registered under the same name [datafusion-comet]
via GitHub
-
[I] Scope the Rust UDF registry to the session that registered the UDF [datafusion-comet]
via GitHub
-
[PR] refactor: hook native Parquet writes into Spark's WriteFilesExec seam [datafusion-comet]
via GitHub
-
Re: [I] Proto: migrate MemorySourceConfig [datafusion]
via GitHub
-
Re: [I] Proto: migrate ParquetSource [datafusion]
via GitHub
-
Re: [I] Proto: migrate the simple file scans (CsvSource, JsonSource, ArrowSource, AvroSource) [datafusion]
via GitHub
-
Re: [I] Proto: migrate the simple file scans (CsvSource, JsonSource, ArrowSource, AvroSource) [datafusion]
via GitHub
-
Re: [I] Proto: migrate the simple file scans (CsvSource, JsonSource, ArrowSource, AvroSource) [datafusion]
via GitHub
-
Re: [I] Proto: migrate the simple file scans (CsvSource, JsonSource, ArrowSource, AvroSource) [datafusion]
via GitHub
-
Re: [I] Proto: migrate the simple file scans (CsvSource, JsonSource, ArrowSource, AvroSource) [datafusion]
via GitHub
-
[PR] Add FFI query planner support with composable extension codecs [datafusion-python]
via GitHub
-
Re: [PR] refactor: use arrow make_comparator for nested structural equality in arrays_overlap and array_position [2/2] [datafusion-comet]
via GitHub
-
[PR] chore(deps): bump the all-other-cargo-deps group across 1 directory with 10 updates [datafusion]
via GitHub
-
[PR] fix: preserve CalendarInterval microseconds across Comet boundaries and native kernels [datafusion-comet]
via GitHub
-
Re: [PR] chore(deps): bump the all-other-cargo-deps group across 1 directory with 9 updates [datafusion]
via GitHub
-
[I] Determine whether "Apache Comet" is a suitable project name [datafusion-comet]
via GitHub
-
Re: [I] Comet native Parquet scan reads full nested columns despite pruned Spark ReadSchema [datafusion-comet]
via GitHub
-
Re: [PR] refactor(physical-plan): Simplify `ExecutionPlan` API with `replace_children` [datafusion]
via GitHub
-
Re: [I] Proto: add DataSource/FileSource try_to_proto hook + port FileScanConfig serde [datafusion]
via GitHub
-
[PR] Add typed Protocols for FFI capsule exports (part of #1577) [datafusion-python]
via GitHub
-
Re: [I] Statistics::calculate_total_byte_size discards a known total_byte_size when num_rows is unknown [datafusion]
via GitHub
-
[I] WindowEvaluator already provides the pure-Python window UDF base class #1577 (item 6) asks for [datafusion-python]
via GitHub
-
Re: [PR] fix: preserve total_byte_size in calculate_total_byte_size when num_r… [datafusion]
via GitHub