Messages by Thread
-
Re: [PR] fix: Avoid pushing down sort under limits [datafusion]
via GitHub
-
[I] Trying to reduce CI time [datafusion-comet]
via GitHub
-
[I] Adaptive partial aggregation can shuffle 10x more rows than 1.0.0, with no supported way to turn it off [datafusion-comet]
via GitHub
-
Re: [PR] feat: route date and timestamp interval arithmetic through codegen dispatch [datafusion-comet]
via GitHub
-
[I] Native explode returns wrong booleans for arrays of structs after the first batch of output [datafusion-comet]
via GitHub
-
Re: [PR] perf: Optimize `array_agg(... IGNORE NULLS)` [datafusion]
via GitHub
-
Re: [PR] bench: add `array_compact` benchmarks [datafusion]
via GitHub
-
Re: [PR] fix: rethrow the exception a JVM input throws instead of a CometNativeException [datafusion-comet]
via GitHub
-
Re: [PR] fix: compact already-sliced batches before the native Iceberg writer counts NaNs [datafusion-comet]
via GitHub
-
[PR] fix: Avoid special-casing `count(1)` in `Count::value_from_stats` [datafusion]
via GitHub
-
[PR] fix: bound spill file batches by bytes so external sorts over wide rows can merge [datafusion]
via GitHub
-
Re: [PR] refactor: Make AggregateExec state modeling and public updates safe (part2: topk) [datafusion]
via GitHub
-
Re: [PR] fix: match iceberg-java for Iceberg partitions past year 262142 in the native writer [datafusion-comet]
via GitHub
-
[PR] build: bump semanticdb to 4.13.10 so scalafix runs on Spark 4.1 and 4.2 [datafusion-comet]
via GitHub
-
[PR] [X-3985] Sync branch-55 with upstream [datafusion]
via GitHub
-
Re: [I] A native plan with no JVM input returns truncated output without an error when the Tokio runtime shuts down [datafusion-comet]
via GitHub
-
[I] to_csv: ignoreLeadingWhiteSpace / ignoreTrailingWhiteSpace trim Unicode whitespace instead of univocity's chars <= ' ' [datafusion-comet]
via GitHub
-
Re: [I] A Rust UDF silently answers calls to an ordinary Scala UDF registered under the same name [datafusion-comet]
via GitHub
-
Re: [I] Reduce retained buffer allocation when list_extract selects short nested arrays [datafusion-comet]
via GitHub
-
Re: [PR] perf: mirror compound ASOF key predicates to right [datafusion]
via GitHub
-
[I] The SQL tab and event log lose the cached plan under CometInMemoryTableScan [datafusion-comet]
via GitHub
-
[I] S3 credential SPI: per-location credentials for native Iceberg reads and writes [datafusion-comet]
via GitHub
-
[PR] fix: run reverse of a binary value through the codegen dispatcher [datafusion-comet]
via GitHub
-
Re: [I] Improve performance of db-benchmark query 8 [datafusion]
via GitHub
-
Re: [I] Incorrect results: ORDER BY drops a sort key based on a nullable UNIQUE constraint [datafusion]
via GitHub
-
Re: [I] [EPIC] cast from string: trim semantics diverge from Spark across all numeric, datetime and boolean targets [datafusion-comet]
via GitHub
-
[I] reverse on a binary column fails with a native UTF-8 error on Spark 4.2 [datafusion-comet]
via GitHub
-
Re: [I] Remove AlignedArrowStreamReader and fix the Native to JVM section of ffi.md [datafusion-comet]
via GitHub
-
[PR] fix: describe CometInMemoryTableScan by the scan it replaces in EXPLAIN [datafusion-comet]
via GitHub
-
Re: [PR] Add related source code locations to errors [datafusion]
via GitHub
-
[PR] fix: fall back to Spark for regr_* aggregates until their merge matches Spark [datafusion-comet]
via GitHub
-
Re: [I] Nine expression benchmark rows labelled "Comet" are measuring Spark [datafusion-comet]
via GitHub
-
[I] Planning time grows quadratically with IN-list size on main [datafusion]
via GitHub
-
[PR] fix: coalesce shuffle partitions under Comet unions the way Spark does [datafusion-comet]
via GitHub
-
Re: [I] With several native plans in one task, the non-zero memory usage warning comes from the wrong plan [datafusion-comet]
via GitHub
-
Re: [I] Reuse the cached dictionary value→slot map in vectorized_equal_to instead of rebuilding it per batch [datafusion]
via GitHub
-
Re: [PR] fix: cancel background batch producers before final metrics [datafusion-comet]
via GitHub
-
Re: [PR] fix: treat typed nulls as missing bounds in aggregate dynamic filter merge [datafusion]
via GitHub
-
[PR] fix: make restricted_column run linearly [datafusion]
via GitHub
-
[PR] fix: cast native IF branches to a common type like CASE WHEN [datafusion-comet]
via GitHub
-
Re: [I] Spark `xxhash64` ignores validity buffer [datafusion]
via GitHub
-
Re: [PR] perf: PartitionedTopK hold retained rows in one shared RecordBatchStore [datafusion]
via GitHub
-
[PR] chore: gate Parquet opener encryption on `parquet_encryption` [datafusion]
via GitHub
-
[PR] fix: match Iceberg's rounding for pre-1970 timestamps in native years/months/days/hours [datafusion-comet]
via GitHub
-
Re: [PR] fix: preserve native sliding integer sum overflow semantics [datafusion-comet]
via GitHub
-
Re: [PR] fix(scheduler): read a swapped broadcast partitioned on the probe side [datafusion-ballista]
via GitHub
-
Re: [PR] fix: to_char returns row 0 for every row with a scalar datetime and a format column [datafusion]
via GitHub
-
[I] INSERT into a partitioned table rejects NOT NULL source columns for optional table columns [datafusion-iceberg]
via GitHub
-
[I] Iceberg plan nodes can't be rebuilt by a PhysicalExtensionCodec [datafusion-iceberg]
via GitHub
-
[I] Expose driver-controlled stage boundaries in physical plans [datafusion]
via GitHub
-
[PR] fix: match Spark's float ordering in min, max, greatest and least [datafusion-comet]
via GitHub
-
Re: [PR] fix: match Spark's float ordering in min, max, greatest and least [datafusion-comet]
via GitHub
-
Re: [PR] fix: match Spark's float ordering in min, max, greatest and least [datafusion-comet]
via GitHub
-
Re: [PR] fix: match Spark's float ordering in min, max, greatest and least [datafusion-comet]
via GitHub
-
Re: [PR] fix: match Spark's float ordering in min, max, greatest and least [datafusion-comet]
via GitHub
-
Re: [PR] fix: match Spark's float ordering in min, max, greatest and least [datafusion-comet]
via GitHub
-
Re: [PR] fix: match Spark's float ordering in min, max, greatest and least [datafusion-comet]
via GitHub
-
Re: [PR] fix: match Spark's float ordering in min, max, greatest and least [datafusion-comet]
via GitHub
-
Re: [I] Hash-based JVM columnar shuffle reports its output as disk spill instead of bytes written [datafusion-comet]
via GitHub
-
Re: [I] MIN dynamic filter race across partitions [datafusion]
via GitHub
-
Re: [PR] fix: reject groups accumulator for NULL-typed bitwise aggregates [datafusion]
via GitHub
-
Re: [I] An input ArrowArrayStream that native never takes is never released [datafusion-comet]
via GitHub
-
[PR] fix: rescale decimal results from dispatched functions to the declared type [datafusion-comet]
via GitHub
-
[I] feat: make `SharedCoalescer` do memory accounting for buffered batches [datafusion]
via GitHub
-
Re: [PR] perf: memoize Comet-subtree check in EliminateRedundantTransitions [datafusion-comet]
via GitHub
-
[I] [Proposal] Readable default column names for SQL query results [datafusion]
via GitHub
-
[I] AQE does not coalesce shuffle partitions under a CometUnion when a branch is not a shuffle [datafusion-comet]
via GitHub
-
[I] CometInMemoryTableScan prints its whole cached plan inline in EXPLAIN [datafusion-comet]
via GitHub
-
[PR] chore(deps): bump datafusion-sqllogictest from 55.0.0 to 55.1.0 [datafusion-iceberg]
via GitHub
-
[PR] chore(deps): bump datafusion-cli from 55.0.0 to 55.1.0 [datafusion-iceberg]
via GitHub
-
[PR] refactor: move the remaining PhysicalExpr impls from core into spark-expr [datafusion-comet]
via GitHub
-
[PR] fix: [branch-1.1] zero sliced boolean offsets at every level before exporting to the JVM (#6339) [datafusion-comet]
via GitHub
-
Re: [PR] perf: classify string-to-timestamp shapes in one byte scan (up to 5.6x faster) [datafusion-comet]
via GitHub
-
[I] Remove the legacy `GroupedHashAggregateStream` and `enable_migration_aggregate` [datafusion]
via GitHub