Messages by Thread
-
-
[I] Native corr, covariance, variance and stddev return wrong values for a constant fractional column merged from several partitions [datafusion-comet]
via GitHub
-
[PR] fix: preserve compound expression equivalences through projection [datafusion]
via GitHub
-
[PR] fix: Route collated instr, substring_index, trim, greatest and least through the codegen dispatcher [datafusion-comet]
via GitHub
-
[I] LEFT JOIN with struct-field ORDER BY fails ordering validation [datafusion]
via GitHub
-
[PR] bench-mark-no-take [datafusion]
via GitHub
-
[PR] feat: per-location credentials for native Iceberg reads and writes [datafusion-comet]
via GitHub
-
[PR] chore(deps): bump urllib3 from 2.7.0 to 2.8.0 [datafusion]
via GitHub
-
[PR] chore(deps): bump pyjwt from 2.13.0 to 2.15.0 [datafusion]
via GitHub
-
[PR] fix: match Spark's float ordering for sort and window keys nested in arrays and structs [datafusion-comet]
via GitHub
-
[PR] chore: require review conversations to be resolved before merging to main [datafusion-comet]
via GitHub
-
[I] RANGE window frames over an array or struct key with a null element span the whole partition [datafusion-comet]
via GitHub
-
[I] Native sort orders null elements of array and struct keys by the key's null order, unlike Spark [datafusion-comet]
via GitHub
-
Re: [I] Preserve scalar UDF output types in property analysis with a central fallback [datafusion]
via GitHub
-
[PR] Add SECURITY.md following ASF guidelines [datafusion]
via GitHub
-
[I] Add a security policy [datafusion]
via GitHub
-
[PR] fix: make adaptive partial aggregation opt-in with spark.comet.exec.aggregate.skipPartial.enabled [datafusion-comet]
via GitHub
-
[PR] Update SECURITY policy to conform to ASF rules [datafusion-sqlparser-rs]
via GitHub
-
[I] Update SECURITY policy to conform to ASF rules [datafusion-sqlparser-rs]
via GitHub
-
[PR] test: cover booleans in native explode output past the first batch [datafusion-comet]
via GitHub
-
[PR] [branch-55] fix: Remove `serde_json/preserve_order` feature from library crates (backport #25884) [datafusion]
via GitHub
-
[I] array_contains, arrays_overlap, array_distinct and array_union ignore string collation [datafusion-comet]
via GitHub
-
[PR] fix: Route collated array element comparisons through the codegen dispatcher [datafusion-comet]
via GitHub
-
[I] instr, substring_index, trim with a trim string, greatest and least ignore string collation [datafusion-comet]
via GitHub
-
Re: [I] Improve performance of conditional expressions [datafusion-comet]
via GitHub
-
[PR] fix: clear_shrink(0) on rows should reset to original size + single group by fixes in clear_shrink [datafusion]
via GitHub
-
[I] Spark `xxhash64` hashes the raw bits of a NaN instead of the canonical NaN [datafusion]
via GitHub
-
[PR] docs: add release notes for 1.1.0 with known regressions and workarounds [datafusion-comet]
via GitHub
-
[PR] fix: fall back to Spark for RANK and DENSE_RANK limits over nested floating-point keys [datafusion-comet]
via GitHub
-
Re: [I] Reduce CI time (Comet is consuming 50% of DataFusion CI, which is a lot) [datafusion-comet]
via GitHub
-
[I] Investigate nested TPC-H q21 slowdown: Comet 37% slower than Spark at SF1000 [datafusion-comet]
via GitHub
-
[PR] perf: optimize dictionary float-zero normalization [datafusion]
via GitHub
-
Re: [PR] fix: Avoid pushing down sort under limits [datafusion]
via GitHub
-
[I] Trying to reduce CI time [datafusion-comet]
via GitHub
-
[I] Adaptive partial aggregation can shuffle 10x more rows than 1.0.0, with no supported way to turn it off [datafusion-comet]
via GitHub
-
[I] Native explode returns wrong booleans for arrays of structs after the first batch of output [datafusion-comet]
via GitHub
-
Re: [PR] perf: Optimize `array_agg(... IGNORE NULLS)` [datafusion]
via GitHub
-
Re: [PR] bench: add `array_compact` benchmarks [datafusion]
via GitHub
-
[PR] fix: Avoid special-casing `count(1)` in `Count::value_from_stats` [datafusion]
via GitHub
-
[PR] fix: bound spill file batches by bytes so external sorts over wide rows can merge [datafusion]
via GitHub
-
Re: [PR] refactor: Make AggregateExec state modeling and public updates safe (part2: topk) [datafusion]
via GitHub
-
[PR] build: bump semanticdb to 4.13.10 so scalafix runs on Spark 4.1 and 4.2 [datafusion-comet]
via GitHub
-
[PR] [X-3985] Sync branch-55 with upstream [datafusion]
via GitHub
-
[I] to_csv: ignoreLeadingWhiteSpace / ignoreTrailingWhiteSpace trim Unicode whitespace instead of univocity's chars <= ' ' [datafusion-comet]
via GitHub
-
Re: [I] A Rust UDF silently answers calls to an ordinary Scala UDF registered under the same name [datafusion-comet]
via GitHub
-
Re: [I] Reduce retained buffer allocation when list_extract selects short nested arrays [datafusion-comet]
via GitHub
-
[I] The SQL tab and event log lose the cached plan under CometInMemoryTableScan [datafusion-comet]
via GitHub
-
[I] S3 credential SPI: per-location credentials for native Iceberg reads and writes [datafusion-comet]
via GitHub
-
[PR] fix: run reverse of a binary value through the codegen dispatcher [datafusion-comet]
via GitHub
-
Re: [I] Improve performance of db-benchmark query 8 [datafusion]
via GitHub
-
Re: [I] Incorrect results: ORDER BY drops a sort key based on a nullable UNIQUE constraint [datafusion]
via GitHub
-
Re: [I] [EPIC] cast from string: trim semantics diverge from Spark across all numeric, datetime and boolean targets [datafusion-comet]
via GitHub
-
[I] reverse on a binary column fails with a native UTF-8 error on Spark 4.2 [datafusion-comet]
via GitHub
-
[PR] fix: describe CometInMemoryTableScan by the scan it replaces in EXPLAIN [datafusion-comet]
via GitHub
-
Re: [PR] Add related source code locations to errors [datafusion]
via GitHub
-
[PR] fix: fall back to Spark for regr_* aggregates until their merge matches Spark [datafusion-comet]
via GitHub
-
Re: [I] Nine expression benchmark rows labelled "Comet" are measuring Spark [datafusion-comet]
via GitHub
-
[I] Planning time grows quadratically with IN-list size on main [datafusion]
via GitHub
-
[PR] fix: coalesce shuffle partitions under Comet unions the way Spark does [datafusion-comet]
via GitHub
-
Re: [I] Reuse the cached dictionary value→slot map in vectorized_equal_to instead of rebuilding it per batch [datafusion]
via GitHub
-
Re: [PR] fix: treat typed nulls as missing bounds in aggregate dynamic filter merge [datafusion]
via GitHub
-
[PR] fix: make restricted_column run linearly [datafusion]
via GitHub
-
[PR] fix: cast native IF branches to a common type like CASE WHEN [datafusion-comet]
via GitHub
-
Re: [I] Spark `xxhash64` ignores validity buffer [datafusion]
via GitHub
-
[PR] chore: gate Parquet opener encryption on `parquet_encryption` [datafusion]
via GitHub
-
[PR] fix: match Iceberg's rounding for pre-1970 timestamps in native years/months/days/hours [datafusion-comet]
via GitHub
-
Re: [PR] fix: preserve native sliding integer sum overflow semantics [datafusion-comet]
via GitHub
-
Re: [PR] fix: to_char returns row 0 for every row with a scalar datetime and a format column [datafusion]
via GitHub
-
[I] INSERT into a partitioned table rejects NOT NULL source columns for optional table columns [datafusion-iceberg]
via GitHub
-
[I] Iceberg plan nodes can't be rebuilt by a PhysicalExtensionCodec [datafusion-iceberg]
via GitHub
-
[I] Expose driver-controlled stage boundaries in physical plans [datafusion]
via GitHub
-
[PR] fix: match Spark's float ordering in min, max, greatest and least [datafusion-comet]
via GitHub
-
Re: [I] MIN dynamic filter race across partitions [datafusion]
via GitHub
-
Re: [PR] fix: reject groups accumulator for NULL-typed bitwise aggregates [datafusion]
via GitHub
-
[PR] fix: rescale decimal results from dispatched functions to the declared type [datafusion-comet]
via GitHub
-
[I] feat: make `SharedCoalescer` do memory accounting for buffered batches [datafusion]
via GitHub
-
Re: [PR] perf: memoize Comet-subtree check in EliminateRedundantTransitions [datafusion-comet]
via GitHub
-
[I] [Proposal] Readable default column names for SQL query results [datafusion]
via GitHub