Messages by Thread
-
-
Re: [I] Push Down Offset to TableScan [datafusion]
via GitHub
-
Re: [PR] feat: Column statistics support v1 [datafusion-ballista]
via GitHub
-
Re: [PR] fix: deduplicate StringView/BinaryView buffer refs in CollectLeft Has... [datafusion]
via GitHub
-
[PR] chore(deps): bump pyjwt from 2.12.0 to 2.15.0 [datafusion-sandbox]
via GitHub
-
[PR] docs: add repository README [datafusion-iceberg]
via GitHub
-
Re: [PR] feat: admit string maps in Spark-to-Comet conversion [datafusion-comet]
via GitHub
-
Re: [I] Native implementation of `get_json_object` returns last value for duplicate keys, Spark returns first [datafusion-comet]
via GitHub
-
Re: [PR] perf(proto): avoid re-normalizing logical plans [datafusion]
via GitHub
-
[PR] fix: redact credentials in object-store URL diagnostics [datafusion]
via GitHub
-
[PR] perf: reduce Parquet runtime filter schema guard overhead [datafusion-comet]
via GitHub
-
Re: [PR] perf: avoid quadratic planning for SELECTs with many aggregates [datafusion]
via GitHub
-
[PR] build(deps): bump pyjwt from 2.13.0 to 2.15.0 [datafusion-python]
via GitHub
-
[PR] test: share collation coverage across Spark 4.x [datafusion-comet]
via GitHub
-
[PR] chore: move native write configs under spark.comet.write [datafusion-comet]
via GitHub
-
[I] Avoid redundant sorts for scalar subquery expressions [datafusion]
via GitHub
-
[I] Add checked time arithmetic regression coverage for scalar UDF type recovery [datafusion]
via GitHub
-
Re: [PR] test: enable SPARK-57298 collect_set tests [datafusion-comet]
via GitHub
-
Re: [I] native: `JVMClasses::with_env: JAVA_VM not initialized` aborts a multi-suite JVM [datafusion-comet]
via GitHub
-
[PR] fix: load the bundled native library only once per class loader [datafusion-comet]
via GitHub
-
Re: [PR] chore(deps): bump setuptools from 82.0.0 to 83.0.0 [datafusion-sandbox]
via GitHub
-
Re: [PR] chore(deps-dev): bump shell-quote from 1.8.3 to 1.10.0 in /datafusion/wasmtest/datafusion-wasm-app [datafusion-sandbox]
via GitHub
-
Re: [PR] feat: Refactor NLJ into an extensible framework for specialized joins [datafusion]
via GitHub
-
Re: [PR] fix: use build_join_schema for CrossJoinExec output schema metadata [datafusion]
via GitHub
-
[PR] test: [branch-1.1] cover sliced boolean arrays in explode (#6473) [datafusion-comet]
via GitHub
-
[PR] fix: [branch-1.1] fall back for incompatible regression aggregates (#6451) [datafusion-comet]
via GitHub
-
[PR] fix: [branch-1.1] rescale decimals in generated dispatchers (#6455) [datafusion-comet]
via GitHub
-
[PR] fix: [branch-1.1] coerce native IF branches to a common type (#6458) [datafusion-comet]
via GitHub
-
[PR] fix: [branch-1.1] make adaptive aggregation skipping opt-in (#6474) [datafusion-comet]
via GitHub
-
[PR] fix: [branch-1.1] fall back for rank limits over nested float keys (#6468) [datafusion-comet]
via GitHub
-
[PR] fix: [branch-1.1] match pre-epoch Iceberg temporal rounding (#6456) [datafusion-comet]
via GitHub
-
[PR] fix: Fix make_array null input handling [datafusion]
via GitHub
-
Re: [I] [Doc] CAST collated-string handling on Spark 4.0+ is implicit and untested [datafusion-comet]
via GitHub
-
[I] Assess native (Rust) code generation for fused expression evaluation [datafusion-comet]
via GitHub
-
Re: [I] Nested duplicate names in a metadata-free Parquet file bypass the native resolver when the file schema equals the requested schema [datafusion-comet]
via GitHub
-
[I] Filtered semi/anti sort-merge join is slow on small key groups and charges each group a whole batch [datafusion]
via GitHub
-
Re: [PR] fix: match Spark statistical aggregate updates and merges [datafusion-comet]
via GitHub
-
Re: [I] `native_datafusion` performance improvement [datafusion-comet]
via GitHub
-
Re: [I] Enable TopK dynamic filter pushdown into native Parquet scans [datafusion-comet]
via GitHub
-
[I] Native existence join enumerates every duplicate build match (N*M candidates for M markers) [datafusion-comet]
via GitHub
-
[PR] docs: request changes for correctness and performance findings in the review skill [datafusion-comet]
via GitHub
-
Re: [PR] fix: make Iceberg delete-file reflection failures fatal [datafusion-comet]
via GitHub
-
[I] Native CASE WHEN names its struct result's fields after the ELSE branch instead of the first THEN branch [datafusion-comet]
via GitHub
-
[I] Native corr, covariance, variance and stddev return wrong values for a constant fractional column merged from several partitions [datafusion-comet]
via GitHub
-
[PR] fix: preserve compound expression equivalences through projection [datafusion]
via GitHub
-
[PR] fix: Route collated instr, substring_index, trim, greatest and least through the codegen dispatcher [datafusion-comet]
via GitHub
-
Re: [PR] fix: Route collated instr, substring_index, trim, greatest and least through the codegen dispatcher [datafusion-comet]
via GitHub
-
Re: [PR] fix: Route collated instr, substring_index, trim, greatest and least through the codegen dispatcher [datafusion-comet]
via GitHub
-
Re: [PR] fix: Route collated instr, substring_index, trim, greatest and least through the codegen dispatcher [datafusion-comet]
via GitHub
-
Re: [PR] fix: Route collated instr, substring_index, trim, greatest and least through the codegen dispatcher [datafusion-comet]
via GitHub
-
[I] LEFT JOIN with struct-field ORDER BY fails ordering validation [datafusion]
via GitHub
-
[PR] bench-mark-no-take [datafusion]
via GitHub
-
[PR] feat: per-location credentials for native Iceberg reads and writes [datafusion-comet]
via GitHub
-
[PR] chore(deps): bump urllib3 from 2.7.0 to 2.8.0 [datafusion]
via GitHub
-
[PR] chore(deps): bump pyjwt from 2.13.0 to 2.15.0 [datafusion]
via GitHub
-
[PR] fix: match Spark's float ordering for sort and window keys nested in arrays and structs [datafusion-comet]
via GitHub
-
[PR] chore: require review conversations to be resolved before merging to main [datafusion-comet]
via GitHub
-
[I] RANGE window frames over an array or struct key with a null element span the whole partition [datafusion-comet]
via GitHub
-
[I] Native sort orders null elements of array and struct keys by the key's null order, unlike Spark [datafusion-comet]
via GitHub
-
Re: [I] Preserve scalar UDF output types in property analysis with a central fallback [datafusion]
via GitHub
-
[PR] Add SECURITY.md following ASF guidelines [datafusion]
via GitHub
-
[I] Add a security policy [datafusion]
via GitHub
-
[PR] fix: make adaptive partial aggregation opt-in with spark.comet.exec.aggregate.skipPartial.enabled [datafusion-comet]
via GitHub
-
[PR] Update SECURITY policy to conform to ASF rules [datafusion-sqlparser-rs]
via GitHub
-
[I] Update SECURITY policy to conform to ASF rules [datafusion-sqlparser-rs]
via GitHub
-
[PR] test: cover booleans in native explode output past the first batch [datafusion-comet]
via GitHub
-
[PR] [branch-55] fix: Remove `serde_json/preserve_order` feature from library crates (backport #25884) [datafusion]
via GitHub
-
[I] array_contains, arrays_overlap, array_distinct and array_union ignore string collation [datafusion-comet]
via GitHub
-
[PR] fix: Route collated array element comparisons through the codegen dispatcher [datafusion-comet]
via GitHub
-
[I] instr, substring_index, trim with a trim string, greatest and least ignore string collation [datafusion-comet]
via GitHub
-
Re: [PR] fix: handle dictionary nulls in null-aware left mark joins [datafusion]
via GitHub
-
Re: [I] Improve performance of conditional expressions [datafusion-comet]
via GitHub
-
[PR] fix: clear_shrink(0) on rows should reset to original size + single group by fixes in clear_shrink [datafusion]
via GitHub
-
[I] Spark `xxhash64` hashes the raw bits of a NaN instead of the canonical NaN [datafusion]
via GitHub
-
[PR] docs: add release notes for 1.1.0 with known regressions and workarounds [datafusion-comet]
via GitHub
-
[PR] fix: fall back to Spark for RANK and DENSE_RANK limits over nested floating-point keys [datafusion-comet]
via GitHub
-
Re: [I] Reduce CI time (Comet is consuming 50% of DataFusion CI, which is a lot) [datafusion-comet]
via GitHub
-
[I] Investigate nested TPC-H q21 slowdown: Comet 37% slower than Spark at SF1000 [datafusion-comet]
via GitHub
-
[PR] perf: optimize dictionary float-zero normalization [datafusion]
via GitHub
-
Re: [PR] fix: Avoid pushing down sort under limits [datafusion]
via GitHub
-
[I] Trying to reduce CI time [datafusion-comet]
via GitHub