Messages by Thread
-
-
[PR] Add a test to enforce compatibility between an aggregate's accumulators [datafusion]
via GitHub
-
Re: [I] [EPIC] Use blocked / chunked memory management in hash aggregation [datafusion]
via GitHub
-
[I] Charge the native shuffle writer's write buffers to the memory pool [datafusion-comet]
via GitHub
-
Re: [PR] chore: Use preinstalled clang for wasm CI job [datafusion]
via GitHub
-
Re: [PR] feat: positional round robin shuffle keyed on a row ordinal [datafusion-comet]
via GitHub
-
Re: [PR] chore: extend pre commit instructions for AI agents [datafusion]
via GitHub
-
Re: [PR] feat: prune-only transfer of parent and dynamic filters across HashJoinExec keys (left, right and mark joins) [datafusion]
via GitHub
-
[PR] docs: correct what the Comet plugin does in the plugin overview [datafusion-comet]
via GitHub
-
[PR] fix: do not run Comet in on-heap mode without `spark.comet.exec.onHeap.enabled` [datafusion-comet]
via GitHub
-
Re: [PR] docs: update the user guide for the 1.1.0 release [datafusion-comet]
via GitHub
-
Re: [PR] feat: add spark.comet.explain.planOnly.enabled [datafusion-comet]
via GitHub
-
Re: [PR] chore: deprecate spark.comet.exec.memoryPool.fraction [datafusion-comet]
via GitHub
-
Re: [PR] fix: preserve ProjectionExec metadata during serialization [datafusion]
via GitHub
-
Re: [PR] Feat(parquet) : Introduce roundtrip number Distinct Valeus [datafusion]
via GitHub
-
Re: [PR] Check in Cargo.lock [datafusion-sqlparser-rs]
via GitHub
-
[I] Native Parquet scan reads nested fields by position when field ids no longer match their names [datafusion-comet]
via GitHub
-
Re: [I] Three dev/diffs weaken a local-shuffle-read assertion into a contradiction, breaking the Spark-only baseline [datafusion-comet]
via GitHub
-
Re: [I] Session-level SecretsStore extension point and CREATE SECRET / DROP SECRET support [datafusion]
via GitHub
-
Re: [PR] Snowflake: Add support for ->> (pipe) operator for chaining SQL stmts [datafusion-sqlparser-rs]
via GitHub
-
[I] `ExtractANSIIntervalDays` and the other interval field extractors have no serde, so `date + <day interval column>` and `extract` of an interval fall back [datafusion-comet]
via GitHub
-
[PR] fix: read shuffle write buffer, spill limit and off-heap sizes in bytes [datafusion-comet]
via GitHub
-
Re: [PR] fix: keep a correlated filter below an aggregate with a grouping set [datafusion]
via GitHub
-
[PR] ci: use RustFS instead of MinIO for datafusion-cli S3 tests [datafusion]
via GitHub
-
[PR] ci: run the S3 integration tests against RustFS instead of MinIO [datafusion-ballista]
via GitHub
-
[I] `spark.comet.maxTempDirectorySize` silently ignores values with a unit [datafusion-comet]
via GitHub
-
[I] `spark.comet.shuffle.native.writeBufferSize` is sent to native code in MiB but used as bytes, so the default write buffer is 1 byte [datafusion-comet]
via GitHub
-
[I] Kubernetes guide example never enables Comet because it sets no off-heap memory [datafusion-comet]
via GitHub
-
Re: [PR] perf: mirror ASOF equality-key filters to right [datafusion]
via GitHub
-
[I] Driver warns that `spark.executor.memoryOverhead` is unset when `spark.executor.memoryOverheadFactor` is set, and in local mode [datafusion-comet]
via GitHub
-
[I] A bare byte count for `spark.memory.offHeap.size` is read as MiB when sizing the Comet memory pool [datafusion-comet]
via GitHub
-
[I] Native memory usage log underestimates the executor overhead for PySpark and SparkR on Kubernetes, and warns on standalone clusters [datafusion-comet]
via GitHub
-
[I] Comet runs in on-heap mode without `spark.comet.exec.onHeap.enabled` when the session extension is registered directly [datafusion-comet]
via GitHub
-
[PR] fix: remove misleading native opt-in for dispatch-only datetime expressions [datafusion-comet]
via GitHub
-
[PR] feat(agg): add initial Blocked aggregate api [datafusion]
via GitHub
-
[I] Decorrelate subqueries whose grouping sets leave out the correlated column [datafusion]
via GitHub
-
Re: [PR] perf: Optimize `array_compact` [datafusion]
via GitHub
-
Re: [PR] docs: make doc changes weekly comet sync [datafusion-comet]
via GitHub
-
Re: [PR] perf: project cached batches by buffer selection, prune on collated strings [datafusion-comet]
via GitHub
-
[I] ci: datafusion-cli S3 tests fail because the MinIO image is not available [datafusion]
via GitHub
-
Re: [PR] refactor: track NestedLoopJoin fallback matches in one left bitmap and drop the cancel protocol [datafusion]
via GitHub
-
[I] `EliminateCrossJoin` drops `NullEqualsNull` from nested inner joins [datafusion]
via GitHub
-
Re: [I] Try and discourage use of `git push --force` [datafusion-comet]
via GitHub
-
[PR] fix: refuse codegen dispatch for a TRY cast that can put a null key in a map [datafusion-comet]
via GitHub
-
Re: [PR] Postgres: Support XML functions (XMLELEMENT, XMLPI, XMLROOT, XMLSERIALIZE, XMLEXISTS) [datafusion-sqlparser-rs]
via GitHub
-
Re: [I] to_unix_timestamp and make_timestamp advertise an allowIncompatible native opt-in that does not exist [datafusion-comet]
via GitHub
-
Re: [PR] fix: avoid full listings for cached pruned partitions [datafusion]
via GitHub
-
Re: [PR] fix: correct `LEAD/LAG IGNORE NULLS` evaluation and limit pushdown [datafusion]
via GitHub
-
Re: [I] `LEAD/LAG IGNORE NULLS` returns incorrect results across `NULL` gaps and with `LIMIT` [datafusion]
via GitHub
-
Re: [PR] fix: reserve the concatenated build batch and computed join keys in HashJoinExec [datafusion]
via GitHub
-
Re: [PR] fix: request a null-aware mark join wherever a NULL mark is observable [datafusion]
via GitHub
-
Re: [PR] Snowflake: parse ORDER BY ALL [datafusion-sqlparser-rs]
via GitHub
-
Re: [I] coalesce(Int, Utf8) coerces to Int [datafusion]
via GitHub
-
Re: [PR] fix: match Spark's duplicate field and field id semantics in parquet field lookup [datafusion-comet]
via GitHub
-
Re: [PR] Fix coalesce string numeric union coercion [datafusion]
via GitHub
-
Re: [I] chore: Re-enable unused-parameter warnings in the strict-warnings profile [datafusion-comet]
via GitHub
-
[PR] Record the first and last token of ALTER statements for their spans [datafusion-sqlparser-rs]
via GitHub
-
Re: [PR] DuckDB: Support SET VARIABLE [datafusion-sqlparser-rs]
via GitHub
-
Re: [PR] feat: recognize more lossless casts for statistics and ordering [datafusion]
via GitHub
-
Re: [PR] perf(functions-aggregate): coalesce small ordered ARRAY_AGG batches [datafusion]
via GitHub
-
Re: [PR] fix: Derive Substrait intersection nullability from every input [datafusion]
via GitHub
-
Re: [PR] DuckDB dialect: support % on LIMIT, like `LIMIT 5% OFFSET 20` [datafusion-sqlparser-rs]
via GitHub
-
Re: [I] Fix licensing issues in source release found during 0.63.0 RC1 vote (missing NOTICE file, flamegraph.svg with CDDL-licensed JavaScript) [datafusion-sqlparser-rs]
via GitHub
-
Re: [PR] Replace benchmark flamegraph.svg with png [datafusion-sqlparser-rs]
via GitHub