github
Thread
Date
Earlier messages
Later messages
Messages by Thread
Re: [PR] fix: reject casts involving non-default collated strings [datafusion-comet]
via GitHub
Re: [PR] fix: reject casts involving non-default collated strings [datafusion-comet]
via GitHub
Re: [PR] fix: reject casts involving non-default collated strings [datafusion-comet]
via GitHub
Re: [I] Make `datafusion_substrait` not depend on `SessionState` [datafusion]
via GitHub
Re: [I] Make `datafusion_substrait` not depend on `SessionState` [datafusion]
via GitHub
Re: [I] Make `datafusion_substrait` not depend on `SessionState` [datafusion]
via GitHub
Re: [PR] fix: coerce timezone-naive minus timezone-aware timestamps at equal time units [datafusion]
via GitHub
Re: [PR] fix: coerce timezone-naive minus timezone-aware timestamps at equal time units [datafusion]
via GitHub
[I] Support decorrelating subqueries with a correlated filter below the nullable side of an outer join [datafusion]
via GitHub
Re: [I] Support decorrelating subqueries with a correlated filter below the nullable side of an outer join [datafusion]
via GitHub
Re: [PR] fix: Check Any, not the delegating downcast, in FFI_ExecutionPlan::new's foreign-echo shortcut [datafusion]
via GitHub
Re: [PR] fix: Check Any, not the delegating downcast, in FFI_ExecutionPlan::new's foreign-echo shortcut [datafusion]
via GitHub
Re: [PR] fix: Check Any, not the delegating downcast, in FFI_ExecutionPlan::new's foreign-echo shortcut [datafusion]
via GitHub
[PR] test: compare in-memory caches with independent answers [datafusion-comet]
via GitHub
Re: [PR] test: compare in-memory caches with independent answers [datafusion-comet]
via GitHub
Re: [PR] test: compare in-memory caches with independent answers [datafusion-comet]
via GitHub
Re: [PR] test: compare in-memory caches with independent answers [datafusion-comet]
via GitHub
Re: [PR] test: compare in-memory caches with independent answers [datafusion-comet]
via GitHub
Re: [I] Estimate Decimal filter selectivity [datafusion]
via GitHub
[PR] fix: Validate `execution.time_zone` at `SET` time [datafusion]
via GitHub
Re: [PR] fix: Validate `execution.time_zone` at `SET` time [datafusion]
via GitHub
Re: [PR] fix: Validate `execution.time_zone` at `SET` time [datafusion]
via GitHub
Re: [PR] fix: Validate `execution.time_zone` at `SET` time [datafusion]
via GitHub
Re: [PR] fix: Validate `execution.time_zone` at `SET` time [datafusion]
via GitHub
Re: [PR] fix: Validate `execution.time_zone` at `SET` time [datafusion]
via GitHub
Re: [PR] feat: add native Delta Lake scan contrib module (page/row-group pruning) [datafusion-comet]
via GitHub
Re: [PR] feat: add native Delta Lake scan contrib module (page/row-group pruning) [datafusion-comet]
via GitHub
Re: [PR] feat: add native Delta Lake scan contrib module (page/row-group pruning) [datafusion-comet]
via GitHub
Re: [PR] feat: add native Delta Lake scan contrib module (page/row-group pruning) [datafusion-comet]
via GitHub
Re: [PR] feat: add native Delta Lake scan contrib module (page/row-group pruning) [datafusion-comet]
via GitHub
Re: [PR] feat: add native Delta Lake scan contrib module (page/row-group pruning) [datafusion-comet]
via GitHub
Re: [PR] feat: add native Delta Lake scan contrib module (page/row-group pruning) [datafusion-comet]
via GitHub
Re: [PR] feat: add native Delta Lake scan contrib module (page/row-group pruning) [datafusion-comet]
via GitHub
Re: [PR] feat: add native Delta Lake scan contrib module (page/row-group pruning) [datafusion-comet]
via GitHub
Re: [I] The config values aren't checked [datafusion]
via GitHub
Re: [I] Untyped aliased placeholder in a `UNION` arm fails with "Schema error: No field named <alias>" when the plan is analyzed or optimized [datafusion]
via GitHub
Re: [PR] test: PostgreSQL differential coverage for timestamps with time zone [datafusion]
via GitHub
Re: [PR] perf: stop holding the fair pool lock across blocking memory calls [datafusion-comet]
via GitHub
Re: [PR] perf: stop holding the fair pool lock across blocking memory calls [datafusion-comet]
via GitHub
Re: [PR] perf: stop holding the fair pool lock across blocking memory calls [datafusion-comet]
via GitHub
Re: [PR] perf: stop holding the fair pool lock across blocking memory calls [datafusion-comet]
via GitHub
Re: [PR] perf: stop holding the fair pool lock across blocking memory calls [datafusion-comet]
via GitHub
Re: [PR] perf: stop holding the fair pool lock across blocking memory calls [datafusion-comet]
via GitHub
Re: [PR] perf: stop holding the fair pool lock across blocking memory calls [datafusion-comet]
via GitHub
Re: [PR] perf: stop holding the fair pool lock across blocking memory calls [datafusion-comet]
via GitHub
Re: [PR] perf: stop holding the fair pool lock across blocking memory calls [datafusion-comet]
via GitHub
Re: [PR] feat: route date and timestamp interval arithmetic through codegen dispatch [datafusion-comet]
via GitHub
Re: [PR] feat: route date and timestamp interval arithmetic through codegen dispatch [datafusion-comet]
via GitHub
Re: [PR] feat: route date and timestamp interval arithmetic through codegen dispatch [datafusion-comet]
via GitHub
Re: [PR] fix: plan UNION of untyped aliased placeholders [datafusion]
via GitHub
Re: [PR] fix: plan UNION of untyped aliased placeholders [datafusion]
via GitHub
Re: [PR] fix: dispatch map lookups with normalized keys and nondeterministic null-guarded children [datafusion-comet]
via GitHub
Re: [PR] fix: dispatch map lookups with normalized keys and nondeterministic null-guarded children [datafusion-comet]
via GitHub
Re: [PR] fix: dispatch map lookups with normalized keys and nondeterministic null-guarded children [datafusion-comet]
via GitHub
Re: [PR] fix: honor the S3 profile name and file and Hadoop's addressing mode for custom endpoints [datafusion-comet]
via GitHub
Re: [PR] fix: honor the S3 profile name and file and Hadoop's addressing mode for custom endpoints [datafusion-comet]
via GitHub
Re: [PR] fix: honor the S3 profile name and file and Hadoop's addressing mode for custom endpoints [datafusion-comet]
via GitHub
Re: [PR] fix: honor the S3 profile name and file and Hadoop's addressing mode for custom endpoints [datafusion-comet]
via GitHub
Re: [PR] fix: honor the S3 profile name and file and Hadoop's addressing mode for custom endpoints [datafusion-comet]
via GitHub
Re: [PR] fix: report task input metrics after the native iterator closes and add to Spark's counters [datafusion-comet]
via GitHub
Re: [PR] fix: report task input metrics after the native iterator closes and add to Spark's counters [datafusion-comet]
via GitHub
Re: [PR] fix: let explicit Hadoop Azure auth outrank ambient AZURE_* environment variables [datafusion-comet]
via GitHub
Re: [PR] fix: let explicit Hadoop Azure auth outrank ambient AZURE_* environment variables [datafusion-comet]
via GitHub
Re: [PR] fix: let explicit Hadoop Azure auth outrank ambient AZURE_* environment variables [datafusion-comet]
via GitHub
Re: [PR] fix: let explicit Hadoop Azure auth outrank ambient AZURE_* environment variables [datafusion-comet]
via GitHub
Re: [PR] fix: let explicit Hadoop Azure auth outrank ambient AZURE_* environment variables [datafusion-comet]
via GitHub
Re: [PR] feat: aggregate large hash tables as hash buckets (experimental, off by default) [datafusion]
via GitHub
Re: [PR] feat: aggregate large hash tables as hash buckets (experimental, off by default) [datafusion]
via GitHub
Re: [PR] feat: aggregate a large final hash table as hash buckets (experimental, off by default) [datafusion]
via GitHub
Re: [PR] feat: aggregate a large final hash table as hash buckets (experimental, off by default) [datafusion]
via GitHub
Re: [PR] feat: aggregate a large final hash table as hash buckets (experimental, off by default) [datafusion]
via GitHub
Re: [PR] feat: aggregate a large final hash table as hash buckets (experimental, off by default) [datafusion]
via GitHub
Re: [PR] feat: aggregate a large final hash table as hash buckets (experimental, off by default) [datafusion]
via GitHub
Re: [PR] feat: aggregate a large final hash table as hash buckets (experimental, off by default) [datafusion]
via GitHub
Re: [PR] feat: aggregate a large final hash table as hash buckets (experimental, off by default) [datafusion]
via GitHub
Re: [PR] feat: aggregate a large final hash table as hash buckets (experimental, off by default) [datafusion]
via GitHub
Re: [PR] feat: aggregate a large final hash table as hash buckets (experimental, off by default) [datafusion]
via GitHub
Re: [I] Field id gating differs from Spark: root-only check and no dependence on fieldId.read.enabled [datafusion-comet]
via GitHub
Re: [PR] fix: reject a file without field ids at any depth whether or not id matching is on [datafusion-comet]
via GitHub
[PR] fix: consider input sizes before reusing join partitioning [datafusion]
via GitHub
Re: [PR] fix: consider input sizes before reusing join partitioning [datafusion]
via GitHub
Re: [I] Estimate group counts without explicit NDVs [datafusion]
via GitHub
[PR] fix: Error if both `group_by` and `aggr` is empty in `AggregateExec::try_new()` [datafusion]
via GitHub
Re: [PR] fix: Error if both `group_by` and `aggr` is empty in `AggregateExec::try_new()` [datafusion]
via GitHub
Re: [PR] fix: Error if both `group_by` and `aggr` is empty in `AggregateExec::try_new()` [datafusion]
via GitHub
Re: [PR] fix: Error if both `group_by` and `aggr` is empty in `AggregateExec::try_new()` [datafusion]
via GitHub
Re: [PR] feat: add opt-in local TopK fusion for native Parquet scans [datafusion-comet]
via GitHub
Re: [PR] feat: add opt-in local TopK fusion for native Parquet scans [datafusion-comet]
via GitHub
Re: [PR] ci: sweep TPC-H/TPC-DS verification over partition counts and empty tables [datafusion-ballista]
via GitHub
Re: [PR] chore(deps): bump quinn-proto from 0.11.14 to 0.11.16 [datafusion-sandbox]
via GitHub
Re: [PR] feat: support pushdown-aware dynamic filter [datafusion]
via GitHub
Re: [PR] fix: reproduce function registry deterministically in SessionStateBuilder::new_from_existing [datafusion]
via GitHub
Re: [PR] fix: reproduce function registry deterministically in SessionStateBuilder::new_from_existing [datafusion]
via GitHub
Re: [PR] add data_page_compression_ratio_threshold to config [datafusion]
via GitHub
Re: [PR] add data_page_compression_ratio_threshold to config [datafusion]
via GitHub
Re: [PR] add data_page_compression_ratio_threshold to config [datafusion]
via GitHub
Re: [PR] add data_page_compression_ratio_threshold to config [datafusion]
via GitHub
[PR] feat: add builders for `CreateExternalCatalog` and `DropCatalog` [datafusion]
via GitHub
Re: [PR] feat: add builders for `CreateExternalCatalog` and `DropCatalog` [datafusion]
via GitHub
Re: [PR] feat: add builders for `CreateExternalCatalog` and `DropCatalog` [datafusion]
via GitHub
Re: [PR] feat: add builders for `CreateExternalCatalog` and `DropCatalog` [datafusion]
via GitHub
[PR] fix: validate parquet encoding when configured [datafusion]
via GitHub
[I] JVM columnar shuffle fails with ArrayIndexOutOfBoundsException when spark.shuffle.checksum.enabled=false [datafusion-comet]
via GitHub
[PR] refactor: move the Flight proxy service into ballista-core [datafusion-ballista]
via GitHub
Re: [PR] refactor: move the Flight proxy service into ballista-core [datafusion-ballista]
via GitHub
Re: [PR] refactor: move the Flight proxy service into ballista-core [datafusion-ballista]
via GitHub
[PR] fix: return a native plan's memory before its Spark task ends [datafusion-comet]
via GitHub
Re: [PR] fix: return a native plan's memory before its Spark task ends [datafusion-comet]
via GitHub
Re: [PR] fix: return a native plan's memory before its Spark task ends [datafusion-comet]
via GitHub
Re: [PR] fix: return a native plan's memory before its Spark task ends [datafusion-comet]
via GitHub
Re: [PR] fix: return a native plan's memory before its Spark task ends [datafusion-comet]
via GitHub
Re: [PR] fix: return a native plan's memory before its Spark task ends [datafusion-comet]
via GitHub
Re: [I] Tuning guide does not explain that Comet's memory comes out of the executor container [datafusion-comet]
via GitHub
Re: [I] Kubernetes guide example never enables Comet because it sets no off-heap memory [datafusion-comet]
via GitHub
Re: [I] Investigate adopting DataFusion's allocator-level memory accounting to replace manual memory tuning [datafusion-comet]
via GitHub
Re: [I] Investigate adopting DataFusion's allocator-level memory accounting to replace manual memory tuning [datafusion-comet]
via GitHub
Re: [I] Consider accounting for Comet in spark.executor.memoryOverhead when off-heap memory is enabled [datafusion-comet]
via GitHub
Re: [I] Consider accounting for Comet in spark.executor.memoryOverhead when off-heap memory is enabled [datafusion-comet]
via GitHub
Re: [I] Memory management cgroup diagram does not show which config value sizes each region [datafusion-comet]
via GitHub
Re: [I] Memory management cgroup diagram does not show which config value sizes each region [datafusion-comet]
via GitHub
Re: [I] Move offHeap enabled check [datafusion-comet]
via GitHub
Re: [I] Move offHeap enabled check [datafusion-comet]
via GitHub
Re: [I] Use Spark ResourceProfiles to give native and JVM stages independent memory configs [datafusion-comet]
via GitHub
Re: [PR] feat: account JVM UDF Arrow allocations in Spark task memory [datafusion-comet]
via GitHub
Re: [PR] feat: account JVM UDF Arrow allocations in Spark task memory [datafusion-comet]
via GitHub
Re: [I] Follow-ups to #6163: correct the `fair_unified` description and prepare `spark.comet.exec.memoryPool.fraction` for removal [datafusion-comet]
via GitHub
Re: [I] Possible memory leak in error cases in JVM-Rust (and vice versa) transition [datafusion-comet]
via GitHub
Re: [I] Possible memory leak in error cases in JVM-Rust (and vice versa) transition [datafusion-comet]
via GitHub
Re: [I] Possible memory leak in error cases in JVM-Rust (and vice versa) transition [datafusion-comet]
via GitHub
Re: [I] Cancel background batch producers before collecting final plan metrics [datafusion-comet]
via GitHub
Re: [I] Cancel background batch producers before collecting final plan metrics [datafusion-comet]
via GitHub
Re: [I] [EPIC] Memory pool and accounting audit sweep [datafusion-comet]
via GitHub
Re: [I] [EPIC] Memory pool and accounting audit sweep [datafusion-comet]
via GitHub
Re: [I] [EPIC] Memory pool and accounting audit sweep [datafusion-comet]
via GitHub
Re: [I] fair_unified: a JVM consumer freeing its last bytes can fail a parked native acquire with NoSuchElementException [datafusion-comet]
via GitHub
Re: [I] ExecutionMemoryPool errors releasing more memory than allocated [datafusion-comet]
via GitHub
Re: [I] ExecutionMemoryPool errors releasing more memory than allocated [datafusion-comet]
via GitHub
[I] spark.comet.shuffle.jvm.batchSize=0 hangs a task, and an unknown spark.comet.exec.memoryPool fails every task [datafusion-comet]
via GitHub
Re: [I] spark.comet.shuffle.jvm.batchSize=0 hangs a task, and an unknown spark.comet.exec.memoryPool fails every task [datafusion-comet]
via GitHub
[I] With several native plans in one task, the non-zero memory usage warning comes from the wrong plan [datafusion-comet]
via GitHub
Re: [I] With several native plans in one task, the non-zero memory usage warning comes from the wrong plan [datafusion-comet]
via GitHub
[I] Hash-based JVM columnar shuffle reports its output as disk spill instead of bytes written [datafusion-comet]
via GitHub
Re: [I] Hash-based JVM columnar shuffle reports its output as disk spill instead of bytes written [datafusion-comet]
via GitHub
[I] The native memory usage log understates untracked memory while a pool is overcommitted [datafusion-comet]
via GitHub
Re: [I] The native memory usage log understates untracked memory while a pool is overcommitted [datafusion-comet]
via GitHub
[I] A native final aggregate that has spilled can fail the task during its replay [datafusion-comet]
via GitHub
Re: [I] A native final aggregate that has spilled can fail the task during its replay [datafusion-comet]
via GitHub
[I] CometTaskMemoryManager logs a warning and a memory dump every time a native reservation is refused [datafusion-comet]
via GitHub
Re: [I] CometTaskMemoryManager logs a warning and a memory dump every time a native reservation is refused [datafusion-comet]
via GitHub
[I] Native window operators reserve no memory for the batches they buffer [datafusion-comet]
via GitHub
[I] Grouped integer SUM doesn't count its per-group state in the aggregate's memory reservation [datafusion-comet]
via GitHub
[PR] fix(core): look inside subqueries when deciding a plan only reads information_schema [datafusion-ballista]
via GitHub
Re: [PR] fix(core): look inside subqueries when deciding a plan only reads information_schema [datafusion-ballista]
via GitHub
Re: [PR] fix(core): look inside subqueries when deciding a plan only reads information_schema [datafusion-ballista]
via GitHub
Re: [PR] fix(core): refuse queries that mix information_schema with other tables [datafusion-ballista]
via GitHub
Re: [PR] fix(core): refuse queries that mix information_schema with other tables [datafusion-ballista]
via GitHub
[PR] fix(scheduler): fail jobs whose tasks cannot be prepared [datafusion-ballista]
via GitHub
Re: [PR] fix(scheduler): fail jobs whose tasks cannot be prepared [datafusion-ballista]
via GitHub
Re: [PR] fix(scheduler): fail jobs whose tasks cannot be prepared [datafusion-ballista]
via GitHub
[PR] fix(scheduler): declare distributed EXPLAIN output columns not-null [datafusion-ballista]
via GitHub
Re: [PR] fix(scheduler): declare distributed EXPLAIN output columns not-null [datafusion-ballista]
via GitHub
Re: [PR] fix(scheduler): declare distributed EXPLAIN output columns not-null [datafusion-ballista]
via GitHub
[PR] fix: emit one row from a no-grouping aggregate with no aggregate expressions [datafusion]
via GitHub
Re: [PR] fix: emit one row from a no-grouping aggregate with no aggregate expressions [datafusion]
via GitHub
Re: [PR] fix: emit one row from a no-grouping aggregate with no aggregate expressions [datafusion]
via GitHub
Re: [PR] feat: Arrow Flight SQL frontend for the scheduler + ADBC support [datafusion-ballista]
via GitHub
Re: [PR] feat: Arrow Flight SQL frontend for the scheduler + ADBC support [datafusion-ballista]
via GitHub
Re: [PR] feat: Arrow Flight SQL frontend for the scheduler + ADBC support [datafusion-ballista]
via GitHub
Re: [PR] feat: Arrow Flight SQL frontend for the scheduler + ADBC support [datafusion-ballista]
via GitHub
Re: [PR] feat: Arrow Flight SQL frontend for the scheduler + ADBC support [datafusion-ballista]
via GitHub
Re: [PR] feat: Arrow Flight SQL frontend for the scheduler + ADBC support [datafusion-ballista]
via GitHub
[I] Instantiate an `AggregateExec` with no group columns + no aggregates triggers `must either specify a row count or at least one column` from RecordBatch [datafusion]
via GitHub
Re: [PR] ci: run ruff over the whole repository [datafusion-ballista]
via GitHub
Re: [PR] fix(scheduler): refund the vcores of running tasks when a job is aborted [datafusion-ballista]
via GitHub
Re: [I] Cancelling a running job leaks its executor vcores and wedges the scheduler under PushStaged [datafusion-ballista]
via GitHub
Re: [I] arrow-ipc-optimizations in ballista-executor does not forward to ballista-core [datafusion-ballista]
via GitHub
Re: [PR] fix(executor): forward arrow-ipc-optimizations to ballista-core [datafusion-ballista]
via GitHub
Re: [I] Decorrelate subqueries whose grouping sets leave out the correlated column [datafusion]
via GitHub
[PR] fix: decorrelate grouping sets that leave out the correlated column when safe [datafusion]
via GitHub
Re: [PR] fix: decorrelate grouping sets that leave out the correlated column when safe [datafusion]
via GitHub
Re: [PR] fix: decorrelate grouping sets that leave out the correlated column when safe [datafusion]
via GitHub
Re: [PR] ci: share one Linux native build across CI workflows [datafusion-comet]
via GitHub
Re: [PR] ci: share one Linux native build across CI workflows [datafusion-comet]
via GitHub
[PR] feat: make the plan nodes inspectable and rebuildable [datafusion-iceberg]
via GitHub
Re: [PR] feat: make the plan nodes inspectable and rebuildable [datafusion-iceberg]
via GitHub
Re: [PR] feat: make the plan nodes inspectable and rebuildable [datafusion-iceberg]
via GitHub
Re: [PR] feat: make the plan nodes inspectable and rebuildable [datafusion-iceberg]
via GitHub
Re: [PR] feat: make the plan nodes inspectable and rebuildable [datafusion-iceberg]
via GitHub
Re: [PR] feat: make the plan nodes inspectable and rebuildable [datafusion-iceberg]
via GitHub
Re: [PR] feat: make the plan nodes inspectable and rebuildable [datafusion-iceberg]
via GitHub
Re: [PR] feat: make the plan nodes inspectable and rebuildable [datafusion-iceberg]
via GitHub
[I] arrays_zip with two same-named inputs fails with "ArrowArray struct has 2 children (expected 1)" [datafusion-comet]
via GitHub
Re: [I] arrays_zip with two same-named inputs fails with "ArrowArray struct has 2 children (expected 1)" [datafusion-comet]
via GitHub
Re: [I] arrays_zip with two same-named inputs fails with "ArrowArray struct has 2 children (expected 1)" [datafusion-comet]
via GitHub
Re: [I] arrays_zip with two same-named inputs fails with "ArrowArray struct has 2 children (expected 1)" [datafusion-comet]
via GitHub
Re: [I] arrays_zip with two same-named inputs fails with "ArrowArray struct has 2 children (expected 1)" [datafusion-comet]
via GitHub
Re: [PR] fix: avoid invalid date_bin ordering after overflow [datafusion]
via GitHub
Re: [PR] fix: avoid invalid date_bin ordering after overflow [datafusion]
via GitHub
Re: [PR] fix: avoid invalid date_bin ordering after overflow [datafusion]
via GitHub
Re: [PR] fix: avoid invalid date_bin ordering after overflow [datafusion]
via GitHub
Earlier messages
Later messages