github
Thread
Date
Earlier messages
Later messages
Messages by Thread
Re: [PR] bench: parquet scan with a table schema narrower than a nested column [datafusion]
via GitHub
Re: [PR] bench: parquet scan with a table schema narrower than a nested column [datafusion]
via GitHub
Re: [PR] fix: avoid panic in array_position start_from near i64::MIN [datafusion]
via GitHub
Re: [I] Add `any_value` aggregate function [datafusion]
via GitHub
[PR] perf(coalesce): make BatchCoalescer bypass threshold configurable (SIGMOD 2025 Binary Compaction) [datafusion]
via GitHub
Re: [PR] perf(coalesce): make BatchCoalescer bypass threshold configurable (SIGMOD 2025 Binary Compaction) [datafusion]
via GitHub
Re: [PR] perf(coalesce): make BatchCoalescer bypass threshold configurable (SIGMOD 2025 Binary Compaction) [datafusion]
via GitHub
Re: [PR] perf(coalesce): make BatchCoalescer bypass threshold configurable (SIGMOD 2025 Binary Compaction) [datafusion]
via GitHub
Re: [PR] perf(coalesce): make BatchCoalescer bypass threshold configurable (SIGMOD 2025 Binary Compaction) [datafusion]
via GitHub
Re: [PR] experiment perf(coalesce): make BatchCoalescer bypass threshold configurable (SIGMOD 2025 Binary Compaction) [datafusion]
via GitHub
Re: [PR] experiment perf(coalesce): make BatchCoalescer bypass threshold configurable (SIGMOD 2025 Binary Compaction) [datafusion]
via GitHub
Re: [PR] experiment perf(coalesce): make BatchCoalescer bypass threshold configurable (SIGMOD 2025 Binary Compaction) [datafusion]
via GitHub
Re: [PR] experiment perf(coalesce): make BatchCoalescer bypass threshold configurable (SIGMOD 2025 Binary Compaction) [datafusion]
via GitHub
Re: [PR] experiment perf(coalesce): make BatchCoalescer bypass threshold configurable (SIGMOD 2025 Binary Compaction) [datafusion]
via GitHub
Re: [PR] experiment perf(coalesce): make BatchCoalescer bypass threshold configurable (SIGMOD 2025 Binary Compaction) [datafusion]
via GitHub
[PR] [branch-54] chore: Update version 54.1.0, add changelog [datafusion]
via GitHub
Re: [PR] [branch-54] chore: Update version 54.1.0, add changelog [datafusion]
via GitHub
[PR] Optimize Spark hex null handling [datafusion]
via GitHub
Re: [PR] Optimize Spark hex null handling [datafusion]
via GitHub
[PR] feat(agg): TopK aggregation for count(*)/count(col) DESC/ASC LIMIT K [datafusion]
via GitHub
Re: [PR] feat(agg): TopK aggregation for count(*)/count(col) DESC/ASC LIMIT K [datafusion]
via GitHub
Re: [PR] feat(agg): TopK aggregation for count(*)/count(col) DESC/ASC LIMIT K [datafusion]
via GitHub
Re: [PR] feat(agg): TopK aggregation for count(*)/count(col) DESC/ASC LIMIT K [datafusion]
via GitHub
Re: [PR] feat(agg): TopK aggregation for count(*)/count(col) DESC/ASC LIMIT K [datafusion]
via GitHub
Re: [PR] feat(agg): TopK aggregation for count(*)/count(col) DESC/ASC LIMIT K [datafusion]
via GitHub
Re: [PR] feat(agg): TopK aggregation for count(*)/count(col) DESC/ASC LIMIT K [datafusion]
via GitHub
Re: [PR] feat(agg): TopK aggregation for count(*)/count(col) DESC/ASC LIMIT K [datafusion]
via GitHub
Re: [PR] feat(agg): TopK aggregation for count(*)/count(col) DESC/ASC LIMIT K [datafusion]
via GitHub
Re: [PR] experiment feat(agg): TopK aggregation for count(*)/count(col) DESC/ASC LIMIT K [datafusion]
via GitHub
Re: [PR] experiment feat(agg): TopK aggregation for count(*)/count(col) DESC/ASC LIMIT K [datafusion]
via GitHub
Re: [PR] experiment feat(agg): TopK aggregation for count(*)/count(col) DESC/ASC LIMIT K [datafusion]
via GitHub
Re: [PR] experiment feat(agg): TopK aggregation for count(*)/count(col) DESC/ASC LIMIT K [datafusion]
via GitHub
Re: [PR] experiment feat(agg): TopK aggregation for count(*)/count(col) DESC/ASC LIMIT K [datafusion]
via GitHub
Re: [PR] experiment feat(agg): TopK aggregation for count(*)/count(col) DESC/ASC LIMIT K [datafusion]
via GitHub
Re: [PR] experiment feat(agg): TopK aggregation for count(*)/count(col) DESC/ASC LIMIT K [datafusion]
via GitHub
Re: [PR] experiment feat(agg): TopK aggregation for count(*)/count(col) DESC/ASC LIMIT K [datafusion]
via GitHub
[PR] Add schema-aware optimizer child rewrites [datafusion]
via GitHub
[PR] Document decimal AVG wrapping arithmetic [datafusion]
via GitHub
Re: [I] DataFusion drops grouped MIN/MAX rows with NULL sort keys under ORDER BY + LIMIT [datafusion]
via GitHub
[PR] fix: TopK aggregation drops groups whose MIN/MAX value is NULL [datafusion]
via GitHub
[PR] Add DataSource/FileSource proto hooks and FileScanConfig serde [datafusion]
via GitHub
Re: [PR] Add DataSource/FileSource proto hooks and FileScanConfig serde [datafusion]
via GitHub
Re: [PR] Add DataSource/FileSource proto hooks and FileScanConfig serde [datafusion]
via GitHub
Re: [PR] Add any_value aggregate function [datafusion]
via GitHub
Re: [PR] Add any_value aggregate function [datafusion]
via GitHub
Re: [PR] Add any_value aggregate function [datafusion]
via GitHub
Re: [PR] Add any_value aggregate function [datafusion]
via GitHub
Re: [PR] Support lower and upper scalar udf on dict arrays [datafusion]
via GitHub
Re: [PR] Support lower and upper scalar udf on dict arrays [datafusion]
via GitHub
[PR] test: promote `try_to_date`/`try_to_timestamp` SQL tests to native coverage [datafusion-comet]
via GitHub
Re: [PR] fix(spark-expr): handle array length mismatch in datediff for dictionary-backed timestamps [datafusion-comet]
via GitHub
Re: [PR] CI: Add workflow to verify release candidate on multiple systems [datafusion-comet]
via GitHub
Re: [PR] fix: multi-insert with native writer in Spark 4.x (#3430) [datafusion-comet]
via GitHub
Re: [PR] Query graph for Join reordering [datafusion]
via GitHub
Re: [PR] feat: Add `FFI_QueryPlanner` to support foreign query planners across shared-library boundaries [datafusion]
via GitHub
Re: [PR] feat: Add `FFI_QueryPlanner` to support foreign query planners across shared-library boundaries [datafusion]
via GitHub
Re: [PR] [Experiment] Adaptive filter pushdown [datafusion]
via GitHub
[PR] perf: vectorize `spark_unscaled_value` (9x faster) [datafusion-comet]
via GitHub
[PR] fix: `get_json_object` returns first value for duplicate keys to match Spark [datafusion-comet]
via GitHub
Re: [I] Support Substrait exchange output for range repartitioning [datafusion]
via GitHub
[PR] feat: add BALLISTA_PROTOCOL_VERSION + k8s health probes [datafusion-ballista]
via GitHub
Re: [PR] feat: add BALLISTA_PROTOCOL_VERSION + k8s health probes [datafusion-ballista]
via GitHub
Re: [PR] feat: add BALLISTA_PROTOCOL_VERSION + k8s health probes [datafusion-ballista]
via GitHub
Re: [PR] feat: add BALLISTA_PROTOCOL_VERSION + k8s health probes [datafusion-ballista]
via GitHub
Re: [PR] feat: add BALLISTA_PROTOCOL_VERSION + k8s health probes [datafusion-ballista]
via GitHub
Re: [PR] feat: add BALLISTA_PROTOCOL_VERSION + k8s health probes [datafusion-ballista]
via GitHub
Re: [PR] feat: add BALLISTA_PROTOCOL_VERSION + k8s health probes [datafusion-ballista]
via GitHub
[PR] feat(optimizer): coalesce peer first_value / last_value into a single struct aggregate [datafusion]
via GitHub
Re: [PR] feat(optimizer): coalesce peer first_value / last_value into a single struct aggregate [datafusion]
via GitHub
[PR] feat(optimizer): coalesce peer first_value / last_value into a single struct aggregate [datafusion]
via GitHub
[PR] Add visitors for ORDER BY and GROUP BY [datafusion-sqlparser-rs]
via GitHub
Re: [PR] Align metadata propagation through Physical and Logical casts [datafusion]
via GitHub
[PR] chore(deps): bump the codeql-actions group with 2 updates [datafusion-comet]
via GitHub
Re: [PR] chore(deps): bump the codeql-actions group with 2 updates [datafusion-comet]
via GitHub
Re: [I] Track Spark 4.2 test failures [datafusion-comet]
via GitHub
Re: [I] Track Spark 4.2 test failures [datafusion-comet]
via GitHub
[I] Spark 4.2: native Iceberg REST catalog scan test fails under Comet [datafusion-comet]
via GitHub
[PR] [WIP] allow range to satisfy key distribution generally [datafusion]
via GitHub
[I] Spark 4.2: BloomFilter tests fail under Comet [datafusion-comet]
via GitHub
[I] Spark 4.2: ANSI arithmetic overflow tests fail under Comet [datafusion-comet]
via GitHub
[I] Spark 4.2: SQL Last Attempt Metric (SLAM) not propagated through Comet operators [datafusion-comet]
via GitHub
[I] Spark 4.2: UnionCodegenSuite partitioning-aware test inspects UnionExec that Comet replaces [datafusion-comet]
via GitHub
[I] Spark 4.2: segment-tree window metrics unavailable under CometWindowExec [datafusion-comet]
via GitHub
[I] Spark 4.2: Comet native collect_set does not normalize NaN / -0.0 (SPARK-57298) [datafusion-comet]
via GitHub
[PR] chore: codeql dependabot fix [datafusion-comet]
via GitHub
Re: [PR] chore: codeql dependabot fix [datafusion-comet]
via GitHub
[PR] Add Comet 1.0.0 release announcement blog post [datafusion-site]
via GitHub
Re: [PR] chore: Test update object store to 0.14.0 [datafusion]
via GitHub
Re: [PR] chore: Test update object store to 0.14.0 [datafusion]
via GitHub
Re: [PR] chore: Test update object store to 0.14.0 [datafusion]
via GitHub
[PR] chore: add 0.17.1 changelog [datafusion-comet]
via GitHub
Re: [PR] docs: add 0.17.1 changelog [datafusion-comet]
via GitHub
[PR] ci: gate expensive workflows on lint by consolidating into rust.yml [datafusion-ballista]
via GitHub
Re: [PR] ci: gate expensive workflows on lint by consolidating into rust.yml [datafusion-ballista]
via GitHub
Re: [I] Create a pipeline/morsel scheduler [datafusion]
via GitHub
[PR] chore: Enable `unused_async` lint [datafusion]
via GitHub
Re: [PR] chore: Enable `unused_async` lint [datafusion]
via GitHub
Re: [PR] chore: Enable `unused_async` lint [datafusion]
via GitHub
[I] Expand Session trait to parity with SessionState [datafusion]
via GitHub
Re: [I] Expand Session trait to parity with SessionState [datafusion]
via GitHub
Re: [I] Expand Session trait to parity with SessionState [datafusion]
via GitHub
Re: [PR] fix: Capture global ORDER BY requirement under ScalarSubqueryExec root [datafusion]
via GitHub
Re: [PR] fix: Capture global ORDER BY requirement under ScalarSubqueryExec root [datafusion]
via GitHub
[PR] fix: Capture global ORDER BY requirement under ScalarSubqueryExec root [datafusion]
via GitHub
Re: [PR] fix: Capture global ORDER BY requirement under ScalarSubqueryExec root [datafusion]
via GitHub
Re: [PR] fix: Capture global ORDER BY requirement under ScalarSubqueryExec root [datafusion]
via GitHub
Re: [PR] fix: Capture global ORDER BY requirement under ScalarSubqueryExec root [datafusion]
via GitHub
Re: [I] Move static support decisions from serde convert into getSupportLevel [datafusion-comet]
via GitHub
Re: [I] [DISCUSSION] Future of Dynamic Filters Sync [datafusion]
via GitHub
Re: [I] [DISCUSSION] Future of Dynamic Filters Sync [datafusion]
via GitHub
Re: [I] [DISCUSSION] Future of Dynamic Filters Sync [datafusion]
via GitHub
Re: [I] [DISCUSSION] Future of Dynamic Filters Sync [datafusion]
via GitHub
Re: [I] [DISCUSSION] Future of Dynamic Filters Sync [datafusion]
via GitHub
Re: [I] [DISCUSSION] Future of Dynamic Filters Sync [datafusion]
via GitHub
Re: [PR] feat(spark): port Spark-compatible Parquet schema adapter from Comet [datafusion]
via GitHub
Re: [PR] feat(spark): port Spark-compatible Parquet schema adapter from Comet [datafusion]
via GitHub
[PR] fix: preserve precision in log() with mixed-width float arguments [datafusion]
via GitHub
Re: [PR] fix: preserve precision in log() with mixed-width float arguments [datafusion]
via GitHub
Re: [PR] fix: preserve precision in log() with mixed-width float arguments [datafusion]
via GitHub
[I] Extract common regex compilation cache for reuse across regexp functions [datafusion]
via GitHub
[PR] chore: bump spark-4.2 profile to the released 4.2.0 [datafusion-comet]
via GitHub
Re: [PR] chore: bump spark-4.2 profile to the released 4.2.0 [datafusion-comet]
via GitHub
Re: [PR] chore: bump spark-4.2 profile to the released 4.2.0 [datafusion-comet]
via GitHub
Re: [I] Implement TimeType support: Infrastructure - shuffle [datafusion-comet]
via GitHub
[PR] feat: make AQE respect broadcast_join_threshold_bytes [datafusion-ballista]
via GitHub
Re: [PR] feat: make broadcast_join_threshold_bytes/rows authoritative under AQE and the static planner [datafusion-ballista]
via GitHub
Re: [PR] feat: make broadcast_join_threshold_bytes/rows authoritative under AQE and the static planner [datafusion-ballista]
via GitHub
Re: [PR] feat: make broadcast_join_threshold_bytes/rows authoritative under AQE and the static planner [datafusion-ballista]
via GitHub
[I] `ballista.optimizer.broadcast_join_threshold_bytes` has no effect when adaptive query planning is enabled [datafusion-ballista]
via GitHub
Re: [I] `ballista.optimizer.broadcast_join_threshold_bytes` has no effect when adaptive query planning is enabled [datafusion-ballista]
via GitHub
[I] Further optimize Spark hex byte encoding: reuse input NullBuffer and special-case the no-nulls path [datafusion]
via GitHub
Re: [I] Further optimize Spark hex byte encoding: reuse input NullBuffer and special-case the no-nulls path [datafusion]
via GitHub
[PR] perf: bulk-append contiguous buffered runs in sort merge join [datafusion]
via GitHub
Re: [PR] perf: bulk-append contiguous buffered runs in sort merge join [datafusion]
via GitHub
Re: [PR] perf: bulk-append contiguous buffered runs in sort merge join [datafusion]
via GitHub
Re: [PR] perf: bulk-append contiguous buffered runs in sort merge join [datafusion]
via GitHub
Re: [PR] perf: bulk-append contiguous buffered runs in sort merge join [datafusion]
via GitHub
Re: [PR] perf: bulk-append contiguous buffered runs in sort merge join [datafusion]
via GitHub
Re: [PR] perf: bulk-append contiguous buffered runs in sort merge join [datafusion]
via GitHub
Re: [PR] perf: bulk-append contiguous buffered runs in sort merge join [datafusion]
via GitHub
Re: [PR] perf: bulk-append contiguous buffered runs in sort merge join [datafusion]
via GitHub
Re: [PR] perf: bulk-append contiguous buffered runs in sort merge join [datafusion]
via GitHub
Re: [PR] perf: bulk-append contiguous buffered runs in sort merge join [datafusion]
via GitHub
Re: [PR] perf: bulk-append contiguous buffered runs in sort merge join [datafusion]
via GitHub
Re: [PR] perf: bulk-append contiguous buffered runs in sort merge join [datafusion]
via GitHub
Re: [PR] perf: bulk-append contiguous buffered runs in sort merge join [datafusion]
via GitHub
Re: [PR] perf: bulk-append contiguous buffered runs in sort merge join [datafusion]
via GitHub
Re: [PR] perf: bulk-append contiguous buffered runs in sort merge join [datafusion]
via GitHub
Re: [PR] perf: bulk-append contiguous buffered runs in sort merge join [datafusion]
via GitHub
Re: [PR] perf: bulk-append contiguous buffered runs in sort merge join [datafusion]
via GitHub
Re: [PR] perf: bulk-append contiguous buffered runs in sort merge join [datafusion]
via GitHub
Re: [PR] perf: bulk-append contiguous buffered runs in sort merge join [datafusion]
via GitHub
Re: [PR] perf: bulk-append contiguous buffered runs in sort merge join [datafusion]
via GitHub
Re: [PR] perf: bulk-append contiguous buffered runs in sort merge join [datafusion]
via GitHub
Re: [PR] perf: bulk-append contiguous buffered runs in sort merge join [datafusion]
via GitHub
Re: [PR] feat: re-enable COUNT for mixed Spark partial / Comet final aggregates [datafusion-comet]
via GitHub
Re: [PR] feat: add comet_version() SQL function [datafusion-comet]
via GitHub
Re: [PR] feat: add comet_version() SQL function [datafusion-comet]
via GitHub
Re: [PR] feat: custom Rust UDFs via arrow-ffi (alternative to #4283) [experimental] [datafusion-comet]
via GitHub
Re: [PR] feat: custom Rust UDFs via arrow-ffi (alternative to #4283) [experimental] [datafusion-comet]
via GitHub
Re: [PR] feat: support Spark 4.1 TIME type and expressions via codegen dispatch [datafusion-comet]
via GitHub
[PR] feat: removed all instances of deprecated virtualtable values field [datafusion]
via GitHub
Re: [PR] feat: removed all instances of deprecated virtualtable values field [datafusion]
via GitHub
Re: [I] [Feature] support size() for MapType inputs [datafusion-comet]
via GitHub
Re: [PR] feat: support size() for MapType input [datafusion-comet]
via GitHub
Re: [I] Unable to use Ballista at scale (e.g. TPC-H @ 1TB) [datafusion-ballista]
via GitHub
Re: [I] Some benchmark queries fail or produce incorrect results [datafusion-ballista]
via GitHub
Re: [I] Some benchmark queries fail or produce incorrect results [datafusion-ballista]
via GitHub
Re: [I] Document how to run TPC-H benchmarks in Kubernetes [datafusion-ballista]
via GitHub
Re: [I] Document how to run TPC-H benchmarks in Kubernetes [datafusion-ballista]
via GitHub
Re: [I] Improve benchmark performance [datafusion-ballista]
via GitHub
Re: [I] Improve benchmark performance [datafusion-ballista]
via GitHub
[I] cast string to boolean: trim ISO control bytes to match Spark's UTF8String.trimAll [datafusion-comet]
via GitHub
[I] Thread `PhysicalOptimizerContext` through `join_selection` stats helpers (CBO extension pattern) [datafusion]
via GitHub
Re: [I] Thread `PhysicalOptimizerContext` through `join_selection` stats helpers (CBO extension pattern) [datafusion]
via GitHub
Re: [PR] fix: avoid full listings for cached pruned partitions [datafusion]
via GitHub
[PR] fix(datasource): avoid over-conservative transformation of num_rows statistics in file scan config [datafusion]
via GitHub
Re: [PR] fix(datasource): avoid over-conservative transformation of num_rows statistics in file scan config [datafusion]
via GitHub
[PR] fix: handle NULL string and format arguments in to_date, to_timestamp, and to_unixtime [datafusion]
via GitHub
Re: [PR] fix: handle NULL string and format arguments in to_date, to_timestamp, and to_unixtime [datafusion]
via GitHub
Re: [PR] fix: handle NULL string and format arguments in to_date, to_timestamp, and to_unixtime [datafusion]
via GitHub
Re: [PR] fix: handle NULL string and format arguments in to_date, to_timestamp, and to_unixtime [datafusion]
via GitHub
Re: [PR] fix: handle NULL string and format arguments in to_date, to_timestamp, and to_unixtime [datafusion]
via GitHub
Re: [I] PostgreSQL compatibility: `time + interval` should wrap within the 24-hour time domain [datafusion]
via GitHub
Re: [I] PostgreSQL compatibility: `time - interval` should wrap within the 24-hour time domain [datafusion]
via GitHub
[I] Centralize aggregate-scope expression rendering in the SQL unparser [datafusion]
via GitHub
[I] Centralize higher-order list lambda evaluation helpers [datafusion]
via GitHub
Re: [I] Centralize higher-order list lambda evaluation helpers [datafusion]
via GitHub
Re: [I] Centralize higher-order list lambda evaluation helpers [datafusion]
via GitHub
[I] Refactor: Add an optimizer-local helper for schema-aware child rewrites [datafusion]
via GitHub
[I] Document decimal AVG wrapping arithmetic behind explicit helpers [datafusion]
via GitHub
Re: [I] Bug in lambda variable resolution [datafusion]
via GitHub
Re: [PR] Feat: add dictionaries as a supported group column type [datafusion]
via GitHub
Re: [PR] Feat: add dictionaries as a supported group column type [datafusion]
via GitHub
Re: [PR] Feat: add dictionaries as a supported group column type [datafusion]
via GitHub
Re: [PR] Feat: add dictionaries as a supported group column type [datafusion]
via GitHub
Re: [PR] Feat: add dictionaries as a supported group column type [datafusion]
via GitHub
Re: [PR] Feat: add dictionaries as a supported group column type [datafusion]
via GitHub
Re: [PR] Feat: add dictionaries as a supported group column type [datafusion]
via GitHub
Re: [PR] Feat: add dictionaries as a supported group column type [datafusion]
via GitHub
Earlier messages
Later messages