github
Thread
Date
Earlier messages
Later messages
Messages by Date
2026/08/18
Re: [PR] docs: add Streamling to known users [datafusion]
via GitHub
2026/08/18
Re: [I] Replace small hand-rolled element loops with arrow kernels and arity helpers [datafusion-comet]
via GitHub
2026/08/18
[I] FuzzDataGenerator silently drops nulls for Boolean/Byte/Short/Integer columns [datafusion-comet]
via GitHub
2026/08/18
Re: [I] FuzzDataGenerator silently drops nulls for Boolean/Byte/Short/Integer columns [datafusion-comet]
via GitHub
2026/08/18
Re: [PR] refactor: replace remaining hand-rolled loops with Arrow kernels [datafusion-comet]
via GitHub
2026/08/18
Re: [PR] (benchmaxx) perf: choose hash join build side at execution time [datafusion]
via GitHub
2026/08/18
Re: [PR] fix: adapt input batches with stricter nested nullability to planned schema in aggregation [datafusion]
via GitHub
2026/08/18
Re: [PR] (benchmaxx) perf: choose hash join build side at execution time [datafusion]
via GitHub
2026/08/18
Re: [PR] (benchmaxx) perf: choose hash join build side at execution time [datafusion]
via GitHub
2026/08/18
Re: [PR] perf: reuse the projected equivalence group when only child orderings change [datafusion]
via GitHub
2026/08/18
Re: [PR] Cache reusable Parquet pruning setup across files with the same schema [datafusion]
via GitHub
2026/08/18
Re: [PR] chore(deps): bump aws-config from 1.10.0 to 1.10.1 [datafusion-ballista]
via GitHub
2026/08/18
Re: [PR] Cache reusable Parquet pruning setup across files with the same schema [datafusion]
via GitHub
2026/08/18
Re: [PR] feat(spark): add `to_binary` and `try_to_binary` [datafusion]
via GitHub
2026/08/18
Re: [PR] perf: raise perfect hash join small build threshold to 256K [datafusion]
via GitHub
2026/08/18
Re: [PR] Cache reusable Parquet pruning setup across files with the same schema [datafusion]
via GitHub
2026/08/18
Re: [PR] fix: log(0.0::float8) should error, not return -inf [datafusion]
via GitHub
2026/08/18
Re: [PR] fix: adapt input batches with stricter nested nullability to planned schema in aggregation [datafusion]
via GitHub
2026/08/18
Re: [PR] perf: raise perfect hash join small build threshold to 256K [datafusion]
via GitHub
2026/08/18
Re: [PR] feat(parquet): clip nested wrappers during schema pruning [datafusion]
via GitHub
2026/08/18
Re: [PR] perf: raise perfect hash join small build threshold to 256K [datafusion]
via GitHub
2026/08/18
Re: [PR] perf: raise perfect hash join small build threshold to 256K [datafusion]
via GitHub
2026/08/18
Re: [PR] perf: raise perfect hash join small build threshold to 256K [datafusion]
via GitHub
2026/08/18
[PR] test(scheduler): pin the expected static and distributed plans [datafusion-ballista]
via GitHub
2026/08/18
Re: [PR] perf: raise perfect hash join small build threshold to 256K [datafusion]
via GitHub
2026/08/18
Re: [PR] perf: make CometShuffleBenchmark completable and fast [datafusion-comet]
via GitHub
2026/08/18
Re: [PR] fix(tui): use the shared API wire types instead of local copies [datafusion-ballista]
via GitHub
2026/08/18
Re: [I] Create a new roadmap for remainder of 2026 [datafusion-ballista]
via GitHub
2026/08/18
Re: [PR] perf: raise perfect hash join small build threshold to 256K [datafusion]
via GitHub
2026/08/18
Re: [PR] perf: raise perfect hash join small build threshold to 256K [datafusion]
via GitHub
2026/08/18
Re: [PR] perf: raise perfect hash join small build threshold to 256K [datafusion]
via GitHub
2026/08/18
Re: [PR] perf: raise perfect hash join small build threshold to 256K [datafusion]
via GitHub
2026/08/18
Re: [PR] chore(ci): bump the github-codeql group across 1 directory with 2 updates [datafusion-ballista]
via GitHub
2026/08/18
Re: [PR] chore(deps): bump futures from 0.3.33 to 0.3.34 [datafusion-ballista]
via GitHub
2026/08/18
Re: [PR] chore(deps): bump async-trait from 0.1.91 to 0.1.92 [datafusion-ballista]
via GitHub
2026/08/18
Re: [PR] chore(ci): bump taiki-e/install-action from 2.85.11 to 2.86.1 [datafusion-ballista]
via GitHub
2026/08/18
Re: [PR] chore(deps): bump uuid from 1.24.0 to 1.24.1 [datafusion-ballista]
via GitHub
2026/08/17
Re: [PR] minor(test): strengthen bitwise sort-merge join spill coverage [datafusion]
via GitHub
2026/08/17
Re: [PR] fix: Support RANGE window frames over binary ORDER BY keys [datafusion]
via GitHub
2026/08/17
Re: [PR] perf: reuse the projected equivalence group when only child orderings change [datafusion]
via GitHub
2026/08/17
[PR] perf(scheduler): only give a broadcast build side its own stage when the probe is partitioned [datafusion-ballista]
via GitHub
2026/08/17
Re: [PR] perf: raise perfect hash join small build threshold to 256K [datafusion]
via GitHub
2026/08/17
Re: [PR] perf: reuse the projected equivalence group when only child orderings change [datafusion]
via GitHub
2026/08/17
Re: [PR] perf: raise perfect hash join small build threshold to 256K [datafusion]
via GitHub
2026/08/17
Re: [PR] perf: raise perfect hash join small build threshold to 256K [datafusion]
via GitHub
2026/08/17
Re: [PR] perf: raise perfect hash join small build threshold to 256K [datafusion]
via GitHub
2026/08/17
Re: [PR] perf: raise perfect hash join small build threshold to 256K [datafusion]
via GitHub
2026/08/17
Re: [PR] perf: raise perfect hash join small build threshold to 256K [datafusion]
via GitHub
2026/08/17
[PR] perf: raise perfect hash join small build threshold to 256K [datafusion]
via GitHub
2026/08/17
Re: [PR] perf: don't manufacture an identity projection in ParquetSource [datafusion]
via GitHub
2026/08/17
Re: [PR] perf: reuse the projected equivalence group when only child orderings change [datafusion]
via GitHub
2026/08/17
Re: [PR] fix: evaluate grouped aggregate arguments after FILTER [datafusion]
via GitHub
2026/08/17
Re: [PR] (benchmaxx) perf: choose hash join build side at execution time [datafusion]
via GitHub
2026/08/17
Re: [PR] (benchmaxx) perf: choose hash join build side at execution time [datafusion]
via GitHub
2026/08/17
Re: [I] Release DataFusion `55.0.0` (Jul / Aug 2026) [datafusion]
via GitHub
2026/08/17
[I] Correctness issue with unparsing UNION & UNION ALL [datafusion]
via GitHub
2026/08/17
Re: [PR] Snowflake: Add support for ->> (pipe) operator for chaining SQL stmts [datafusion-sqlparser-rs]
via GitHub
2026/08/17
Re: [PR] (benchmaxx) perf: choose hash join build side at execution time [datafusion]
via GitHub
2026/08/17
Re: [PR] (benchmaxx) perf: choose hash join build side at execution time [datafusion]
via GitHub
2026/08/17
Re: [PR] (benchmaxx) perf: choose hash join build side at execution time [datafusion]
via GitHub
2026/08/17
Re: [PR] (benchmaxx) perf: choose hash join build side at execution time [datafusion]
via GitHub
2026/08/17
[PR] perf: reuse the projected equivalence group when only child orderings change [datafusion]
via GitHub
2026/08/17
[PR] fix: evaluate grouped aggregate arguments after FILTER [datafusion]
via GitHub
2026/08/17
Re: [PR] (benchmaxx) perf: choose hash join build side at execution time [datafusion]
via GitHub
2026/08/17
Re: [I] Grouped aggregate FILTER evaluates arguments for rejected rows [datafusion]
via GitHub
2026/08/17
Re: [PR] (benchmaxx) perf: choose hash join build side at execution time [datafusion]
via GitHub
2026/08/17
Re: [PR] (benchmaxx) perf: choose hash join build side at execution time [datafusion]
via GitHub
2026/08/17
[I] Grouped aggregate FILTER evaluates arguments for rejected rows [datafusion]
via GitHub
2026/08/17
Re: [PR] bench(functions): group the 63 bench targets into 9 [datafusion]
via GitHub
2026/08/17
Re: [PR] (benchmax) perf: bound multi-column join cardinality by the composite key [datafusion]
via GitHub
2026/08/17
Re: [PR] feat: add basic MemTable MERGE INTO support [datafusion]
via GitHub
2026/08/17
Re: [PR] Add reservation-backed memory accounting to LimitedBatchCoalescer [datafusion]
via GitHub
2026/08/17
Re: [PR] perf: choose hash join build side at execution time [datafusion]
via GitHub
2026/08/17
Re: [PR] perf: choose hash join build side at execution time [datafusion]
via GitHub
2026/08/17
[PR] perf: choose hash join build side at execution time [datafusion]
via GitHub
2026/08/17
Re: [PR] feat(parquet): clip nested wrappers during schema pruning [datafusion]
via GitHub
2026/08/17
Re: [I] Unify integer decimal sign and width assembly with numeric formatting [datafusion]
via GitHub
2026/08/17
Re: [I] Unify integer decimal sign and width assembly with numeric formatting [datafusion]
via GitHub
2026/08/17
Re: [PR] chore: drop Spark 3.4 support [datafusion-comet]
via GitHub
2026/08/17
Re: [PR] chore: drop Spark 3.4 support [datafusion-comet]
via GitHub
2026/08/17
Re: [PR] Experimental: support parquet partition write [datafusion-comet]
via GitHub
2026/08/17
Re: [PR] fix(spark): match Spark abs overflow errors [datafusion]
via GitHub
2026/08/17
Re: [PR] fix(spark): match Spark abs overflow errors [datafusion]
via GitHub
2026/08/17
Re: [PR] Fix physical distinctness expression nullability [datafusion]
via GitHub
2026/08/17
Re: [PR] introduce new dictionary benchmarks [datafusion]
via GitHub
2026/08/17
Re: [PR] introduce new dictionary benchmarks [datafusion]
via GitHub
2026/08/17
Re: [PR] fix: format Bytes-category custom metrics with byte units [datafusion]
via GitHub
2026/08/17
Re: [PR] introduce new dictionary benchmarks [datafusion]
via GitHub
2026/08/17
Re: [PR] fix: format Bytes-category custom metrics with byte units [datafusion]
via GitHub
2026/08/17
[PR] chore(deps): bump uuid from 1.24.0 to 1.24.1 [datafusion-ballista]
via GitHub
2026/08/17
Re: [PR] chore(ci): bump taiki-e/install-action from 2.85.11 to 2.85.13 [datafusion-ballista]
via GitHub
2026/08/17
Re: [PR] chore(ci): bump taiki-e/install-action from 2.85.11 to 2.85.13 [datafusion-ballista]
via GitHub
2026/08/17
[PR] chore(ci): bump taiki-e/install-action from 2.85.11 to 2.86.1 [datafusion-ballista]
via GitHub
2026/08/17
Re: [PR] minor(test): strengthen bitwise sort-merge join spill coverage [datafusion]
via GitHub
2026/08/17
Re: [PR] Enable more clippy lints [datafusion]
via GitHub
2026/08/17
Re: [PR] introduce new dictionary benchmarks [datafusion]
via GitHub
2026/08/17
Re: [PR] perf: Extend WindowTopN to `dense_rank` [datafusion]
via GitHub
2026/08/17
Re: [PR] fix(spark): match Spark abs overflow errors [datafusion]
via GitHub
2026/08/17
Re: [PR] perf: Extend WindowTopN to `dense_rank` [datafusion]
via GitHub
2026/08/17
Re: [PR] perf: don't manufacture an identity projection in ParquetSource [datafusion]
via GitHub
2026/08/17
Re: [I] Release DataFusion `55.0.0` (Jul / Aug 2026) [datafusion]
via GitHub
2026/08/17
[PR] perf: don't manufacture an identity projection in ParquetSource [datafusion]
via GitHub
2026/08/17
Re: [PR] docs: add Streamling to known users [datafusion]
via GitHub
2026/08/17
Re: [PR] docs: add Streamling to known users [datafusion]
via GitHub
2026/08/17
Re: [PR] Doris SQL: create table column options [datafusion-sqlparser-rs]
via GitHub
2026/08/17
Re: [PR] Experimental: support parquet partition write [datafusion-comet]
via GitHub
2026/08/17
Re: [PR] Add reservation-backed memory accounting to LimitedBatchCoalescer [datafusion]
via GitHub
2026/08/17
Re: [PR] fix: plumb SchemaProvider table_type through FFI [datafusion]
via GitHub
2026/08/17
Re: [PR] PoC: Blocked state management for hash aggregation [datafusion]
via GitHub
2026/08/17
Re: [PR] fix: RIGHT/FULL/NATURAL JOIN USING does not coalesce the join key (returns NULL for right-only rows) [datafusion]
via GitHub
2026/08/17
Re: [PR] docs: move TopK user defined operator example into extending-operators guide [datafusion]
via GitHub
2026/08/17
Re: [PR] feat: decimal support for percentile_cont and median [datafusion]
via GitHub
2026/08/17
Re: [PR] refactor: CrossJoinStream (simplifying, less state, async generator pattern) [datafusion]
via GitHub
2026/08/17
Re: [PR] feat: Sort-aware Iceberg reads in Comet via a per-partition streaming merge [datafusion-comet]
via GitHub
2026/08/17
Re: [PR] feat: Sort-aware Iceberg reads in Comet via a per-partition streaming merge [datafusion-comet]
via GitHub
2026/08/17
Re: [PR] feat: add native Delta Lake scan contrib module (page/row-group pruning) [datafusion-comet]
via GitHub
2026/08/17
Re: [PR] feat: Sort-aware Iceberg reads in Comet via a per-partition streaming merge [datafusion-comet]
via GitHub
2026/08/17
Re: [PR] fix: expand float fuzz coverage and correct discovered inconsistencies [datafusion]
via GitHub
2026/08/17
Re: [PR] perf: skip building FilterRemapper when there are no parent filters [datafusion]
via GitHub
2026/08/17
Re: [PR] fix: expand float fuzz coverage and correct discovered inconsistencies [datafusion]
via GitHub
2026/08/17
Re: [PR] fix: use total_cmp for grouped float MIN/MAX accumulators (#24432) [datafusion]
via GitHub
2026/08/17
Re: [PR] fix: use to_bits() comparison for float types in eq_array (#24431) [datafusion]
via GitHub
2026/08/17
Re: [PR] fix: use to_bits() comparison for float types in eq_array (#24431) [datafusion]
via GitHub
2026/08/17
Re: [PR] fix: use total_cmp for grouped float MIN/MAX accumulators (#24432) [datafusion]
via GitHub
2026/08/17
Re: [PR] fix: reject Parquet files with duplicate column names instead of silently dropping data [datafusion]
via GitHub
2026/08/17
Re: [PR] fix: reject Parquet files with duplicate column names instead of silently dropping data [datafusion]
via GitHub
2026/08/17
Re: [PR] fix: preserve Catalyst nullability and field IDs in native Parquet writes [datafusion-comet]
via GitHub
2026/08/17
Re: [PR] fix: preserve Catalyst nullability and field IDs in native Parquet writes [datafusion-comet]
via GitHub
2026/08/17
Re: [PR] feat: Sort-aware Iceberg reads in Comet via a per-partition streaming merge [datafusion-comet]
via GitHub
2026/08/17
Re: [PR] feat: build gate + inert wiring for contrib Delta scans [Delta contrib split, part 2] [datafusion-comet]
via GitHub
2026/08/17
Re: [PR] fix: preserve Catalyst nullability and field IDs in native Parquet writes [datafusion-comet]
via GitHub
2026/08/17
Re: [PR] feat: support max_by and min_by aggregate expressions [datafusion-comet]
via GitHub
2026/08/17
Re: [PR] docs: refresh benchmarking.md with SF1000 results after #2315 (AQE default-on) [datafusion-ballista]
via GitHub
2026/08/17
Re: [I] Support streaming aggregates when partitions are unsorted but non‑overlapping [datafusion]
via GitHub
2026/08/17
Re: [I] Support streaming aggregates when partitions are unsorted but non‑overlapping [datafusion]
via GitHub
2026/08/17
Re: [I] Support streaming aggregates when partitions are unsorted but non‑overlapping [datafusion]
via GitHub
2026/08/17
Re: [PR] feat: stream aggregates when group keys are partition-disjoint [datafusion]
via GitHub
2026/08/17
[PR] feat: stream aggregates when group keys are partition-disjoint [datafusion]
via GitHub
2026/08/17
Re: [PR] fix: make CometExplodeExec respect batch size [datafusion-comet]
via GitHub
2026/08/17
Re: [PR] fix: make CometExplodeExec respect batch size [datafusion-comet]
via GitHub
2026/08/17
Re: [PR] fix: make CometExplodeExec respect batch size [datafusion-comet]
via GitHub
2026/08/17
Re: [PR] fix: make CometExplodeExec respect batch size [datafusion-comet]
via GitHub
2026/08/17
[PR] docs: add Streamling to known users [datafusion]
via GitHub
2026/08/17
Re: [PR] feat: build gate + inert wiring for contrib Delta scans [Delta contrib split, part 2] [datafusion-comet]
via GitHub
2026/08/17
Re: [PR] [Tracking] feat(contrib): Native Delta Lake scan via delta-kernel-rs (Iceberg-style contrib) [datafusion-comet]
via GitHub
2026/08/17
Re: [PR] feat: add micro benchmark runner and EC2 guide [datafusion-comet]
via GitHub
2026/08/17
Re: [PR] (benchmax) perf: bound multi-column join cardinality by the composite key [datafusion]
via GitHub
2026/08/17
Re: [PR] (benchmax) perf: bound multi-column join cardinality by the composite key [datafusion]
via GitHub
2026/08/17
Re: [PR] (benchmax) perf: bound multi-column join cardinality by the composite key [datafusion]
via GitHub
2026/08/17
Re: [PR] perf(core): share a broadcast read between the tasks on an executor [datafusion-ballista]
via GitHub
2026/08/17
Re: [PR] perf: bound multi-column join cardinality by the composite key [datafusion]
via GitHub
2026/08/17
Re: [PR] perf: bound multi-column join cardinality by the composite key [datafusion]
via GitHub
2026/08/17
Re: [PR] perf: bound multi-column join cardinality by the composite key [datafusion]
via GitHub
2026/08/17
Re: [PR] perf: bound multi-column join cardinality by the composite key [datafusion]
via GitHub
2026/08/17
[PR] perf: make CometShuffleBenchmark completable and fast [datafusion-comet]
via GitHub
2026/08/17
Re: [I] Materialize Dictionaries in Group Keys [datafusion]
via GitHub
2026/08/17
Re: [I] Materialize Dictionaries in Group Keys [datafusion]
via GitHub
2026/08/17
Re: [I] Boolean-to-decimal cast produces invalid Decimal128 when 10^scale does not fit precision [datafusion-comet]
via GitHub
2026/08/17
Re: [PR] feat: remove native cast from boolean to decimal [datafusion-comet]
via GitHub
2026/08/17
Re: [I] Support streaming aggregates when partitions are unsorted but non‑overlapping [datafusion]
via GitHub
2026/08/17
Re: [I] Support streaming aggregates when partitions are unsorted but non‑overlapping [datafusion]
via GitHub
2026/08/17
[I] Support streaming aggregates when partitions are unsorted but non‑overlapping [datafusion]
via GitHub
2026/08/17
Re: [PR] perf: bound multi-column join cardinality by the composite key [datafusion]
via GitHub
2026/08/17
Re: [PR] perf: cache dictionary arc pointer [datafusion]
via GitHub
2026/08/17
Re: [I] Significant overhead in datafusion_common::utils::memory::get_record_batch_memory_size [datafusion]
via GitHub
2026/08/17
Re: [PR] perf: Reduce record batch memory accounting overhead [datafusion]
via GitHub
2026/08/17
[PR] perf: grow the native shuffle writer's reservation in steps [datafusion-comet]
via GitHub
2026/08/17
Re: [PR] Add atomic SessionContext.with_extensions API [datafusion-python]
via GitHub
2026/08/17
Re: [PR] feat: support `WindowGroupLimitExec` [datafusion-comet]
via GitHub
2026/08/17
Re: [I] Extend native Arrow UDF path to scalar Python UDFs (ArrowEvalPythonExec) [datafusion-comet]
via GitHub
2026/08/17
[I] Extend native Arrow UDF path to scalar Python UDFs (ArrowEvalPythonExec) [datafusion-comet]
via GitHub
2026/08/17
[PR] perf: bound multi-column join cardinality by the composite key [datafusion]
via GitHub
2026/08/17
[PR] perf(core): share a broadcast read between the tasks on an executor [datafusion-ballista]
via GitHub
2026/08/17
[PR] fix: use to_bits() comparison for float types in eq_array (#24431) [datafusion]
via GitHub
2026/08/17
Re: [I] Reduce JNI round-trips in the unified memory pools (batching, hysteresis, cheaper call path) [datafusion-comet]
via GitHub
2026/08/17
Re: [PR] Improve planning speed: Fast path for `union_schema` when all children share a schema [datafusion]
via GitHub
2026/08/17
Re: [I] feat: add generic FallbackGroupColumn so any Arrow type uses GroupValuesColumn (no fallback to GroupValuesRows) [datafusion]
via GitHub
2026/08/17
Re: [I] feat: add generic FallbackGroupColumn so any Arrow type uses GroupValuesColumn (no fallback to GroupValuesRows) [datafusion]
via GitHub
2026/08/17
Re: [PR] perf: Reduce record batch memory accounting overhead [datafusion]
via GitHub
2026/08/17
Re: [PR] Add dfbench statistics command [datafusion]
via GitHub
2026/08/17
[PR] fix: use total_cmp for grouped float MIN/MAX accumulators (#24432) [datafusion]
via GitHub
2026/08/17
Re: [PR] perf(scheduler): read an unfiltered small scan inline instead of staging it [datafusion-ballista]
via GitHub
2026/08/17
Re: [I] [DISCUSSION] PR Review Process / States / "How to find the PRs in the list that need review" [datafusion]
via GitHub
2026/08/17
Re: [I] Consider PR labels to make reviewing easier [datafusion]
via GitHub
2026/08/17
Re: [I] Add AI tooling disclosure text to contributor guide and fields to PR templates [datafusion]
via GitHub
2026/08/17
Re: [PR] docs: document regular expression function arguments [datafusion-python]
via GitHub
2026/08/17
Re: [D] Policy on uptick of LLM PRs from new contributors [datafusion]
via GitHub
2026/08/17
[PR] perf(scheduler): read an unfiltered small scan inline instead of staging it [datafusion-ballista]
via GitHub
2026/08/17
Re: [I] Grouped approx_distinct can overflow Arrow BinaryArray state offsets [datafusion]
via GitHub
2026/08/17
Re: [PR] Apply same sparse routing for range partition dynamic filter as hash partition and restructure `build_filter` [datafusion]
via GitHub
2026/08/17
Re: [PR] docs(dataframe): improve dataframe! and from_columns documentation (#… [datafusion]
via GitHub
2026/08/17
Re: [PR] [NOT READY FOR REVIEW] Add DataFusion 55.0.0 release blog post [datafusion-site]
via GitHub
2026/08/17
Re: [PR] [NOT READY FOR REVIEW] Add DataFusion 55.0.0 release blog post [datafusion-site]
via GitHub
2026/08/17
[PR] docs(dataframe): improve dataframe! and from_columns documentation (#… [datafusion]
via GitHub
2026/08/17
[PR] fix: expand float fuzz coverage and correct discovered inconsistencies [datafusion]
via GitHub
2026/08/17
Re: [I] Allow individual cast conversions to be enabled/disabled [datafusion-comet]
via GitHub
2026/08/17
[I] Grouped floating-point `MIN` and `MAX` return order-dependent results for NaNs and signed zeros [datafusion]
via GitHub
2026/08/17
[I] `ScalarValue::eq_array` is inconsistent with `ScalarValue` equality for floating-point values [datafusion]
via GitHub
2026/08/17
Re: [PR] Honour operator precedence in `IS [NOT] DISTINCT FROM` [datafusion-sqlparser-rs]
via GitHub
2026/08/17
Re: [I] to_time / try_to_time: native parser rejects 'T12' and '12:30:45.' which Spark accepts [datafusion-comet]
via GitHub
Earlier messages
Later messages