github
Thread
Date
Earlier messages
Later messages
Messages by Thread
Re: [PR] chore: Fix Scala code warnings [datafusion-comet]
via GitHub
Re: [PR] chore: Fix Scala code warnings [datafusion-comet]
via GitHub
Re: [PR] chore: Fix Scala code warnings [datafusion-comet]
via GitHub
Re: [PR] chore: Fix Scala code warnings [datafusion-comet]
via GitHub
Re: [PR] chore: Fix Scala code warnings [datafusion-comet]
via GitHub
[I] `array_append` mishandles NULL values with non-empty offset ranges [datafusion]
via GitHub
Re: [I] `array_append` mishandles NULL values with non-empty offset ranges [datafusion]
via GitHub
Re: [I] Move Extended tests to the standard ones? [datafusion]
via GitHub
Re: [I] [DISCUSS] Separate stable API into separate crates [datafusion]
via GitHub
[I] CAST statistics can produce incorrect MIN/MAX results [datafusion]
via GitHub
Re: [I] CAST statistics can produce incorrect MIN/MAX results [datafusion]
via GitHub
Re: [PR] feat: Arrow Flight SQL frontend for the scheduler + ADBC support [datafusion-ballista]
via GitHub
Re: [PR] feat: Add Iceberg Support [datafusion-ballista]
via GitHub
Re: [PR] feat: Add Iceberg Support [datafusion-ballista]
via GitHub
Re: [PR] feat: Add Iceberg Support [datafusion-ballista]
via GitHub
Re: [PR] feat: Pinot-style colocated-join optimizer for hash-bucketed tables [datafusion-ballista]
via GitHub
Re: [PR] feat: Pinot-style colocated-join optimizer for hash-bucketed tables [datafusion-ballista]
via GitHub
[I] Use exact column statistics to prove ordering through integer arithmetic [datafusion]
via GitHub
Re: [I] Use exact column statistics to prove ordering through integer arithmetic [datafusion]
via GitHub
Re: [PR] feat(executor): use default memory pool in executor if no config provided [datafusion-ballista]
via GitHub
Re: [PR] [WIP]feat(aqe): port Spark's OptimizeSkewedJoin [datafusion-ballista]
via GitHub
Re: [PR] [WIP]feat(aqe): port Spark's OptimizeSkewedJoin [datafusion-ballista]
via GitHub
Re: [PR] [WIP] feat(aqe): early-stop on global LIMIT [datafusion-ballista]
via GitHub
Re: [PR] [WIP] feat(aqe): early-stop on global LIMIT [datafusion-ballista]
via GitHub
Re: [PR] test(python): add datafusion-python compatibility tests [datafusion-ballista]
via GitHub
Re: [PR] test(python): add datafusion-python compatibility tests [datafusion-ballista]
via GitHub
Re: [PR] feat(aqe): support executor failure in AdaptiveExecutionGraph [datafusion-ballista]
via GitHub
Re: [PR] feat(aqe): support executor failure in AdaptiveExecutionGraph [datafusion-ballista]
via GitHub
[I] Reuse the cached dictionary value→slot map in vectorized_equal_to instead of rebuilding it per batch [datafusion]
via GitHub
Re: [I] Reuse the cached dictionary value→slot map in vectorized_equal_to instead of rebuilding it per batch [datafusion]
via GitHub
Re: [PR] fix(physical-plan): honor distinct soft limits in SingleHashAggregateStream [datafusion]
via GitHub
Re: [PR] fix(physical-plan): honor distinct soft limits in SingleHashAggregateStream [datafusion]
via GitHub
Re: [PR] fix(physical-plan): honor distinct soft limits in SingleHashAggregateStream [datafusion]
via GitHub
Re: [PR] fix(physical-plan): honor distinct soft limits in SingleHashAggregateStream [datafusion]
via GitHub
Re: [PR] fix(physical-plan): honor distinct soft limits in SingleHashAggregateStream [datafusion]
via GitHub
Re: [PR] fix(physical-plan): honor distinct soft limits in SingleHashAggregateStream [datafusion]
via GitHub
Re: [PR] fix(physical-plan): honor distinct soft limits in SingleHashAggregateStream [datafusion]
via GitHub
Re: [PR] fix(physical-plan): honor distinct soft limits in SingleHashAggregateStream [datafusion]
via GitHub
Re: [PR] fix(physical-plan): honor distinct soft limits in SingleHashAggregateStream [datafusion]
via GitHub
Re: [PR] fix(physical-plan): honor distinct soft limits in SingleHashAggregateStream [datafusion]
via GitHub
Re: [PR] fix(physical-plan): honor distinct soft limits in SingleHashAggregateStream [datafusion]
via GitHub
Re: [PR] fix(physical-plan): honor distinct soft limits in SingleHashAggregateStream [datafusion]
via GitHub
Re: [PR] fix(physical-plan): honor distinct soft limits in SingleHashAggregateStream [datafusion]
via GitHub
Re: [PR] fix(physical-plan): honor distinct soft limits in SingleHashAggregateStream [datafusion]
via GitHub
Re: [PR] fix(physical-plan): honor distinct soft limits in SingleHashAggregateStream [datafusion]
via GitHub
[PR] feat: sort-merge fallback for hash joins under memory pressure [datafusion]
via GitHub
Re: [PR] feat: sort-merge fallback for hash joins under memory pressure [datafusion]
via GitHub
Re: [PR] feat: sort-merge fallback for partitioned hash joins under memory pressure [datafusion]
via GitHub
Re: [PR] feat: sort-merge fallback for partitioned hash joins under memory pressure [datafusion]
via GitHub
Re: [PR] feat: sort-merge fallback for partitioned hash joins under memory pressure [datafusion]
via GitHub
[PR] fix(ci): pull the MinIO test image from quay.io [datafusion-ballista]
via GitHub
Re: [PR] fix(ci): pull the MinIO test image from quay.io [datafusion-ballista]
via GitHub
Re: [PR] fix(ci): pull the MinIO test image from quay.io [datafusion-ballista]
via GitHub
[I] ci: object store tests fail to pull the MinIO docker image [datafusion-ballista]
via GitHub
Re: [I] ci: object store tests fail to pull the MinIO docker image [datafusion-ballista]
via GitHub
[PR] fix(ci): pull the MinIO test image from quay.io [datafusion]
via GitHub
Re: [PR] fix(ci): pull the MinIO test image from quay.io [datafusion]
via GitHub
Re: [PR] fix(ci): pull the MinIO test image from quay.io [datafusion]
via GitHub
Re: [PR] fix(ci): pull the MinIO test image from quay.io [datafusion]
via GitHub
Re: [PR] perf: Optimize prefix-group processing in `PartialSortExec` [datafusion]
via GitHub
Re: [PR] perf: Optimize prefix-group processing in `PartialSortExec` [datafusion]
via GitHub
Re: [PR] perf: Optimize prefix-group processing in `PartialSortExec` [datafusion]
via GitHub
[PR] feat: run length on binary input natively [datafusion-comet]
via GitHub
Re: [PR] feat: run length on binary input natively [datafusion-comet]
via GitHub
Re: [PR] feat: run length on binary input natively [datafusion-comet]
via GitHub
Re: [PR] feat: run length on binary input natively [datafusion-comet]
via GitHub
Re: [PR] feat: run length on binary input natively [datafusion-comet]
via GitHub
Re: [PR] feat: run length on binary input natively [datafusion-comet]
via GitHub
[I] ci: `datafusion-cli` MinIO docker login failed [datafusion]
via GitHub
Re: [I] ci: `datafusion-cli` MinIO docker login failed [datafusion]
via GitHub
Re: [I] ci: `datafusion-cli` MinIO docker login failed [datafusion]
via GitHub
Re: [I] [EPIC] Add public APIs required for scheduler high availability [datafusion-ballista]
via GitHub
Re: [I] [EPIC] Add public APIs required for scheduler high availability [datafusion-ballista]
via GitHub
Re: [I] [EPIC] Add public APIs required for scheduler high availability [datafusion-ballista]
via GitHub
Re: [I] [EPIC] Add public APIs required for scheduler high availability [datafusion-ballista]
via GitHub
[PR] chore: drop the redundant width_bucket shim registrations and guard serde uniqueness [datafusion-comet]
via GitHub
Re: [PR] chore: drop the redundant width_bucket shim registrations and guard serde uniqueness [datafusion-comet]
via GitHub
[PR] chore: guard RangeExpr protobuf fields exhaustively [datafusion]
via GitHub
Re: [PR] chore: guard RangeExpr protobuf fields exhaustively [datafusion]
via GitHub
Re: [I] FuzzDataGenerator silently drops nulls for Boolean/Byte/Short/Integer columns [datafusion-comet]
via GitHub
Re: [PR] fix: preserve nulls for Boolean/Byte/Short/Integer columns in FuzzDataGenerator [datafusion-comet]
via GitHub
Re: [I] Benchmark map hashing separately from map normalization for native shuffle [datafusion-comet]
via GitHub
[PR] fix: honor the S3 profile name and file and Hadoop's addressing mode for custom endpoints [datafusion-comet]
via GitHub
Re: [PR] fix: honor the S3 profile name and file and Hadoop's addressing mode for custom endpoints [datafusion-comet]
via GitHub
Re: [PR] fix: honor the S3 profile name and file and Hadoop's addressing mode for custom endpoints [datafusion-comet]
via GitHub
Re: [PR] fix: honor the S3 profile name and file and Hadoop's addressing mode for custom endpoints [datafusion-comet]
via GitHub
Re: [PR] fix: honor the S3 profile name and file and Hadoop's addressing mode for custom endpoints [datafusion-comet]
via GitHub
Re: [PR] fix: honor the S3 profile name and file and Hadoop's addressing mode for custom endpoints [datafusion-comet]
via GitHub
Re: [PR] fix: honor the S3 profile name and file and Hadoop's addressing mode for custom endpoints [datafusion-comet]
via GitHub
Re: [PR] fix: honor the S3 profile name and file and Hadoop's addressing mode for custom endpoints [datafusion-comet]
via GitHub
Re: [PR] fix: honor the S3 profile name and file and Hadoop's addressing mode for custom endpoints [datafusion-comet]
via GitHub
Re: [PR] fix: honor the S3 profile name and file and Hadoop's addressing mode for custom endpoints [datafusion-comet]
via GitHub
[I] CSV null_regex is applied to schema inference but never to the reader, so matching values are not null [datafusion]
via GitHub
Re: [I] CSV null_regex is applied to schema inference but never to the reader, so matching values are not null [datafusion]
via GitHub
[PR] fix: drop partition columns already present in the file schema of a ListingTable [datafusion]
via GitHub
Re: [PR] fix: drop partition columns already present in the file schema of a ListingTable [datafusion]
via GitHub
Re: [PR] fix: drop partition columns already present in the file schema of a ListingTable [datafusion]
via GitHub
[I] CsvReadOptions.null_regex has no effect: matching values are read as literal strings, and fail the read in a numeric column [datafusion-python]
via GitHub
Re: [I] CsvReadOptions.null_regex has no effect: matching values are read as literal strings, and fail the read in a numeric column [datafusion-python]
via GitHub
Re: [PR] fix: reject invalid placeholders in CREATE FUNCTION bodies at definition time [datafusion]
via GitHub
Re: [PR] fix: reject invalid placeholders in CREATE FUNCTION bodies at definition time [datafusion]
via GitHub
Re: [I] [EPIC] Use blocked / chunked memory management in hash aggregation [datafusion]
via GitHub
Re: [I] [EPIC] Use blocked / chunked memory management in hash aggregation [datafusion]
via GitHub
Re: [I] [EPIC] Use blocked / chunked memory management in hash aggregation [datafusion]
via GitHub
Re: [I] [EPIC] Use blocked / chunked memory management in hash aggregation [datafusion]
via GitHub
Re: [I] [EPIC] Use blocked / chunked memory management in hash aggregation [datafusion]
via GitHub
Re: [I] [EPIC] Use blocked / chunked memory management in hash aggregation [datafusion]
via GitHub
Re: [I] [EPIC] Use blocked / chunked memory management in hash aggregation [datafusion]
via GitHub
Re: [I] [EPIC] Use blocked / chunked memory management in hash aggregation [datafusion]
via GitHub
Re: [I] [EPIC] Use blocked / chunked memory management in hash aggregation [datafusion]
via GitHub
Re: [I] [EPIC] Use blocked / chunked memory management in hash aggregation [datafusion]
via GitHub
Re: [I] [EPIC] Use blocked / chunked memory management in hash aggregation [datafusion]
via GitHub
Re: [I] [EPIC] Use blocked / chunked memory management in hash aggregation [datafusion]
via GitHub
Re: [I] [EPIC] Use blocked / chunked memory management in hash aggregation [datafusion]
via GitHub
Re: [I] [EPIC] Use blocked / chunked memory management in hash aggregation [datafusion]
via GitHub
Re: [I] [EPIC] Use blocked / chunked memory management in hash aggregation [datafusion]
via GitHub
Re: [I] [EPIC] Use blocked / chunked memory management in hash aggregation [datafusion]
via GitHub
[PR] test: cover native memory accounting boundaries [datafusion-comet]
via GitHub
Re: [PR] test: cover native memory accounting boundaries [datafusion-comet]
via GitHub
Re: [PR] test: cover native memory accounting boundaries [datafusion-comet]
via GitHub
Re: [I] docs: Update "Adding a New Expression" to reference SQL file tests [datafusion-comet]
via GitHub
Re: [I] `LeftSemi` hash and nested loop joins report `EmissionType::Incremental` but emit only after the probe side is exhausted [datafusion]
via GitHub
Re: [PR] fix(physical-plan): report Final emission for LeftSemi hash and neste… [datafusion]
via GitHub
Re: [PR] perf: vectorize the native map lookup behind element_at and GetMapValue [datafusion-comet]
via GitHub
Re: [I] Map `element_at` is ~35x more expensive than every other map kernel, and slower than Spark [datafusion-comet]
via GitHub
[PR] DuckDB dialect: support % on LIMIT, like `LIMIT 5% OFFSET 20` [datafusion-sqlparser-rs]
via GitHub
Re: [PR] fix: preserve EXTRACT year comparisons with unrepresentable bounds [datafusion]
via GitHub
Re: [PR] fix: preserve EXTRACT year comparisons with unrepresentable bounds [datafusion]
via GitHub
Re: [PR] fix: preserve EXTRACT year comparisons with unrepresentable bounds [datafusion]
via GitHub
Re: [I] Support column reference as percentile arg for `percentile_cont` and `approx_percentile_cont` [datafusion]
via GitHub
Re: [I] Support column reference as percentile arg for `percentile_cont` and `approx_percentile_cont` [datafusion]
via GitHub
Re: [I] Support column reference as percentile arg for `percentile_cont` and `approx_percentile_cont` [datafusion]
via GitHub
Re: [I] Add support for Spark 4.2.0-preview4 [datafusion-comet]
via GitHub
[I] Give the two FFI example crates a runnable entry point [datafusion-python]
via GitHub
[I] Design note: composing multi-partition tasks with plan-level AQE in DataFusion [datafusion-ballista]
via GitHub
Re: [PR] feat(pwmj): support RightSemi/RightAnti existence joins [datafusion]
via GitHub
Re: [I] Chore: Move BloomFilterAgg to spark-expr crate [datafusion-comet]
via GitHub
Re: [I] Chore: Move BloomFilterAgg to spark-expr crate [datafusion-comet]
via GitHub
Re: [PR] perf: optimizing take_n for DictionaryGroupValuesColumn [datafusion]
via GitHub
[I] Ordered ARRAY_AGG accumulator retains one Arrow array per update_batch call, inflating per-row memory for grouped aggregation [datafusion]
via GitHub
Re: [I] Ordered ARRAY_AGG accumulator retains one Arrow array per update_batch call, inflating per-row memory for grouped aggregation [datafusion]
via GitHub
Re: [PR] ci: retry the Maven wrapper bootstrap in every job that calls ./mvnw directly [datafusion-comet]
via GitHub
Re: [PR] ci: retry the Maven wrapper bootstrap in every job that calls ./mvnw directly [datafusion-comet]
via GitHub
Re: [PR] ci: retry the Maven wrapper bootstrap in every job that calls ./mvnw directly [datafusion-comet]
via GitHub
[PR] feat: support direct Variant projection in native Parquet scans [datafusion-comet]
via GitHub
Re: [PR] feat: support direct Variant projection in native Parquet scans [datafusion-comet]
via GitHub
Re: [PR] perf: buffer the NestedLoopJoin build side as coalesced chunks instead of one concat_batches allocation [datafusion]
via GitHub
Re: [PR] perf: buffer the NestedLoopJoin build side as coalesced chunks instead of one concat_batches allocation [datafusion]
via GitHub
Re: [PR] perf: buffer the NestedLoopJoin build side as coalesced chunks instead of one concat_batches allocation [datafusion]
via GitHub
Re: [PR] perf: buffer the NestedLoopJoin build side as coalesced chunks instead of one concat_batches allocation [datafusion]
via GitHub
Re: [PR] perf: buffer the NestedLoopJoin build side as coalesced chunks instead of one concat_batches allocation [datafusion]
via GitHub
Re: [PR] perf: buffer the NestedLoopJoin build side as coalesced chunks instead of one concat_batches allocation [datafusion]
via GitHub
Re: [PR] perf: buffer the NestedLoopJoin build side as coalesced chunks instead of one concat_batches allocation [datafusion]
via GitHub
Re: [PR] perf: buffer the NestedLoopJoin build side as coalesced chunks instead of one concat_batches allocation [datafusion]
via GitHub
Re: [PR] feat: serialize ASOF join plans [datafusion]
via GitHub
Re: [PR] feat: serialize ASOF join plans [datafusion]
via GitHub
Re: [PR] feat: serialize ASOF join plans [datafusion]
via GitHub
Re: [I] width_bucket bypasses CometExpressionSerde framework [datafusion-comet]
via GitHub
Re: [PR] Support predicate subqueries in projections [datafusion]
via GitHub
Re: [PR] Support predicate subqueries in projections [datafusion]
via GitHub
Re: [PR] Support predicate subqueries in projections [datafusion]
via GitHub
[PR] fix: dispatch map lookups with normalized keys and nondeterministic null-guarded children [datafusion-comet]
via GitHub
Re: [PR] fix: dispatch map lookups with normalized keys and nondeterministic null-guarded children [datafusion-comet]
via GitHub
Re: [PR] fix: dispatch map lookups with normalized keys and nondeterministic null-guarded children [datafusion-comet]
via GitHub
Re: [I] Support higher-order array functions via JVM UDF bridge [datafusion-comet]
via GitHub
[PR] DRAFT - Test aggmetrics2 [datafusion]
via GitHub
Re: [PR] DRAFT - Test aggmetrics2 [datafusion]
via GitHub
Re: [PR] DRAFT - Test aggmetrics2 [datafusion]
via GitHub
Re: [PR] DRAFT - Test aggmetrics2 [datafusion]
via GitHub
Re: [I] Remove orphaned files in comet iceberg split writers [datafusion-comet]
via GitHub
Re: [PR] fix: report a negative array_resize size as a user error [datafusion]
via GitHub
Re: [PR] fix: report a negative array_resize size as a user error [datafusion]
via GitHub
Re: [PR] fix: report a negative array_resize size as a user error [datafusion]
via GitHub
Re: [PR] fix: report a negative array_resize size as a user error [datafusion]
via GitHub
Re: [PR] fix: apply Spark's Parquet conversion rules to nested struct/list/map fields [datafusion-comet]
via GitHub
Re: [PR] fix: apply Spark's Parquet conversion rules to nested struct/list/map fields [datafusion-comet]
via GitHub
Re: [I] Nested (struct/list/map) Parquet schema-evolution conversions bypass Spark's type-conversion rules: silent NULLs, string parsing, and a native panic [datafusion-comet]
via GitHub
[PR] perf: reduce Spark cache row conversion overhead [datafusion-comet]
via GitHub
Re: [PR] perf: fuse Comet cache vector reads into Spark codegen [datafusion-comet]
via GitHub
Re: [PR] perf: fuse Comet cache vector reads into Spark codegen [datafusion-comet]
via GitHub
Re: [PR] test: restore ANSI array access error coverage [datafusion-comet]
via GitHub
Re: [PR] test: restore ANSI array access error coverage [datafusion-comet]
via GitHub
Re: [PR] test: restore ANSI array access error coverage [datafusion-comet]
via GitHub
[PR] ci: require extended tests in the merge queue [datafusion]
via GitHub
Re: [PR] ci: require extended tests in the merge queue [datafusion]
via GitHub
Re: [PR] ci: require extended tests in the merge queue [datafusion]
via GitHub
Re: [PR] ci: require extended tests in the merge queue [datafusion]
via GitHub
Re: [PR] ci: require extended tests in the merge queue [datafusion]
via GitHub
Re: [PR] ci: require extended tests in the merge queue [datafusion]
via GitHub
Re: [PR] ci: require extended tests in the merge queue [datafusion]
via GitHub
Re: [PR] ci: require extended tests in the merge queue [datafusion]
via GitHub
Re: [PR] ci: require extended tests in the merge queue [datafusion]
via GitHub
Re: [PR] ci: require extended tests in the merge queue [datafusion]
via GitHub
Re: [PR] ci: require extended tests in the merge queue [datafusion]
via GitHub
Re: [PR] feat: route MapSort fallback through the codegen dispatcher [datafusion-comet]
via GitHub
Re: [PR] feat: route MapSort fallback through the codegen dispatcher [datafusion-comet]
via GitHub
[I] Substrait: the grouping set column holds `__grouping_id`, not the set's index [datafusion]
via GitHub
[PR] DuckDB dialect support for escape string literals [datafusion-sqlparser-rs]
via GitHub
Re: [PR] fix: prevent panic and incorrect results for COUNT with ORDER BY [datafusion]
via GitHub
Re: [I] Tracking fs-hdfs issues [datafusion-comet]
via GitHub
Earlier messages
Later messages