Messages by Thread
-
-
[PR] chore: Add helpers for computing buffer offset span, length [datafusion]
via GitHub
-
[I] Wrong results: correlated `NOT IN` loses the correlation when it is an equality on the `IN` value column [datafusion]
via GitHub
-
Re: [PR] chore(deps): bump reqwest from 0.13.4 to 0.13.5 [datafusion-ballista]
via GitHub
-
Re: [I] Comet 1.1.0 Release (September) [datafusion-comet]
via GitHub
-
[PR] perf: Improve efficiency of `array_resize` for sliced arrays [datafusion]
via GitHub
-
[PR] docs: explain allocator hazards and diagram where memory is allocated [datafusion-comet]
via GitHub
-
Re: [PR] chore(ci): bump the github-codeql group across 1 directory with 2 updates [datafusion-ballista]
via GitHub
-
[I] Wrong results: nullable constant `NOT IN (subquery)` returns every row when the subquery column is `NOT NULL` [datafusion]
via GitHub
-
Re: [PR] chore(ci): bump astral-sh/setup-uv from 9.0.0 to 10.1.0 [datafusion-ballista]
via GitHub
-
[I] perf: broadcast hash join rebuilds the build-side hash table in every task [datafusion-comet]
via GitHub
-
[PR] fix: keep renamed columns intact in leaf projection pushdown [datafusion]
via GitHub
-
Re: [PR] IN LIST: retain short lists with specialized filters [datafusion]
via GitHub
-
[I] Wrong results: `NOT IN (subquery)` ignores NULLs from a nullable outer expression over `NOT NULL` columns [datafusion]
via GitHub
-
[I] Wrong results: COALESCE on a volatile operand evaluates it two times, and fails when the output is declared non-nullable [datafusion]
via GitHub
-
[PR] fix: evaluate a volatile `BETWEEN` value one time [datafusion]
via GitHub
-
Re: [PR] feat: Fix output bytes metric in hash agg [datafusion]
via GitHub
-
Re: [I] [DISCUSSION] Rethinking Code Review and Testing in the Agent Era [datafusion]
via GitHub
-
[PR] fix: preserve null treatment when unparsing window and aggregate functions [datafusion]
via GitHub
-
[PR] fix: correct `LEAD/LAG IGNORE NULLS` evaluation and limit pushdown [datafusion]
via GitHub
-
[I] `LEAD/LAG IGNORE NULLS` returns incorrect results across `NULL` gaps and with `LIMIT` [datafusion]
via GitHub
-
[PR] chore(deps): bump the codeql-actions group with 2 updates [datafusion-comet]
via GitHub
-
[PR] chore(deps): bump actions/setup-java from 4 to 6 [datafusion-comet]
via GitHub
-
[PR] Refactor: replace `dialect_of!` table partition check with `Dialect` trait method [datafusion-sqlparser-rs]
via GitHub
-
[PR] feat: add GroupColumn support for Decimal32/Decimal64 [datafusion]
via GitHub
-
[PR] feat: derive per-column pruning guarantees from tuple IN lists [datafusion]
via GitHub
-
[PR] interleave retained rows at emit instead of a take per group entry [datafusion]
via GitHub
-
[PR] feat(pwmj): support left/right mark joins [datafusion]
via GitHub
-
[PR] dev: include examples README checks in the local lint suite [datafusion]
via GitHub
-
[I] Provide typed evaluate_bounds for date_bin and from_unixtime [datafusion]
via GitHub
-
[I] Restore date_trunc ordering propagation for safe timezone-aware timestamps [datafusion]
via GitHub
-
Re: [I] Nested array comparison does not match Spark for signed zero [datafusion-comet]
via GitHub
-
[I] Enable Bloom-filter pruning for joins on multiple columns [datafusion]
via GitHub
-
Re: [PR] perf: add direct native reads for broadcast exchange [datafusion-comet]
via GitHub
-
[PR] chore(deps): bump soupsieve from 2.8.3 to 2.9 [datafusion-sandbox]
via GitHub
-
Re: [PR] feat: Lambda function support from DataFusion, illustrated with array_filter [datafusion-comet]
via GitHub
-
Re: [PR] fix: include ABFS container in object store cache key [datafusion-comet]
via GitHub
-
[PR] chore(deps): bump clap from 4.6.6 to 4.6.7 [datafusion-ballista]
via GitHub
-
[PR] chore(deps): bump rustls from 0.23.44 to 0.23.45 [datafusion-ballista]
via GitHub
-
[PR] chore(ci): bump taiki-e/install-action from 2.87.7 to 2.87.13 [datafusion-ballista]
via GitHub
-
[PR] chore(deps): bump graphviz-rust from 0.9.8 to 0.9.9 [datafusion-ballista]
via GitHub
-
Re: [I] Duplicate field ids inside a struct are not validated when the file schema equals the requested schema and no predicate is pushed [datafusion-comet]
via GitHub
-
[I] SQL unparser drops IGNORE NULLS / RESPECT NULLS on window and aggregate functions [datafusion]
via GitHub
-
Re: [PR] fix: Derive Substrait intersection nullability from every input [datafusion]
via GitHub
-
Re: [PR] chore(ci): Register the substrait check job with runs-on, route spark checks through xtask [datafusion]
via GitHub
-
[PR] fix: keep a DISTINCT aggregate's arity in `SingleDistinctToGroupBy` [datafusion]
via GitHub
-
Re: [I] Integer interval inference can incorrectly remove ORDER BY [datafusion]
via GitHub
-
[PR] feat: Load Parquet page indexes for external row selections [datafusion]
via GitHub
-
[I] [EPIC] Leaf expression pushdown and filter pushdown: bugs, decisions and tests [datafusion]
via GitHub
-
[PR] feat: add a "no new evaluations" optimizer invariant, switched off [datafusion]
via GitHub
-
[I] BETWEEN evaluates its operand two times, so a volatile operand returns wrong rows [datafusion]
via GitHub
-
Re: [PR] fix: `get_json_object` returns first value for duplicate keys to match Spark [datafusion-comet]
via GitHub
-
[PR] fix: one inlining policy for the rules that inline a projection column [datafusion]
via GitHub
-
[PR] fix: decide the precedence between PushDownFilter and PushDownLeafProjections [datafusion]
via GitHub
-
[I] Support Parquet page index loading for external row selections [datafusion]
via GitHub
-
[PR] test: differential fuzz for leaf expression pushdown and parquet filter pushdown [datafusion]
via GitHub
-
[PR] test: close five mutation-testing gaps in leaf expression pushdown guards [datafusion]
via GitHub
-
[I] Test gaps in PushDownFilter and leaf expression pushdown found by mutation testing [datafusion]
via GitHub
-
Re: [I] Iceberg scan falls back to Spark on IS NULL/IS NOT NULL over list/map columns (stale complex-type check) [datafusion-comet]
via GitHub
-
[I] Let the data source veto `ExpressionPlacement::MoveTowardsLeafNodes` per function [datafusion]
via GitHub
-
[I] Replace the `__datafusion_extracted` / `__common_expr` alias prefix protocol with a typed marker [datafusion]
via GitHub
-
[I] Resolve columns by schema index, not by name, in ExtractLeafExpressions and PushDownLeafProjections [datafusion]
via GitHub
-
[I] Leaf extraction computes the same expression two times in one projection [datafusion]
via GitHub
-
[I] `push_down_leaf_projections` fails with an ambiguous schema when a sub-query projection renames columns [datafusion]
via GitHub
-
Re: [PR] feat(parquet): Enable Parquet `filter_pushdown` by default, with heurstic fallback when projecting few non-filter columns [datafusion]
via GitHub
-
Re: [PR] feat: reuse partial aggregate hashes across repartition [datafusion]
via GitHub
-
[PR] fix: keep the recovery projection when leaf pushdown would drop a computed same-name column [datafusion]
via GitHub
-
Re: [PR] refactor: separate compact IN-list pruning threshold from the default cap [datafusion]
via GitHub
-
[PR] build(deps): bump soupsieve from 2.8.4 to 2.9 [datafusion-python]
via GitHub
-
[PR] chore(deps): bump soupsieve from 2.8.4 to 2.9 [datafusion]
via GitHub
-
Re: [PR] feat: wire native existence join support [datafusion-comet]
via GitHub