jonkeane commented on code in PR #51657:
URL: https://github.com/apache/arrow/pull/51657#discussion_r4144872331
##########
r/NEWS.md:
##########
@@ -19,13 +19,43 @@
# arrow 25.0.1.9000
+## Breaking changes
+
+- `read_feather()` and `write_feather()` now warn that they are deprecated. Use
+ `read_ipc_file()` and `write_ipc_file()` instead. Similarly,
+ `format = "feather"` in `open_dataset()` and `write_dataset()` is deprecated
+ in favour of `format = "ipc"`, and extra arguments passed via `...` to
+ `read_ipc_stream()` and `write_ipc_stream()` are deprecated and ignored
+ (#49237).
+
+## New features
+
+- New `AzureFileSystem` class and `az_container()` helper for working with
+ Azure Blob Storage, analogous to `S3FileSystem` and `s3_bucket()`. Azure
+ support is enabled by default when building from source on Linux and macOS,
+ provided the required system libraries are available; see
+ `vignette("install", package = "arrow")` (@marberts, #32123).
+
## Minor improvements and fixes
+- Variables with the same name as a function, such as `date`, can now be used
+ in dplyr verbs (#39688).
+- Reading Parquet files with `float16` columns now returns the correct values
+ (#50378).
+- `if_else()` now works when one branch is a bare `NA` and the other is a date
+ or timestamp (#38358).
+- `mutate()` with `if_any()` or `if_all()` now gives the new column the correct
+ name (#34860).
- Factor levels inside list columns are now unified across the whole column
when converting to R, so data read in multiple batches (e.g. via
`read_ipc_stream()` or `open_dataset()`) produces valid factors that can be
unnested. Similarly, `int64` and `uint32` values inside list columns are
converted to a single R type across the column (#50514).
+- `register_scalar_function()` now checks that the names in `in_type` match
+ the arguments of `fun`, instead of silently ignoring them (#37761).
+- `str_replace()` with an `NA` replacement now returns `NA` for matched
+ elements, matching stringr (@Gosling-dude, #33432).
+- `summarise()` after `arrange()` now works (#45373).
Review Comment:
Same with this, this seems bigger than (the last item of) minor improvements
##########
r/NEWS.md:
##########
@@ -19,13 +19,43 @@
# arrow 25.0.1.9000
+## Breaking changes
+
+- `read_feather()` and `write_feather()` now warn that they are deprecated. Use
+ `read_ipc_file()` and `write_ipc_file()` instead. Similarly,
+ `format = "feather"` in `open_dataset()` and `write_dataset()` is deprecated
+ in favour of `format = "ipc"`, and extra arguments passed via `...` to
+ `read_ipc_stream()` and `write_ipc_stream()` are deprecated and ignored
+ (#49237).
+
+## New features
+
+- New `AzureFileSystem` class and `az_container()` helper for working with
+ Azure Blob Storage, analogous to `S3FileSystem` and `s3_bucket()`. Azure
+ support is enabled by default when building from source on Linux and macOS,
+ provided the required system libraries are available; see
+ `vignette("install", package = "arrow")` (@marberts, #32123).
+
## Minor improvements and fixes
+- Variables with the same name as a function, such as `date`, can now be used
+ in dplyr verbs (#39688).
+- Reading Parquet files with `float16` columns now returns the correct values
+ (#50378).
+- `if_else()` now works when one branch is a bare `NA` and the other is a date
+ or timestamp (#38358).
+- `mutate()` with `if_any()` or `if_all()` now gives the new column the correct
+ name (#34860).
- Factor levels inside list columns are now unified across the whole column
when converting to R, so data read in multiple batches (e.g. via
`read_ipc_stream()` or `open_dataset()`) produces valid factors that can be
unnested. Similarly, `int64` and `uint32` values inside list columns are
converted to a single R type across the column (#50514).
+- `register_scalar_function()` now checks that the names in `in_type` match
+ the arguments of `fun`, instead of silently ignoring them (#37761).
Review Comment:
This might be more than a minor improvement, IMO. At the very least, it
might be good to be at the top fo the list?
--
This is an automated message from the Apache Git Service.
To respond to the message, please log on to GitHub and use the
URL above to go to the specific comment.
To unsubscribe, e-mail: [email protected]
For queries about this service, please contact Infrastructure at:
[email protected]