fkjellberg opened a new pull request, #812:
URL: https://github.com/apache/commons-compress/pull/812

   <!--
     Licensed to the Apache Software Foundation (ASF) under one
     or more contributor license agreements.  See the NOTICE file
     distributed with this work for additional information
     regarding copyright ownership.  The ASF licenses this file
     to you under the Apache License, Version 2.0 (the
     "License"); you may not use this file except in compliance
     with the License.  You may obtain a copy of the License at
   
       https://www.apache.org/licenses/LICENSE-2.0
   
     Unless required by applicable law or agreed to in writing,
     software distributed under the License is distributed on an
     "AS IS" BASIS, WITHOUT WARRANTIES OR CONDITIONS OF ANY
     KIND, either express or implied.  See the License for the
     specific language governing permissions and limitations
     under the License.
   -->
   
   Thanks for your contribution to [Apache 
Commons](https://commons.apache.org/)! Your help is appreciated!
   
   Before you push a pull request, review this list:
   
   - [X] Read the [contribution guidelines](CONTRIBUTING.md) for this project.
   - [ ] Read the [ASF Generative Tooling 
Guidance](https://www.apache.org/legal/generative-tooling.html) if you use 
Artificial Intelligence (AI).
   - [ ] I used AI to create any part of, or all of, this pull request. Which 
AI tool was used to create this pull request, and to what extent did it 
contribute?
   - [X] Run a successful build using the default 
[Maven](https://maven.apache.org/) goal with `mvn`; that's `mvn` on the command 
line by itself.
   - [X] Write unit tests that match behavioral changes, where the tests fail 
if the changes to the runtime are not applied. This may not always be possible, 
but it is a best practice.
   - [X] Write a pull request description that is detailed enough to understand 
what the pull request does, how, and why.
   - [X] Each commit in the pull request should have a meaningful subject line 
and body. Note that a maintainer may squash commits during the merge process.
   
   This PR replaces the `BinaryTree` Huffman decoder with the new canonical 
`HuffmanDecoder` already used by bzip2 and deflate64. Running the existing 
benchmark test shows a small performance gain.
   
   `BinaryTree` implementation:
   
   ```
   Benchmark                                         Mode  Cnt     Score    
Error  Units
   Lh5CompressorInputStreamBenchmark.testDecompress  avgt   25  1542,895 ± 
41,839  us/op
   ```
   
   `LhaHuffmanDecoder` implementation:
   
   ```
   Benchmark                                         Mode  Cnt     Score    
Error  Units
   Lh5CompressorInputStreamBenchmark.testDecompress  avgt   25  1515,176 ± 
41,576  us/op
   ```
   
   The main benefit of switching to the `HuffmanDecoder` apart from better code 
reuse is lower RAM usage. For the benchmark test, the `BinaryTree` 
implementation allocated 131,068 and 4,092 bytes to hold the two huffman trees 
during decompression while the `HuffmanDecoder` allocates 1,116 and 84 bytes.
   
   The `BinaryTree` implementation was originally copied from the zip package 
and adapted for LHA. The original class is still used by 
`ExplodingInputStream`, and I think it's worth investigating whether it could 
be replaced by `HuffmanDecoder` as well.
   


-- 
This is an automated message from the Apache Git Service.
To respond to the message, please log on to GitHub and use the
URL above to go to the specific comment.

To unsubscribe, e-mail: [email protected]

For queries about this service, please contact Infrastructure at:
[email protected]

Reply via email to