> It turns out he 2-step FP16 to integral conversion process is more beneficial 
> on platforms that support AVX10.2 instructions. This is mainly due to the 
> reduced instruction count and automatic special case handling (e.g., NaN). 
> With that in mind, the changes in this PR go back to the original approach 
> when AVX10.2 is detected in the C2 compiler. There are also some updates to 
> the JTREG tests and JMH benchmarks.
> 
> The JTREG test listed below was used to verify correctness with 
> `-XX:-UseSuperWord` and `-XX:+UseSuperWord` JVM options applied. All 
> modifications and tests used [OpenJDK 
> v28-b15](https://github.com/openjdk/jdk/releases/tag/jdk-28%2B15) as the 
> baseline build.
> 
> 1. 
> `jtreg:test/hotspot/jtreg/compiler/vectorapi/TestFloat16ToIntegralConv.java`
> 
> ---------
> - [x] I confirm that I make this contribution in accordance with the [OpenJDK 
> Interim AI Policy](https://openjdk.org/legal/ai).

Mohamed Issa has updated the pull request incrementally with one additional 
commit since the last revision:

  Use AVX10.2 fp16 to byte direct conversion vector instruction whenever 
possible.

-------------

Changes:
  - all: https://git.openjdk.org/jdk/pull/32957/files
  - new: https://git.openjdk.org/jdk/pull/32957/files/14abc268..84e80949

Webrevs:
 - full: https://webrevs.openjdk.org/?repo=jdk&pr=32957&range=01
 - incr: https://webrevs.openjdk.org/?repo=jdk&pr=32957&range=00-01

  Stats: 37 lines in 7 files changed: 34 ins; 2 del; 1 mod
  Patch: https://git.openjdk.org/jdk/pull/32957.diff
  Fetch: git fetch https://git.openjdk.org/jdk.git pull/32957/head:pull/32957

PR: https://git.openjdk.org/jdk/pull/32957

Reply via email to