> Patch optimizes Float16 to integral conversion operations. Currently, its a > two step process where by first a Float16 value is > converted to a single precision floating point value followed by a conversion > to an integral value. > > x86 targets supporting AVX512-FP16 feature (Intel Sapphire Rapids+ and > upcoming AMD Zen6) provides direct instruction to convert a Float16 value to > integral value. > > Following are the performance numbers of micro benchmark included with the > patch on Granite Rapids with and without auto-vectorization. > > <img width="1125" height="636" alt="image" > src="https://github.com/user-attachments/assets/ca6e6757-1579-475f-8307-9454c7c025c1" > /> > > Kindly review and share your feedback. > > Best Regards, > Jatin > > --------- > - [x] I confirm that I make this contribution in accordance with the [OpenJDK > Interim AI Policy](https://openjdk.org/legal/ai).
Jatin Bhateja has updated the pull request with a new target base due to a merge or a rebase. The incremental webrev excludes the unrelated changes brought in by the merge/rebase. The pull request contains seven additional commits since the last revision: - Merge branch 'master' of http://github.com/openjdk/jdk into JDK-8382523 - Review comments resolution - Review comments resolution - Review comments resolution - Review comments resolution - Review comments resolution - 8382523: Optimize Float16 to integral conversion operations for AVX512-FP16 targets ------------- Changes: - all: https://git.openjdk.org/jdk/pull/30928/files - new: https://git.openjdk.org/jdk/pull/30928/files/24ec724a..5af9c53c Webrevs: - full: https://webrevs.openjdk.org/?repo=jdk&pr=30928&range=06 - incr: https://webrevs.openjdk.org/?repo=jdk&pr=30928&range=05-06 Stats: 628843 lines in 7107 files changed: 395362 ins; 195145 del; 38336 mod Patch: https://git.openjdk.org/jdk/pull/30928.diff Fetch: git fetch https://git.openjdk.org/jdk.git pull/30928/head:pull/30928 PR: https://git.openjdk.org/jdk/pull/30928
