On Wed, 22 Jul 2026 11:24:30 GMT, Jatin Bhateja <[email protected]> wrote:
>> - Currently for masked Float16 intrinsified vector operation we emit a >> sequence of BLEND instruction which merges the result of complete vector >> operation with passthrough vector under the influence of mask. >> - Targets supporting AVX512-FP16 feature offers direct predicated >> instructions. >> - This patch adds the support to infer predicated vector >> ADD/SUB/MUL/DIV/FMA/SQRT/MIN/MAX operation on AVX512-FP16 targets >> >> Following are the performance numbers of benchmark included with the patch >> on AVX512-FP16 target (Intel Granite Rapids) >> <img width="1497" height="857" alt="image" >> src="https://github.com/user-attachments/assets/5600449e-e946-4bab-ae99-2e13dc454b56" >> /> >> >> Kindly review and share your feedback. >> >> Best Regards, >> Jatin >> >> >> >> --------- >> - [x] I confirm that I make this contribution in accordance with the >> [OpenJDK Interim AI Policy](https://openjdk.org/legal/ai). > > Jatin Bhateja has refreshed the contents of this pull request, and previous > commits have been removed. Incremental views are not available. The pull > request now contains one commit: > > 8386957: C2 VectorAPI: Predicated operation support for intrinsified > Float16Vector unary/binary/ternary operations on AVX512-FP16 targets src/hotspot/cpu/x86/x86.ad line 25022: > 25020: match(Set dst (SubVHF (Binary dst src2) mask)); > 25021: match(Set dst (MulVHF (Binary dst src2) mask)); > 25022: match(Set dst (DivVHF (Binary dst src2) mask)); These could be made part of the corresponding vadd_reg_masked, vsub_reg_masked, ... for consistency with existing code. Likewise the other instructs below. ------------- PR Review Comment: https://git.openjdk.org/jdk/pull/32004#discussion_r4109292594
