These instructions are part of the ACE set, so I assume they are useful for AI - there are format conversions, rounding, comparison, and, of course, operations on these formats in the vector and mask registers. Looking at the documentation, which is huge, gives me the impression this will improve quantized models the most, but this is a guess.
What heavy CPU workloads AVX10 V2 AUX could help accelerate significantly?
It has a smells of 'we need to push back our IP deadlines' thingy.
These instructions are part of the ACE set, so I assume they are useful for AI - there are format conversions, rounding, comparison, and, of course, operations on these formats in the vector and mask registers. Looking at the documentation, which is huge, gives me the impression this will improve quantized models the most, but this is a guess.