[PATCH 2/2] AArch64: Optimise SVE FP64 Hyperbolics
Wilco Dijkstra
Wilco.Dijkstra@arm.com
Wed Jun 18 17:33:34 GMT 2025
Hi Dylan,
> Reworked SVE FP64 hyperbolics to use the SVE FEXPA
> instruction.
>
> Also updated the special case handelling for large
> inputs to be entirely vectorised.
>
> Performance improvements on Neoverse V1:
>
> cosh_sve: 19% for |x| < 709, 5x otherwise
> sinh_sve: 24% for |x| < 709, 5.9x otherwise
> tanh_sve: 12% for |x| < 19, 9x otherwise
OK - committed.
Cheers,
Wilco
More information about the Libc-alpha
mailing list