[PATCH 2/2] AArch64: Optimise SVE FP64 Hyperbolics

Wilco Dijkstra Wilco.Dijkstra@arm.com
Wed Jun 18 17:33:34 GMT 2025


Hi Dylan,

> Reworked SVE FP64 hyperbolics to use the SVE FEXPA
> instruction.
>
> Also updated the special case handelling for large
> inputs to be entirely vectorised.
>
> Performance improvements on Neoverse V1:
>
> cosh_sve: 19% for |x| < 709, 5x otherwise
> sinh_sve: 24% for |x| < 709, 5.9x otherwise
> tanh_sve: 12% for |x| < 19,  9x otherwise

OK - committed.

Cheers,
Wilco


More information about the Libc-alpha mailing list