[PATCH 2/2] aarch64: Improve codegen for SVE log1pf users
Wilco Dijkstra
Wilco.Dijkstra@arm.com
Fri Jan 3 21:21:13 GMT 2025
Hi Long,
> Reduce memory access by using lanewise MLA and reduce number of MOVPRFXs.
> Move log1pf implementation to inline helper function.
> Speedup on Neoverse V1 for log1pf (10%), acoshf (-1%), atanhf (2%), asinhf (2%).
> ---
> OK for master? If so please commit for as I don't have commit rights.
OK - committed.
Cheers,
Wilco
More information about the Libc-alpha
mailing list