[PATCH] aarch64: Improve codegen of AdvSIMD atan(2)(f)

Wilco Dijkstra Wilco.Dijkstra@arm.com
Tue Dec 17 15:31:28 GMT 2024


Hi Joana,

> Load the polynomial evaluation coefficients into 2 vectors and use lanewise MLAs.
> 8% improvement in throughput microbenchmark on Neoverse V1.

OK - pushed (needed an extra return at the end).

Cheers,
Wilco


More information about the Libc-alpha mailing list