[PATCH] aarch64: optimized memcpy implementation for thunderx2

Siddhesh Poyarekar siddhesh@gotplt.org
Tue Oct 2 03:35:00 GMT 2018


On 01/10/18 9:52 PM, Anton Youdkevitch wrote:
> memcpy-large, the original T2 implementation as the baseline
> 
<snip>
> 
> 
> memcpy-walk, the original T2 implementation as the baseline

Very nice, thanks for posting this!  Please add a subset of these 
results (highlights or at least specific names of the benchmarks with 
the result ranges) to your commit log.  I'll leave the actual code 
review to Steve.

Thanks,
Siddhesh



More information about the Libc-alpha mailing list