git.ipfire.org Git - thirdparty/glibc.git/commit

elf: Optimize _dl_new_hash in dl-new-hash.h

Unroll slightly and enforce good instruction scheduling. This improves
performance on out-of-order machines. The unrolling allows for
pipelined multiplies.

As well, as an optional sysdep, reorder the operations and prevent
reassosiation for better scheduling and higher ILP. This commit
only adds the barrier for x86, although it should be either no
change or a win for any architecture.

Unrolling further started to induce slowdowns for sizes [0, 4]
but can help the loop so if larger sizes are the target further
unrolling can be beneficial.

Results for _dl_new_hash
Benchmarked on Tigerlake: 11th Gen Intel(R) Core(TM) i7-1165G7 @ 2.80GHz

Time as Geometric Mean of N=30 runs
Geometric of all benchmark New / Old: 0.674
  type, length, New Time, Old Time, New Time / Old Time
fixed,      0,    2.865,     2.72,               1.053
fixed,      1,    3.567,    2.489,               1.433
fixed,      2,    2.577,    3.649,               0.706
fixed,      3,    3.644,    5.983,               0.609
fixed,      4,    4.211,    6.833,               0.616
fixed,      5,    4.741,    9.372,               0.506
fixed,      6,    5.415,    9.561,               0.566
fixed,      7,    6.649,   10.789,               0.616
fixed,      8,    8.081,   11.808,               0.684
fixed,      9,    8.427,   12.935,               0.651
fixed,     10,    8.673,   14.134,               0.614
fixed,     11,    10.69,   15.408,               0.694
fixed,     12,   10.789,   16.982,               0.635
fixed,     13,   12.169,   18.411,               0.661
fixed,     14,   12.659,   19.914,               0.636
fixed,     15,   13.526,   21.541,               0.628
fixed,     16,   14.211,   23.088,               0.616
fixed,     32,   29.412,   52.722,               0.558
fixed,     64,    65.41,  142.351,               0.459
fixed,    128,  138.505,  295.625,               0.469
fixed,    256,  291.707,  601.983,               0.485
random,      2,   12.698,   12.849,               0.988
random,      4,   16.065,   15.857,               1.013
random,      8,   19.564,   21.105,               0.927
random,     16,   23.919,   26.823,               0.892
random,     32,   31.987,   39.591,               0.808
random,     64,   49.282,   71.487,               0.689
random,    128,    82.23,  145.364,               0.566
random,    256,  152.209,  298.434,                0.51

Co-authored-by: Alexander Monakov <amonakov@ispras.ru>
Reviewed-by: Siddhesh Poyarekar <siddhesh@sourceware.org>

author	Noah Goldstein <goldstein.w.n@gmail.com>
	Thu, 19 May 2022 22:18:03 +0000 (17:18 -0500)
committer	Noah Goldstein <goldstein.w.n@gmail.com>
	Mon, 23 May 2022 15:38:40 +0000 (10:38 -0500)
commit	9a421348cd7d0704663e26e6171828bed6e0a2cf
tree	8bdbf5e9d77420298f1795467dda50ed4e39afb0	tree
parent	3d155d4b6c29ddfd0b3318fa58dbf8ef20e7bca0	commit \| diff

benchtests/bench-dl-new-hash.c		diff \| blob \| blame \| history
elf/simple-dl-new-hash.h	[moved from elf/dl-new-hash.h with 75% similarity]	diff \| blob \| blame \| history
elf/tst-dl-hash.c		diff \| blob \| blame \| history
sysdeps/generic/dl-new-hash.h	[new file with mode: 0644]	blob
sysdeps/x86/dl-new-hash.h	[new file with mode: 0644]	blob