Implementation notes: amd64, pmnod076, crypto_aead/pi32cipher128v2

Computer: pmnod076
Architecture: amd64
CPU ID: GenuineIntel-00050657-bfebfbff
SUPERCOP version: 20191017
Operation: crypto_aead
Primitive: pi32cipher128v2
TimeObject sizeTest sizeImplementationCompilerBenchmark dateSUPERCOP version
4988415843 8 028974 832 896goptvgcc_-march=native_-mtune=native_-O2_-fomit-frame-pointer_-fwrapv_-fPIC_-fPIE2019102320191017
5055015026 8 027982 832 896goptvgcc_-march=native_-mtune=native_-O_-fomit-frame-pointer_-fwrapv_-fPIC_-fPIE2019102320191017
5119618723 8 034611 840 896goptvgcc_-march=native_-mtune=native_-O3_-fomit-frame-pointer_-fwrapv_-fPIC_-fPIE2019102320191017
949327272 8 019302 816 896goptvgcc_-march=native_-mtune=native_-Os_-fomit-frame-pointer_-fwrapv_-fPIC_-fPIE2019102320191017
1959687070 8 022915 840 896ref3gcc_-march=native_-mtune=native_-O3_-fomit-frame-pointer_-fwrapv_-fPIC_-fPIE2019102320191017
2005445766 8 018886 832 896ref3gcc_-march=native_-mtune=native_-O2_-fomit-frame-pointer_-fwrapv_-fPIC_-fPIE2019102320191017
2032644801 8 017758 832 896ref3gcc_-march=native_-mtune=native_-O_-fomit-frame-pointer_-fwrapv_-fPIC_-fPIE2019102320191017
21528247940 8 066877 864 928ref3icc_-march=skylake-avx512_-mtune=skylake-avx512_-O2_-fomit-frame-pointer2019102320191017
2252224183 8 016118 816 896ref3gcc_-march=native_-mtune=native_-Os_-fomit-frame-pointer_-fwrapv_-fPIC_-fPIE2019102320191017
23842451540 8 071005 864 928ref3icc_-march=icelake-client_-mtune=icelake-client_-O2_-fomit-frame-pointer2019102320191017
24338246836 8 063829 864 864ref3icc_-march=corei7_-mtune=corei7_-O2_-fomit-frame-pointer2019102320191017
24383647972 8 066277 864 864ref3icc_-march=core-avx-i_-mtune=core-avx-i_-O2_-fomit-frame-pointer2019102320191017
24437647972 8 066277 864 864ref3icc_-march=sandybridge_-mtune=sandybridge_-O2_-fomit-frame-pointer2019102320191017
24468047972 8 066277 864 864ref3icc_-march=ivybridge_-mtune=ivybridge_-O2_-fomit-frame-pointer2019102320191017
24482247972 8 066277 864 864ref3icc_-march=corei7-avx_-mtune=corei7-avx_-O2_-fomit-frame-pointer2019102320191017
24568249092 8 067749 864 864ref3icc_-march=skylake_-mtune=skylake_-O2_-fomit-frame-pointer2019102320191017
24578649092 8 067749 864 864ref3icc_-march=haswell_-mtune=haswell_-O2_-fomit-frame-pointer2019102320191017
24598849092 8 067749 864 864ref3icc_-march=core-avx2_-mtune=core-avx2_-O2_-fomit-frame-pointer2019102320191017
24629249092 8 067749 864 864ref3icc_-march=broadwell_-mtune=broadwell_-O2_-fomit-frame-pointer2019102320191017
31700430964 8 049909 864 928ref2icc_-march=skylake-avx512_-mtune=skylake-avx512_-O2_-fomit-frame-pointer2019102320191017
35586629644 8 049117 864 928ref2icc_-march=icelake-client_-mtune=icelake-client_-O2_-fomit-frame-pointer2019102320191017
40388833244 8 051909 864 864ref2icc_-march=broadwell_-mtune=broadwell_-O2_-fomit-frame-pointer2019102320191017
40478633260 8 051925 864 864ref2icc_-march=haswell_-mtune=haswell_-O2_-fomit-frame-pointer2019102320191017
40797233244 8 051909 864 864ref2icc_-march=skylake_-mtune=skylake_-O2_-fomit-frame-pointer2019102320191017
40825033260 8 051925 864 864ref2icc_-march=core-avx2_-mtune=core-avx2_-O2_-fomit-frame-pointer2019102320191017
43409431988 8 050301 864 864ref2icc_-march=corei7-avx_-mtune=corei7-avx_-O2_-fomit-frame-pointer2019102320191017
43410431988 8 050301 864 864ref2icc_-march=core-avx-i_-mtune=core-avx-i_-O2_-fomit-frame-pointer2019102320191017
43412831988 8 050301 864 864ref2icc_-march=ivybridge_-mtune=ivybridge_-O2_-fomit-frame-pointer2019102320191017
43540230996 8 048013 864 864ref2icc_-march=corei7_-mtune=corei7_-O2_-fomit-frame-pointer2019102320191017
43591031988 8 050301 864 864ref2icc_-march=sandybridge_-mtune=sandybridge_-O2_-fomit-frame-pointer2019102320191017
5642286071 8 019230 832 896ref2gcc_-march=native_-mtune=native_-O2_-fomit-frame-pointer_-fwrapv_-fPIC_-fPIE2019102320191017
5683665523 8 018502 832 896ref2gcc_-march=native_-mtune=native_-O_-fomit-frame-pointer_-fwrapv_-fPIC_-fPIE2019102320191017
5922684634 8 016590 816 896ref2gcc_-march=native_-mtune=native_-Os_-fomit-frame-pointer_-fwrapv_-fPIC_-fPIE2019102320191017
8713849114 8 024947 840 896ref2gcc_-march=native_-mtune=native_-O3_-fomit-frame-pointer_-fwrapv_-fPIC_-fPIE2019102320191017

Checksum failure

Implementation: crypto_aead/pi32cipher128v2/ref2
Compiler: icc -march=broadwell -mtune=broadwell -O3 -fomit-frame-pointer
bee7ef5815d3c849beb2297b37e3cab1d5cee87e6cfbead22d9682f9dc97bba2
Number of similar (compiler,implementation) pairs: 22, namely:
CompilerImplementations
icc -march=broadwell -mtune=broadwell -O3 -fomit-frame-pointer ref2 ref3
icc -march=core-avx-i -mtune=core-avx-i -O3 -fomit-frame-pointer ref2 ref3
icc -march=core-avx2 -mtune=core-avx2 -O3 -fomit-frame-pointer ref2 ref3
icc -march=corei7-avx -mtune=corei7-avx -O3 -fomit-frame-pointer ref2 ref3
icc -march=corei7 -mtune=corei7 -O3 -fomit-frame-pointer ref2 ref3
icc -march=haswell -mtune=haswell -O3 -fomit-frame-pointer ref2 ref3
icc -march=icelake-client -mtune=icelake-client -O3 -fomit-frame-pointer ref2 ref3
icc -march=ivybridge -mtune=ivybridge -O3 -fomit-frame-pointer ref2 ref3
icc -march=sandybridge -mtune=sandybridge -O3 -fomit-frame-pointer ref2 ref3
icc -march=skylake-avx512 -mtune=skylake-avx512 -O3 -fomit-frame-pointer ref2 ref3
icc -march=skylake -mtune=skylake -O3 -fomit-frame-pointer ref2 ref3

Test failure

Implementation: crypto_aead/pi32cipher128v2/optimized_nonSSE
Compiler: gcc -march=native -mtune=native -O2 -fomit-frame-pointer -fwrapv -fPIC -fPIE
error 111
crypto_aead_encrypt returns more than crypto_aead_ABYTES extra bytes

Number of similar (compiler,implementation) pairs: 56, namely:
CompilerImplementations
gcc -march=native -mtune=native -O2 -fomit-frame-pointer -fwrapv -fPIC -fPIE optimized_nonSSE ref
gcc -march=native -mtune=native -O3 -fomit-frame-pointer -fwrapv -fPIC -fPIE optimized_nonSSE ref
gcc -march=native -mtune=native -O -fomit-frame-pointer -fwrapv -fPIC -fPIE optimized_nonSSE ref
gcc -march=native -mtune=native -Os -fomit-frame-pointer -fwrapv -fPIC -fPIE optimized_nonSSE ref
icc -march=broadwell -mtune=broadwell -O2 -fomit-frame-pointer optimized_nonSSE ref
icc -march=broadwell -mtune=broadwell -O3 -fomit-frame-pointer optimized_nonSSE ref
icc -march=cannonlake -mtune=cannonlake -O2 -fomit-frame-pointer optimized_nonSSE ref
icc -march=cannonlake -mtune=cannonlake -O3 -fomit-frame-pointer optimized_nonSSE ref
icc -march=core-avx-i -mtune=core-avx-i -O2 -fomit-frame-pointer optimized_nonSSE ref
icc -march=core-avx-i -mtune=core-avx-i -O3 -fomit-frame-pointer optimized_nonSSE ref
icc -march=core-avx2 -mtune=core-avx2 -O2 -fomit-frame-pointer optimized_nonSSE ref
icc -march=core-avx2 -mtune=core-avx2 -O3 -fomit-frame-pointer optimized_nonSSE ref
icc -march=corei7-avx -mtune=corei7-avx -O2 -fomit-frame-pointer optimized_nonSSE ref
icc -march=corei7-avx -mtune=corei7-avx -O3 -fomit-frame-pointer optimized_nonSSE ref
icc -march=corei7 -mtune=corei7 -O2 -fomit-frame-pointer optimized_nonSSE ref
icc -march=corei7 -mtune=corei7 -O3 -fomit-frame-pointer optimized_nonSSE ref
icc -march=haswell -mtune=haswell -O2 -fomit-frame-pointer optimized_nonSSE ref
icc -march=haswell -mtune=haswell -O3 -fomit-frame-pointer optimized_nonSSE ref
icc -march=icelake-client -mtune=icelake-client -O2 -fomit-frame-pointer optimized_nonSSE ref
icc -march=icelake-client -mtune=icelake-client -O3 -fomit-frame-pointer optimized_nonSSE ref
icc -march=ivybridge -mtune=ivybridge -O2 -fomit-frame-pointer optimized_nonSSE ref
icc -march=ivybridge -mtune=ivybridge -O3 -fomit-frame-pointer optimized_nonSSE ref
icc -march=sandybridge -mtune=sandybridge -O2 -fomit-frame-pointer optimized_nonSSE ref
icc -march=sandybridge -mtune=sandybridge -O3 -fomit-frame-pointer optimized_nonSSE ref
icc -march=skylake-avx512 -mtune=skylake-avx512 -O2 -fomit-frame-pointer optimized_nonSSE ref
icc -march=skylake-avx512 -mtune=skylake-avx512 -O3 -fomit-frame-pointer optimized_nonSSE ref
icc -march=skylake -mtune=skylake -O2 -fomit-frame-pointer optimized_nonSSE ref
icc -march=skylake -mtune=skylake -O3 -fomit-frame-pointer optimized_nonSSE ref

Test failure

Implementation: crypto_aead/pi32cipher128v2/ref2
Compiler: icc -march=cannonlake -mtune=cannonlake -O2 -fomit-frame-pointer
error 111

Number of similar (compiler,implementation) pairs: 4, namely:
CompilerImplementations
icc -march=cannonlake -mtune=cannonlake -O2 -fomit-frame-pointer ref2 ref3
icc -march=cannonlake -mtune=cannonlake -O3 -fomit-frame-pointer ref2 ref3

Compiler output

Implementation: crypto_aead/pi32cipher128v2/goptv
Compiler: icc -march=broadwell -mtune=broadwell -O2 -fomit-frame-pointer
pi-cipher.c: pi-cipher.c(273): error: vector operation not supported with these operand types
pi-cipher.c: return (x << n) | (x >> ((PI_WORD_SIZE) - n));
pi-cipher.c: ^
pi-cipher.c:
pi-cipher.c: pi-cipher.c(273): error: vector operation not supported with these operand types
pi-cipher.c: return (x << n) | (x >> ((PI_WORD_SIZE) - n));
pi-cipher.c: ^
pi-cipher.c:
pi-cipher.c: pi-cipher.c(286): error: "__builtin_shuffle" is not supported
pi-cipher.c: n_t += __builtin_shuffle(y, g_mask);
pi-cipher.c: ^
pi-cipher.c:
pi-cipher.c: pi-cipher.c(287): error: "__builtin_shuffle" is not supported
pi-cipher.c: n_t += __builtin_shuffle(y, n_mask);
pi-cipher.c: ^
pi-cipher.c:
pi-cipher.c: pi-cipher.c(290): error: "__builtin_shuffle" is not supported
pi-cipher.c: n_t ^= __builtin_shuffle(n_t, n_x_1) ^ __builtin_shuffle(n_t, n_x_2);
pi-cipher.c: ^
pi-cipher.c:
pi-cipher.c: pi-cipher.c(290): error: "__builtin_shuffle" is not supported
pi-cipher.c: n_t ^= __builtin_shuffle(n_t, n_x_1) ^ __builtin_shuffle(n_t, n_x_2);
pi-cipher.c: ^
pi-cipher.c:
pi-cipher.c: pi-cipher.c(305): error: "__builtin_shuffle" is not supported
pi-cipher.c: ...

Number of similar (compiler,implementation) pairs: 24, namely:
CompilerImplementations
icc -march=broadwell -mtune=broadwell -O2 -fomit-frame-pointer goptv
icc -march=broadwell -mtune=broadwell -O3 -fomit-frame-pointer goptv
icc -march=cannonlake -mtune=cannonlake -O2 -fomit-frame-pointer goptv
icc -march=cannonlake -mtune=cannonlake -O3 -fomit-frame-pointer goptv
icc -march=core-avx-i -mtune=core-avx-i -O2 -fomit-frame-pointer goptv
icc -march=core-avx-i -mtune=core-avx-i -O3 -fomit-frame-pointer goptv
icc -march=core-avx2 -mtune=core-avx2 -O2 -fomit-frame-pointer goptv
icc -march=core-avx2 -mtune=core-avx2 -O3 -fomit-frame-pointer goptv
icc -march=corei7-avx -mtune=corei7-avx -O2 -fomit-frame-pointer goptv
icc -march=corei7-avx -mtune=corei7-avx -O3 -fomit-frame-pointer goptv
icc -march=corei7 -mtune=corei7 -O2 -fomit-frame-pointer goptv
icc -march=corei7 -mtune=corei7 -O3 -fomit-frame-pointer goptv
icc -march=haswell -mtune=haswell -O2 -fomit-frame-pointer goptv
icc -march=haswell -mtune=haswell -O3 -fomit-frame-pointer goptv
icc -march=icelake-client -mtune=icelake-client -O2 -fomit-frame-pointer goptv
icc -march=icelake-client -mtune=icelake-client -O3 -fomit-frame-pointer goptv
icc -march=ivybridge -mtune=ivybridge -O2 -fomit-frame-pointer goptv
icc -march=ivybridge -mtune=ivybridge -O3 -fomit-frame-pointer goptv
icc -march=sandybridge -mtune=sandybridge -O2 -fomit-frame-pointer goptv
icc -march=sandybridge -mtune=sandybridge -O3 -fomit-frame-pointer goptv
icc -march=skylake-avx512 -mtune=skylake-avx512 -O2 -fomit-frame-pointer goptv
icc -march=skylake-avx512 -mtune=skylake-avx512 -O3 -fomit-frame-pointer goptv
icc -march=skylake -mtune=skylake -O2 -fomit-frame-pointer goptv
icc -march=skylake -mtune=skylake -O3 -fomit-frame-pointer goptv