Implementation notes: amd64, scw1b63b1, crypto_kem/ntskem13136

Computer: scw1b63b1
Architecture: amd64
CPU ID: GenuineIntel-000506f1-0f8bfbff
SUPERCOP version: 20191017
Operation: crypto_kem
Primitive: ntskem13136
TimeObject sizeTest sizeImplementationCompilerBenchmark dateSUPERCOP version
337321697946 84 16114262 976 1576sse2clang_-march=native_-O2_-fomit-frame-pointer_-fwrapv_-Qunused-arguments_-fPIC_-fPIE2019122320191017
340994899180 84 16116326 976 1576sse2clang_-march=native_-O3_-fomit-frame-pointer_-fwrapv_-Qunused-arguments_-fPIC_-fPIE2019122320191017
341877083131 84 1698608 968 1576sse2clang_-march=native_-Os_-fomit-frame-pointer_-fwrapv_-Qunused-arguments_-fPIC_-fPIE2019122320191017
342307697946 84 16114262 976 1576sse2clang_-march=native_-O_-fomit-frame-pointer_-fwrapv_-Qunused-arguments_-fPIC_-fPIE2019122320191017
3715144104464 84 16122910 960 1640sse2gcc_-march=native_-mtune=native_-O3_-fomit-frame-pointer_-fwrapv_-fPIC_-fPIE2019122320191017
374180083475 84 1699598 960 1640sse2gcc_-march=native_-mtune=native_-O_-fomit-frame-pointer_-fwrapv_-fPIC_-fPIE2019122320191017
374960684509 84 16100846 960 1640sse2gcc_-march=native_-mtune=native_-O2_-fomit-frame-pointer_-fwrapv_-fPIC_-fPIE2019122320191017
382583675933 84 1691293 952 1608sse2gcc_-march=native_-mtune=native_-Os_-fomit-frame-pointer_-fwrapv_-fPIC_-fPIE2019122320191017
561925277260 84 1695262 960 1640optgcc_-march=native_-mtune=native_-O2_-fomit-frame-pointer_-fwrapv_-fPIC_-fPIE2019122320191017
5669624111932 84 16131990 960 1640optgcc_-march=native_-mtune=native_-O3_-fomit-frame-pointer_-fwrapv_-fPIC_-fPIE2019122320191017
573200093754 84 16112422 976 1576optclang_-march=native_-O3_-fomit-frame-pointer_-fwrapv_-Qunused-arguments_-fPIC_-fPIE2019122320191017
573963476401 84 1694078 960 1640optgcc_-march=native_-mtune=native_-O_-fomit-frame-pointer_-fwrapv_-fPIC_-fPIE2019122320191017
575066277819 84 1694800 968 1576optclang_-march=native_-Os_-fomit-frame-pointer_-fwrapv_-Qunused-arguments_-fPIC_-fPIE2019122320191017
576520092879 84 16110694 976 1576optclang_-march=native_-O_-fomit-frame-pointer_-fwrapv_-Qunused-arguments_-fPIC_-fPIE2019122320191017
579502692879 84 16110694 976 1576optclang_-march=native_-O2_-fomit-frame-pointer_-fwrapv_-Qunused-arguments_-fPIC_-fPIE2019122320191017
588611496576 84 16115310 976 1576optclang_-mcpu=native_-O3_-fomit-frame-pointer_-fwrapv_-Qunused-arguments_-fPIC_-fPIE2019122320191017
621462270052 84 1686813 952 1608optgcc_-march=native_-mtune=native_-Os_-fomit-frame-pointer_-fwrapv_-fPIC_-fPIE2019122320191017
3516850434643 76 1653194 936 1576refclang_-march=native_-O3_-fomit-frame-pointer_-fwrapv_-Qunused-arguments_-fPIC_-fPIE2019122320191017
3529054231405 76 1649106 936 1576refclang_-march=native_-O_-fomit-frame-pointer_-fwrapv_-Qunused-arguments_-fPIC_-fPIE2019122320191017
3530774831405 76 1649106 936 1576refclang_-march=native_-O2_-fomit-frame-pointer_-fwrapv_-Qunused-arguments_-fPIC_-fPIE2019122320191017
3683716835981 76 1654626 936 1576refclang_-mcpu=native_-O3_-fomit-frame-pointer_-fwrapv_-Qunused-arguments_-fPIC_-fPIE2019122320191017
3797529439810 76 1659790 928 1640refgcc_-march=native_-mtune=native_-O3_-fomit-frame-pointer_-fwrapv_-fPIC_-fPIE2019122320191017
3797754822832 76 1639588 928 1576refclang_-march=native_-Os_-fomit-frame-pointer_-fwrapv_-Qunused-arguments_-fPIC_-fPIE2019122320191017
4043687225296 76 1643230 928 1640refgcc_-march=native_-mtune=native_-O2_-fomit-frame-pointer_-fwrapv_-fPIC_-fPIE2019122320191017
4195738823754 76 1641382 928 1640refgcc_-march=native_-mtune=native_-O_-fomit-frame-pointer_-fwrapv_-fPIC_-fPIE2019122320191017
4487034420037 76 1636677 920 1608refgcc_-march=native_-mtune=native_-Os_-fomit-frame-pointer_-fwrapv_-fPIC_-fPIE2019122320191017

Compiler output

Implementation: crypto_kem/ntskem13136/avx2
Compiler: clang -march=native -O2 -fomit-frame-pointer -fwrapv -Qunused-arguments -fPIC -fPIE
bitslice_fft_256.c: bitslice_fft_256.c:103:29: error: always_inline function '_mm256_setr_epi32' requires target feature 'xsave', but would be inlined into function 'bitslice_butterflies13_256' that is compiled without support for 'xsave'
bitslice_fft_256.c: const __m256i mask256 = _mm256_setr_epi32(0, 2, 4, 6, 1, 3, 5, 7);
bitslice_fft_256.c: ^
bitslice_fft_256.c: bitslice_fft_256.c:127:25: error: always_inline function '_mm256_set_epi32' requires target feature 'xsave', but would be inlined into function 'bitslice_butterflies13_256' that is compiled without support for 'xsave'
bitslice_fft_256.c: out[j][i] = _mm256_set_epi32(-u32[7], -u32[6], -u32[5], -u32[4],
bitslice_fft_256.c: ^
bitslice_fft_256.c: bitslice_fft_256.c:140:20: error: always_inline function '_mm256_permutevar8x32_epi32' requires target feature 'xsave', but would be inlined into function 'bitslice_butterflies13_256' that is compiled without support for 'xsave'
bitslice_fft_256.c: t[0] = _mm256_permutevar8x32_epi32(out[ j ][b], mask256);
bitslice_fft_256.c: ^
bitslice_fft_256.c: bitslice_fft_256.c:141:20: error: always_inline function '_mm256_permutevar8x32_epi32' requires target feature 'xsave', but would be inlined into function 'bitslice_butterflies13_256' that is compiled without support for 'xsave'
bitslice_fft_256.c: ...
bitslice_fft_256.c: clang: note: diagnostic msg:
bitslice_fft_256.c: ********************
bitslice_fft_256.c:
bitslice_fft_256.c: PLEASE ATTACH THE FOLLOWING FILES TO THE BUG REPORT:
bitslice_fft_256.c: Preprocessed source(s) and associated run script(s) are located at:
bitslice_fft_256.c: clang: note: diagnostic msg: /tmp/bitslice_fft_256-8b4c88.c
bitslice_fft_256.c: clang: note: diagnostic msg: /tmp/bitslice_fft_256-8b4c88.sh
bitslice_fft_256.c: clang: note: diagnostic msg:
bitslice_fft_256.c:
bitslice_fft_256.c: ********************

Number of similar (compiler,implementation) pairs: 1, namely:
CompilerImplementations
clang -march=native -O2 -fomit-frame-pointer -fwrapv -Qunused-arguments -fPIC -fPIE avx2

Compiler output

Implementation: crypto_kem/ntskem13136/avx2
Compiler: clang -march=native -O3 -fomit-frame-pointer -fwrapv -Qunused-arguments -fPIC -fPIE
bitslice_fft_256.c: bitslice_fft_256.c:103:29: error: always_inline function '_mm256_setr_epi32' requires target feature 'xsave', but would be inlined into function 'bitslice_butterflies13_256' that is compiled without support for 'xsave'
bitslice_fft_256.c: const __m256i mask256 = _mm256_setr_epi32(0, 2, 4, 6, 1, 3, 5, 7);
bitslice_fft_256.c: ^
bitslice_fft_256.c: bitslice_fft_256.c:127:25: error: always_inline function '_mm256_set_epi32' requires target feature 'xsave', but would be inlined into function 'bitslice_butterflies13_256' that is compiled without support for 'xsave'
bitslice_fft_256.c: out[j][i] = _mm256_set_epi32(-u32[7], -u32[6], -u32[5], -u32[4],
bitslice_fft_256.c: ^
bitslice_fft_256.c: bitslice_fft_256.c:140:20: error: always_inline function '_mm256_permutevar8x32_epi32' requires target feature 'xsave', but would be inlined into function 'bitslice_butterflies13_256' that is compiled without support for 'xsave'
bitslice_fft_256.c: t[0] = _mm256_permutevar8x32_epi32(out[ j ][b], mask256);
bitslice_fft_256.c: ^
bitslice_fft_256.c: bitslice_fft_256.c:141:20: error: always_inline function '_mm256_permutevar8x32_epi32' requires target feature 'xsave', but would be inlined into function 'bitslice_butterflies13_256' that is compiled without support for 'xsave'
bitslice_fft_256.c: ...
bitslice_fft_256.c: clang: note: diagnostic msg:
bitslice_fft_256.c: ********************
bitslice_fft_256.c:
bitslice_fft_256.c: PLEASE ATTACH THE FOLLOWING FILES TO THE BUG REPORT:
bitslice_fft_256.c: Preprocessed source(s) and associated run script(s) are located at:
bitslice_fft_256.c: clang: note: diagnostic msg: /tmp/bitslice_fft_256-34f6ed.c
bitslice_fft_256.c: clang: note: diagnostic msg: /tmp/bitslice_fft_256-34f6ed.sh
bitslice_fft_256.c: clang: note: diagnostic msg:
bitslice_fft_256.c:
bitslice_fft_256.c: ********************

Number of similar (compiler,implementation) pairs: 1, namely:
CompilerImplementations
clang -march=native -O3 -fomit-frame-pointer -fwrapv -Qunused-arguments -fPIC -fPIE avx2

Compiler output

Implementation: crypto_kem/ntskem13136/avx2
Compiler: clang -march=native -O -fomit-frame-pointer -fwrapv -Qunused-arguments -fPIC -fPIE
bitslice_fft_256.c: bitslice_fft_256.c:103:29: error: always_inline function '_mm256_setr_epi32' requires target feature 'xsave', but would be inlined into function 'bitslice_butterflies13_256' that is compiled without support for 'xsave'
bitslice_fft_256.c: const __m256i mask256 = _mm256_setr_epi32(0, 2, 4, 6, 1, 3, 5, 7);
bitslice_fft_256.c: ^
bitslice_fft_256.c: bitslice_fft_256.c:127:25: error: always_inline function '_mm256_set_epi32' requires target feature 'xsave', but would be inlined into function 'bitslice_butterflies13_256' that is compiled without support for 'xsave'
bitslice_fft_256.c: out[j][i] = _mm256_set_epi32(-u32[7], -u32[6], -u32[5], -u32[4],
bitslice_fft_256.c: ^
bitslice_fft_256.c: bitslice_fft_256.c:140:20: error: always_inline function '_mm256_permutevar8x32_epi32' requires target feature 'xsave', but would be inlined into function 'bitslice_butterflies13_256' that is compiled without support for 'xsave'
bitslice_fft_256.c: t[0] = _mm256_permutevar8x32_epi32(out[ j ][b], mask256);
bitslice_fft_256.c: ^
bitslice_fft_256.c: bitslice_fft_256.c:141:20: error: always_inline function '_mm256_permutevar8x32_epi32' requires target feature 'xsave', but would be inlined into function 'bitslice_butterflies13_256' that is compiled without support for 'xsave'
bitslice_fft_256.c: ...
bitslice_fft_256.c: clang: note: diagnostic msg:
bitslice_fft_256.c: ********************
bitslice_fft_256.c:
bitslice_fft_256.c: PLEASE ATTACH THE FOLLOWING FILES TO THE BUG REPORT:
bitslice_fft_256.c: Preprocessed source(s) and associated run script(s) are located at:
bitslice_fft_256.c: clang: note: diagnostic msg: /tmp/bitslice_fft_256-f1d2a3.c
bitslice_fft_256.c: clang: note: diagnostic msg: /tmp/bitslice_fft_256-f1d2a3.sh
bitslice_fft_256.c: clang: note: diagnostic msg:
bitslice_fft_256.c:
bitslice_fft_256.c: ********************

Number of similar (compiler,implementation) pairs: 1, namely:
CompilerImplementations
clang -march=native -O -fomit-frame-pointer -fwrapv -Qunused-arguments -fPIC -fPIE avx2

Compiler output

Implementation: crypto_kem/ntskem13136/avx2
Compiler: clang -march=native -Os -fomit-frame-pointer -fwrapv -Qunused-arguments -fPIC -fPIE
bitslice_fft_256.c: bitslice_fft_256.c:103:29: error: always_inline function '_mm256_setr_epi32' requires target feature 'xsave', but would be inlined into function 'bitslice_butterflies13_256' that is compiled without support for 'xsave'
bitslice_fft_256.c: const __m256i mask256 = _mm256_setr_epi32(0, 2, 4, 6, 1, 3, 5, 7);
bitslice_fft_256.c: ^
bitslice_fft_256.c: bitslice_fft_256.c:127:25: error: always_inline function '_mm256_set_epi32' requires target feature 'xsave', but would be inlined into function 'bitslice_butterflies13_256' that is compiled without support for 'xsave'
bitslice_fft_256.c: out[j][i] = _mm256_set_epi32(-u32[7], -u32[6], -u32[5], -u32[4],
bitslice_fft_256.c: ^
bitslice_fft_256.c: bitslice_fft_256.c:140:20: error: always_inline function '_mm256_permutevar8x32_epi32' requires target feature 'xsave', but would be inlined into function 'bitslice_butterflies13_256' that is compiled without support for 'xsave'
bitslice_fft_256.c: t[0] = _mm256_permutevar8x32_epi32(out[ j ][b], mask256);
bitslice_fft_256.c: ^
bitslice_fft_256.c: bitslice_fft_256.c:141:20: error: always_inline function '_mm256_permutevar8x32_epi32' requires target feature 'xsave', but would be inlined into function 'bitslice_butterflies13_256' that is compiled without support for 'xsave'
bitslice_fft_256.c: ...
bitslice_fft_256.c: clang: note: diagnostic msg:
bitslice_fft_256.c: ********************
bitslice_fft_256.c:
bitslice_fft_256.c: PLEASE ATTACH THE FOLLOWING FILES TO THE BUG REPORT:
bitslice_fft_256.c: Preprocessed source(s) and associated run script(s) are located at:
bitslice_fft_256.c: clang: note: diagnostic msg: /tmp/bitslice_fft_256-e82322.c
bitslice_fft_256.c: clang: note: diagnostic msg: /tmp/bitslice_fft_256-e82322.sh
bitslice_fft_256.c: clang: note: diagnostic msg:
bitslice_fft_256.c:
bitslice_fft_256.c: ********************

Number of similar (compiler,implementation) pairs: 1, namely:
CompilerImplementations
clang -march=native -Os -fomit-frame-pointer -fwrapv -Qunused-arguments -fPIC -fPIE avx2

Compiler output

Implementation: crypto_kem/ntskem13136/avx2
Compiler: clang -mcpu=native -O3 -fomit-frame-pointer -fwrapv -Qunused-arguments -fPIC -fPIE
bitslice_bma_128.c: bitslice_bma_128.c:315:21: error: always_inline function '_mm_shuffle_epi8' requires target feature 'ssse3', but would be inlined into function 'bitslice_bma' that is compiled without support for 'ssse3'
bitslice_bma_128.c: out[1][j] = _mm_shuffle_epi8(out[0][j], mask);
bitslice_bma_128.c: ^
bitslice_bma_128.c: bitslice_bma_128.c:316:21: error: '__builtin_ia32_palignr128' needs target feature ssse3
bitslice_bma_128.c: out[0][j] = _mm_alignr_epi8(out[0][j], psi[0][j], 15);
bitslice_bma_128.c: ^
bitslice_bma_128.c: /usr/lib/llvm-3.8/bin/../lib/clang/3.8.1/include/tmmintrin.h:69:12: note: expanded from macro '_mm_alignr_epi8'
bitslice_bma_128.c: (__m128i)__builtin_ia32_palignr128((__v16qi)(__m128i)(a), ^
bitslice_bma_128.c: 2 errors generated.

Number of similar (compiler,implementation) pairs: 2, namely:
CompilerImplementations
clang -mcpu=native -O3 -fomit-frame-pointer -fwrapv -Qunused-arguments -fPIC -fPIE avx2 sse2

Compiler output

Implementation: crypto_kem/ntskem13136/avx2
Compiler: gcc -march=native -mtune=native -O2 -fomit-frame-pointer -fwrapv -fPIC -fPIE
bitslice_fft_256.c: bitslice_fft_256.c: In function ‘bitslice_butterflies13_256’:
bitslice_fft_256.c: bitslice_fft_256.c:103:19: warning: AVX vector return without AVX enabled changes the ABI [-Wpsabi]
bitslice_fft_256.c: const __m256i mask256 = _mm256_setr_epi32(0, 2, 4, 6, 1, 3, 5, 7);
bitslice_fft_256.c: ^~~~~~~
bitslice_fft_256.c: In file included from /usr/lib/gcc/x86_64-linux-gnu/6/include/immintrin.h:41:0,
bitslice_fft_256.c: from bitslice_fft_256.h:21,
bitslice_fft_256.c: from bitslice_fft_256.c:19:
bitslice_fft_256.c: /usr/lib/gcc/x86_64-linux-gnu/6/include/avxintrin.h:1338:1: error: inlining failed in call to always_inline ‘_mm256_setr_epi32’: target specific option mismatch
bitslice_fft_256.c: _mm256_setr_epi32 (int __A, int __B, int __C, int __D,
bitslice_fft_256.c: ^~~~~~~~~~~~~~~~~
bitslice_fft_256.c: ...
bitslice_fft_256.c: ^~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~
bitslice_fft_256.c: In file included from /usr/lib/gcc/x86_64-linux-gnu/6/include/immintrin.h:43:0,
bitslice_fft_256.c: from bitslice_fft_256.h:21,
bitslice_fft_256.c: from bitslice_fft_256.c:19:
bitslice_fft_256.c: /usr/lib/gcc/x86_64-linux-gnu/6/include/avx2intrin.h:913:1: error: inlining failed in call to always_inline ‘_mm256_xor_si256’: target specific option mismatch
bitslice_fft_256.c: _mm256_xor_si256 (__m256i __A, __m256i __B)
bitslice_fft_256.c: ^~~~~~~~~~~~~~~~
bitslice_fft_256.c: bitslice_fft_256.c:198:35: note: called from here
bitslice_fft_256.c: out[k][b] = _mm256_xor_si256(out[k][b], t[b]);
bitslice_fft_256.c: ^~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~

Number of similar (compiler,implementation) pairs: 4, namely:
CompilerImplementations
gcc -march=native -mtune=native -O2 -fomit-frame-pointer -fwrapv -fPIC -fPIE avx2
gcc -march=native -mtune=native -O3 -fomit-frame-pointer -fwrapv -fPIC -fPIE avx2
gcc -march=native -mtune=native -O -fomit-frame-pointer -fwrapv -fPIC -fPIE avx2
gcc -march=native -mtune=native -Os -fomit-frame-pointer -fwrapv -fPIC -fPIE avx2