SPEC CPU®2017 Floating Point Rate Result

Lenovo Global Technology

ThinkSystem SR665
3.50 GHz, AMD EPYC 7F52

SPECrate®2017_fp_base = 34300

SPECrate®2017_fp_peak = 34700

CPU2017 License:	9017	Test Date:	Jul-2020
Test Sponsor:	Lenovo Global Technology	Hardware Availability:	Jun-2020
Tested by:	Lenovo Global Technology	Software Availability:	Nov-2019

Benchmark result graphs are available in the PDF report.

Hardware
CPU Name:	AMD EPYC 7F52
Max MHz:	3900
Nominal:	3500
Enabled:	32 cores, 2 chips, 2 threads/core
Orderable:	1,2 chips
Cache L1:	32 KB I + 32 KB D on chip per core
L2:	512 KB I+D on chip per core
L3:	256 MB I+D on chip per chip, 16 MB per core
Other:	None
Memory:	1 TB (32 x 32 GB 2Rx8 PC4-3200AA-R)
Storage:	1 x 960 GB SATA SSD
Other:	None

Software
OS:	Red Hat Enterprise Linux 8.1 (Ootpa) Kernel 4.18.0-147.el8.x86_64
Compiler:	C/C++/Fortran: Version 2.0.0 of AOCC
Parallel:	No
Firmware:	Lenovo BIOS Version D8E105P 1.00 released May-2020
File System:	xfs
System State:	Run level 3 (multi-user)
Base Pointers:	64-bit
Peak Pointers:	64-bit
Other:	jemalloc: jemalloc memory allocator library v5.2.0
Power Management:	BIOS set to prefer performance at the cost of additional power usage

Results Table

Benchmark	Base							Peak
Benchmark	Copies	Seconds	Ratio	Seconds	Ratio	Seconds	Ratio	Copies	Seconds	Ratio	Seconds	Ratio	Seconds	Ratio
SPECrate®2017_fp_base			34300
SPECrate®2017_fp_peak			34700
Results appear in the order in which they were run. Bold underlined text indicates a median measurement.
503.bwaves_r	64	873	735	877	732	877	732	64	873	735	877	732	877	732
507.cactuBSSN_r	64	163	496	164	495	163	497	64	163	496	163	497	163	496
508.namd_r	64	283	215	283	215	282	215	64	283	215	283	215	282	215
510.parest_r	64	474	354	471	355	473	354	64	474	354	471	355	473	354
511.povray_r	64	473	316	474	315	474	315	64	467	320	470	318	471	317
519.lbm_r	64	398	169	397	170	395	171	64	398	169	397	170	395	171
521.wrf_r	64	397	361	406	353	396	362	64	397	361	406	353	396	362
526.blender_r	64	287	339	288	339	289	337	64	287	339	288	339	289	337
527.cam4_r	64	371	302	372	301	369	303	64	368	304	373	300	369	303
538.imagick_r	64	185	860	185	862	184	863	64	183	869	183	870	185	861
544.nab_r	64	278	387	276	390	276	390	64	278	387	276	390	276	390
549.fotonik3d_r	64	1042	239	1039	240	1041	240	64	1042	239	1039	240	1041	240
554.roms_r	64	491	207	491	207	490	208	32	220	231	219	232	219	232

Compiler Notes

The AMD64 AOCC Compiler Suite is available at
http://developer.amd.com/amd-aocc/

Submit Notes

The config file option 'submit' was used.
'numactl' was used to bind copies to the cores.
See the configuration file for details.

Operating System Notes

'ulimit -s unlimited' was used to set environment stack size
'ulimit -l 2097152' was used to set environment locked pages in memory limit

runcpu command invoked through numactl i.e.:
numactl --interleave=all runcpu <etc>

Set dirty_ratio=8 to limit dirty cache to 8% of memory
Set swappiness=1 to swap only if necessary
Set zone_reclaim_mode=1 to free local node memory and avoid remote memory
sync then drop_caches=3 to reset caches before invoking runcpu

dirty_ratio, swappiness, zone_reclaim_mode and drop_caches were
all set using privileged echo (e.g. echo 1 > /proc/sys/vm/swappiness).

Transparent huge pages set to 'always' for this run (OS default)

Environment Variables Notes

Environment variables set by runcpu before the start of the run:
LD_LIBRARY_PATH =
     "/home/cpu2017-1.1.0-amd-rome-aocc200-C1/amd_rate_aocc200_rome_C_lib/64;
     /home/cpu2017-1.1.0-amd-rome-aocc200-C1/amd_rate_aocc200_rome_C_lib/32:"
MALLOC_CONF = "retain:true"

General Notes

Binaries were compiled on a system with 2x AMD EPYC 7601 CPU + 512GB Memory using Fedora 26

NA: The test sponsor attests, as of date of publication, that CVE-2017-5754 (Meltdown)
is mitigated in the system as tested and documented.
Yes: The test sponsor attests, as of date of publication, that CVE-2017-5753 (Spectre variant 1)
is mitigated in the system as tested and documented.
Yes: The test sponsor attests, as of date of publication, that CVE-2017-5715 (Spectre variant 2)
is mitigated in the system as tested and documented.
jemalloc: configured and built with GCC v9.1.0 in Ubuntu 19.04 with -O3 -znver2 -flto
jemalloc 5.2.0 is available here:
https://github.com/jemalloc/jemalloc/releases/download/5.2.0/jemalloc-5.2.0.tar.bz2

Platform Notes

 BIOS settings:
 Choose Operating Mode set to Maximum Performance and then set it to Custom Mode
 NUMA nodes per socket set to NPS4

 Sysinfo program /home/cpu2017-1.1.0-amd-rome-aocc200-C1/bin/sysinfo
 Rev: r6365 of 2019-08-21 295195f888a3d7edb1e6e46a485a0011
 running on localhost.localdomain Wed Jul 15 02:58:40 2020

 SUT (System Under Test) info as seen by some common utilities.
 For more information on this section, see
    https://www.spec.org/cpu2017/Docs/config.html#sysinfo

 From /proc/cpuinfo
    model name : AMD EPYC 7F52 16-Core Processor
       2  "physical id"s (chips)
       64 "processors"
    cores, siblings (Caution: counting these is hw and system dependent. The following
    excerpts from /proc/cpuinfo might not be reliable.  Use with caution.)
       cpu cores : 16
       siblings  : 32
       physical 0: cores 0 4 8 12 16 20 24 28 32 36 40 44 48 52 56 60
       physical 1: cores 0 4 8 12 16 20 24 28 32 36 40 44 48 52 56 60

 From lscpu:
      Architecture:        x86_64
      CPU op-mode(s):      32-bit, 64-bit
      Byte Order:          Little Endian
      CPU(s):              64
      On-line CPU(s) list: 0-63
      Thread(s) per core:  2
      Core(s) per socket:  16
      Socket(s):           2
      NUMA node(s):        8
      Vendor ID:           AuthenticAMD
      CPU family:          23
      Model:               49
      Model name:          AMD EPYC 7F52 16-Core Processor
      Stepping:            0
      CPU MHz:             2455.416
      CPU max MHz:         3500.0000
      CPU min MHz:         2500.0000
      BogoMIPS:            6987.46
      Virtualization:      AMD-V
      L1d cache:           32K
      L1i cache:           32K
      L2 cache:            512K
      L3 cache:            16384K
      NUMA node0 CPU(s):   0-3,32-35
      NUMA node1 CPU(s):   4-7,36-39
      NUMA node2 CPU(s):   8-11,40-43
      NUMA node3 CPU(s):   12-15,44-47
      NUMA node4 CPU(s):   16-19,48-51
      NUMA node5 CPU(s):   20-23,52-55
      NUMA node6 CPU(s):   24-27,56-59
      NUMA node7 CPU(s):   28-31,60-63
      Flags:               fpu vme de pse tsc msr pae mce cx8 apic sep mtrr pge mca cmov
      pat pse36 clflush mmx fxsr sse sse2 ht syscall nx mmxext fxsr_opt pdpe1gb rdtscp lm
      constant_tsc rep_good nopl xtopology nonstop_tsc cpuid extd_apicid aperfmperf pni
      pclmulqdq monitor ssse3 fma cx16 sse4_1 sse4_2 movbe popcnt aes xsave avx f16c
      rdrand lahf_lm cmp_legacy svm extapic cr8_legacy abm sse4a misalignsse 3dnowprefetch
      osvw ibs skinit wdt tce topoext perfctr_core perfctr_nb bpext perfctr_llc mwaitx cpb
      cat_l3 cdp_l3 hw_pstate sme ssbd mba sev ibrs ibpb stibp vmmcall fsgsbase bmi1 avx2
      smep bmi2 cqm rdt_a rdseed adx smap clflushopt clwb sha_ni xsaveopt xsavec xgetbv1
      xsaves cqm_llc cqm_occup_llc cqm_mbm_total cqm_mbm_local clzero irperf xsaveerptr
      wbnoinvd arat npt lbrv svm_lock nrip_save tsc_scale vmcb_clean flushbyasid
      decodeassists pausefilter pfthreshold avic v_vmsave_vmload vgif umip rdpid
      overflow_recov succor smca

 /proc/cpuinfo cache data
    cache size : 512 KB

 From numactl --hardware  WARNING: a numactl 'node' might or might not correspond to a
 physical chip.
   available: 8 nodes (0-7)
   node 0 cpus: 0 1 2 3 32 33 34 35
   node 0 size: 128801 MB
   node 0 free: 128640 MB
   node 1 cpus: 4 5 6 7 36 37 38 39
   node 1 size: 129020 MB
   node 1 free: 128847 MB
   node 2 cpus: 8 9 10 11 40 41 42 43
   node 2 size: 129020 MB
   node 2 free: 128886 MB
   node 3 cpus: 12 13 14 15 44 45 46 47
   node 3 size: 129008 MB
   node 3 free: 128881 MB
   node 4 cpus: 16 17 18 19 48 49 50 51
   node 4 size: 129020 MB
   node 4 free: 128855 MB
   node 5 cpus: 20 21 22 23 52 53 54 55
   node 5 size: 129020 MB
   node 5 free: 128866 MB
   node 6 cpus: 24 25 26 27 56 57 58 59
   node 6 size: 129020 MB
   node 6 free: 128866 MB
   node 7 cpus: 28 29 30 31 60 61 62 63
   node 7 size: 129020 MB
   node 7 free: 128832 MB
   node distances:
   node   0   1   2   3   4   5   6   7
     0:  10  12  12  12  32  32  32  32
     1:  12  10  12  12  32  32  32  32
     2:  12  12  10  12  32  32  32  32
     3:  12  12  12  10  32  32  32  32
     4:  32  32  32  32  10  12  12  12
     5:  32  32  32  32  12  10  12  12
     6:  32  32  32  32  12  12  10  12
     7:  32  32  32  32  12  12  12  10

 From /proc/meminfo
    MemTotal:       1056697444 kB
    HugePages_Total:       0
    Hugepagesize:       2048 kB

 From /etc/*release* /etc/*version*
    os-release:
       NAME="Red Hat Enterprise Linux"
       VERSION="8.1 (Ootpa)"
       ID="rhel"
       ID_LIKE="fedora"
       VERSION_ID="8.1"
       PLATFORM_ID="platform:el8"
       PRETTY_NAME="Red Hat Enterprise Linux 8.1 (Ootpa)"
       ANSI_COLOR="0;31"
    redhat-release: Red Hat Enterprise Linux release 8.1 (Ootpa)
    system-release: Red Hat Enterprise Linux release 8.1 (Ootpa)
    system-release-cpe: cpe:/o:redhat:enterprise_linux:8.1:ga

 uname -a:
    Linux localhost.localdomain 4.18.0-147.el8.x86_64 #1 SMP Thu Sep 26 15:52:44 UTC 2019
    x86_64 x86_64 x86_64 GNU/Linux

 Kernel self-reported vulnerability status:

 CVE-2018-3620 (L1 Terminal Fault):        Not affected
 Microarchitectural Data Sampling:         Not affected
 CVE-2017-5754 (Meltdown):                 Not affected
 CVE-2018-3639 (Speculative Store Bypass): Mitigation: Speculative Store Bypass disabled
                                           via prctl and seccomp
 CVE-2017-5753 (Spectre variant 1):        Mitigation: usercopy/swapgs barriers and __user
                                           pointer sanitization
 CVE-2017-5715 (Spectre variant 2):        Mitigation: Full AMD retpoline, IBPB:
                                           conditional, IBRS_FW, STIBP: conditional, RSB
                                           filling

 run-level 3 Jul 15 02:56

 SPEC is set to: /home/cpu2017-1.1.0-amd-rome-aocc200-C1
    Filesystem     Type  Size  Used Avail Use% Mounted on
    /dev/sdb2      xfs   838G  124G  714G  15% /home

 From /sys/devices/virtual/dmi/id
     BIOS:    Lenovo D8E105P-1.00 05/08/2020
     Vendor:  Lenovo
     Product: ThinkSystem SR665 MB
     Product Family: ThinkSystem
     Serial:  1234567890

 Additional information from dmidecode follows.  WARNING: Use caution when you interpret
 this section. The 'dmidecode' program reads system data which is "intended to allow
 hardware to be accurately determined", but the intent may not be met, as there are
 frequent changes to hardware, firmware, and the "DMTF SMBIOS" standard.
   Memory:
     32x Samsung M393A4G43AB3-CWE 32 kB 2 rank 3200

 (End of data from sysinfo program)
This system support 16 DIMMs per processor, total 32 DIMMs.
32 DIMM slots installed with 32 GB DIMM for this run.

Compiler Version Notes

==============================================================================
C               | 519.lbm_r(base, peak) 538.imagick_r(base, peak)
                | 544.nab_r(base, peak)
------------------------------------------------------------------------------
AOCC.LLVM.2.0.0.B191.2019_07_19 clang version 8.0.0 (CLANG: Jenkins
  AOCC_2_0_0-Build#191) (based on LLVM AOCC.LLVM.2.0.0.B191.2019_07_19)
Target: x86_64-unknown-linux-gnu
Thread model: posix
InstalledDir: /sppo/dev/compilers/aocc-compiler-2.0.0/bin
------------------------------------------------------------------------------

==============================================================================
C++             | 508.namd_r(base, peak) 510.parest_r(base, peak)
------------------------------------------------------------------------------
AOCC.LLVM.2.0.0.B191.2019_07_19 clang version 8.0.0 (CLANG: Jenkins
  AOCC_2_0_0-Build#191) (based on LLVM AOCC.LLVM.2.0.0.B191.2019_07_19)
Target: x86_64-unknown-linux-gnu
Thread model: posix
InstalledDir: /sppo/dev/compilers/aocc-compiler-2.0.0/bin
------------------------------------------------------------------------------

==============================================================================
C++, C          | 511.povray_r(base, peak) 526.blender_r(base, peak)
------------------------------------------------------------------------------
AOCC.LLVM.2.0.0.B191.2019_07_19 clang version 8.0.0 (CLANG: Jenkins
  AOCC_2_0_0-Build#191) (based on LLVM AOCC.LLVM.2.0.0.B191.2019_07_19)
Target: x86_64-unknown-linux-gnu
Thread model: posix
InstalledDir: /sppo/dev/compilers/aocc-compiler-2.0.0/bin
AOCC.LLVM.2.0.0.B191.2019_07_19 clang version 8.0.0 (CLANG: Jenkins
  AOCC_2_0_0-Build#191) (based on LLVM AOCC.LLVM.2.0.0.B191.2019_07_19)
Target: x86_64-unknown-linux-gnu
Thread model: posix
InstalledDir: /sppo/dev/compilers/aocc-compiler-2.0.0/bin
------------------------------------------------------------------------------

==============================================================================
C++, C, Fortran | 507.cactuBSSN_r(base, peak)
------------------------------------------------------------------------------
AOCC.LLVM.2.0.0.B191.2019_07_19 clang version 8.0.0 (CLANG: Jenkins
  AOCC_2_0_0-Build#191) (based on LLVM AOCC.LLVM.2.0.0.B191.2019_07_19)
Target: x86_64-unknown-linux-gnu
Thread model: posix
InstalledDir: /sppo/dev/compilers/aocc-compiler-2.0.0/bin
AOCC.LLVM.2.0.0.B191.2019_07_19 clang version 8.0.0 (CLANG: Jenkins
  AOCC_2_0_0-Build#191) (based on LLVM AOCC.LLVM.2.0.0.B191.2019_07_19)
Target: x86_64-unknown-linux-gnu
Thread model: posix
InstalledDir: /sppo/dev/compilers/aocc-compiler-2.0.0/bin
AOCC.LLVM.2.0.0.B191.2019_07_19 clang version 8.0.0 (CLANG: Jenkins
  AOCC_2_0_0-Build#191) (based on LLVM AOCC.LLVM.2.0.0.B191.2019_07_19)
Target: x86_64-unknown-linux-gnu
Thread model: posix
InstalledDir: /sppo/dev/compilers/aocc-compiler-2.0.0/bin
------------------------------------------------------------------------------

==============================================================================
Fortran         | 503.bwaves_r(base, peak) 549.fotonik3d_r(base, peak)
                | 554.roms_r(base, peak)
------------------------------------------------------------------------------
AOCC.LLVM.2.0.0.B191.2019_07_19 clang version 8.0.0 (CLANG: Jenkins
  AOCC_2_0_0-Build#191) (based on LLVM AOCC.LLVM.2.0.0.B191.2019_07_19)
Target: x86_64-unknown-linux-gnu
Thread model: posix
InstalledDir: /sppo/dev/compilers/aocc-compiler-2.0.0/bin
------------------------------------------------------------------------------

==============================================================================
Fortran, C      | 521.wrf_r(base, peak) 527.cam4_r(base, peak)
------------------------------------------------------------------------------
AOCC.LLVM.2.0.0.B191.2019_07_19 clang version 8.0.0 (CLANG: Jenkins
  AOCC_2_0_0-Build#191) (based on LLVM AOCC.LLVM.2.0.0.B191.2019_07_19)
Target: x86_64-unknown-linux-gnu
Thread model: posix
InstalledDir: /sppo/dev/compilers/aocc-compiler-2.0.0/bin
AOCC.LLVM.2.0.0.B191.2019_07_19 clang version 8.0.0 (CLANG: Jenkins
  AOCC_2_0_0-Build#191) (based on LLVM AOCC.LLVM.2.0.0.B191.2019_07_19)
Target: x86_64-unknown-linux-gnu
Thread model: posix
InstalledDir: /sppo/dev/compilers/aocc-compiler-2.0.0/bin
------------------------------------------------------------------------------

Base Compiler Invocation

C benchmarks:

clang

C++ benchmarks:

clang++

Fortran benchmarks:

flang

Benchmarks using both Fortran and C:

flang clang

Benchmarks using both C and C++:

clang++ clang

Benchmarks using Fortran, C, and C++:

clang++ clang flang

Base Portability Flags

503.bwaves_r:	-DSPEC_LP64
507.cactuBSSN_r:	-DSPEC_LP64
508.namd_r:	-DSPEC_LP64
510.parest_r:	-DSPEC_LP64
511.povray_r:	-DSPEC_LP64
519.lbm_r:	-DSPEC_LP64
521.wrf_r:	-DSPEC_CASE_FLAG -Mbyteswapio -DSPEC_LP64
526.blender_r:	-funsigned-char -D__BOOL_DEFINED -DSPEC_LP64
527.cam4_r:	-DSPEC_CASE_FLAG -DSPEC_LP64
538.imagick_r:	-DSPEC_LP64
544.nab_r:	-DSPEC_LP64
549.fotonik3d_r:	-DSPEC_LP64
554.roms_r:	-DSPEC_LP64

Base Optimization Flags

C benchmarks:

C++ benchmarks:

Fortran benchmarks:

-flto -Wl,-mllvm -Wl,-function-specialize -Wl,-mllvm -Wl,-region-vectorize -Wl,-mllvm -Wl,-vector-library=LIBMVEC -Wl,-mllvm -Wl,-reduce-array-computations=3 -O3 -march=znver2 -funroll-loops -Mrecursive -mllvm -vector-library=LIBMVEC -z muldefs -Kieee -fno-finite-math-only -lmvec -lamdlibm -ljemalloc -lflang

Benchmarks using both Fortran and C:

Benchmarks using both C and C++:

-std=c++98 -flto -Wl,-mllvm -Wl,-function-specialize -Wl,-mllvm -Wl,-region-vectorize -Wl,-mllvm -Wl,-vector-library=LIBMVEC -Wl,-mllvm -Wl,-reduce-array-computations=3 -Wl,-mllvm -Wl,-suppress-fmas -O3 -ffast-math -march=znver2 -fstruct-layout=3 -mllvm -unroll-threshold=50 -fremap-arrays -mllvm -function-specialize -mllvm -enable-gvn-hoist -mllvm -reduce-array-computations=3 -mllvm -global-vectorize-slp -mllvm -vector-library=LIBMVEC -mllvm -inline-threshold=1000 -flv-function-specialization -mllvm -loop-unswitch-threshold=200000 -mllvm -unroll-threshold=100 -mllvm -enable-partial-unswitch -z muldefs -lmvec -lamdlibm -ljemalloc -lflang

Benchmarks using Fortran, C, and C++:

Peak Compiler Invocation

C benchmarks:

clang

C++ benchmarks:

clang++

Fortran benchmarks:

flang

Benchmarks using both Fortran and C:

flang clang

Benchmarks using both C and C++:

clang++ clang

Benchmarks using Fortran, C, and C++:

clang++ clang flang

Peak Portability Flags

Same as Base Portability Flags

Peak Optimization Flags

C benchmarks:

519.lbm_r:	basepeak = yes
538.imagick_r:	-flto -Wl,-mllvm -Wl,-function-specialize -Wl,-mllvm -Wl,-region-vectorize -Wl,-mllvm -Wl,-vector-library=LIBMVEC -Wl,-mllvm -Wl,-reduce-array-computations=3 -Ofast -march=znver2 -mno-sse4a -fstruct-layout=5 -mllvm -vectorize-memory-aggressively -mllvm -function-specialize -mllvm -enable-gvn-hoist -mllvm -unroll-threshold=50 -fremap-arrays -mllvm -vector-library=LIBMVEC -mllvm -reduce-array-computations=3 -mllvm -global-vectorize-slp -mllvm -inline-threshold=1000 -flv-function-specialization -lmvec -lamdlibm -ljemalloc -lflang
544.nab_r:	basepeak = yes

C++ benchmarks:

508.namd_r:	basepeak = yes
510.parest_r:	basepeak = yes

Fortran benchmarks:

503.bwaves_r:	basepeak = yes
549.fotonik3d_r:	basepeak = yes
554.roms_r:	-flto -Wl,-mllvm -Wl,-function-specialize -Wl,-mllvm -Wl,-region-vectorize -Wl,-mllvm -Wl,-vector-library=LIBMVEC -Wl,-mllvm -Wl,-reduce-array-computations=3 -Wl,-mllvm -Wl,-enable-X86-prefetching -O3 -march=znver2 -funroll-loops -Mrecursive -mllvm -vector-library=LIBMVEC -Kieee -fno-finite-math-only -lmvec -lamdlibm -ljemalloc -lflang

Benchmarks using both Fortran and C:

521.wrf_r:

basepeak = yes

527.cam4_r:

-flto -Wl,-mllvm -Wl,-function-specialize -Wl,-mllvm -Wl,-region-vectorize -Wl,-mllvm -Wl,-vector-library=LIBMVEC -Wl,-mllvm -Wl,-reduce-array-computations=3 -Ofast -march=znver2 -mno-sse4a -fstruct-layout=5 -mllvm -vectorize-memory-aggressively -mllvm -function-specialize -mllvm -enable-gvn-hoist -mllvm -unroll-threshold=50 -fremap-arrays -mllvm -vector-library=LIBMVEC -mllvm -reduce-array-computations=3 -mllvm -global-vectorize-slp -mllvm -inline-threshold=1000 -flv-function-specialization -O3 -funroll-loops -Mrecursive -Kieee -fno-finite-math-only -lmvec -lamdlibm -ljemalloc -lflang

Benchmarks using both C and C++:

511.povray_r:

-std=c++98 -flto -Wl,-mllvm -Wl,-function-specialize -Wl,-mllvm -Wl,-region-vectorize -Wl,-mllvm -Wl,-vector-library=LIBMVEC -Wl,-mllvm -Wl,-reduce-array-computations=3 -Wl,-mllvm -Wl,-x86-use-vzeroupper=false -Ofast -march=znver2 -mno-sse4a -fstruct-layout=5 -mllvm -vectorize-memory-aggressively -mllvm -function-specialize -mllvm -enable-gvn-hoist -mllvm -unroll-threshold=50 -fremap-arrays -mllvm -vector-library=LIBMVEC -mllvm -reduce-array-computations=3 -mllvm -global-vectorize-slp -mllvm -inline-threshold=1000 -flv-function-specialization -mllvm -unroll-threshold=100 -mllvm -enable-partial-unswitch -mllvm -loop-unswitch-threshold=200000 -lmvec -lamdlibm -ljemalloc -lflang

526.blender_r:

basepeak = yes

Benchmarks using Fortran, C, and C++:

-std=c++98 -flto -Wl,-mllvm -Wl,-function-specialize -Wl,-mllvm -Wl,-region-vectorize -Wl,-mllvm -Wl,-vector-library=LIBMVEC -Wl,-mllvm -Wl,-reduce-array-computations=3 -Ofast -march=znver2 -mno-sse4a -fstruct-layout=5 -mllvm -vectorize-memory-aggressively -mllvm -function-specialize -mllvm -enable-gvn-hoist -mllvm -unroll-threshold=50 -fremap-arrays -mllvm -vector-library=LIBMVEC -mllvm -reduce-array-computations=3 -mllvm -global-vectorize-slp -mllvm -inline-threshold=1000 -flv-function-specialization -mllvm -unroll-threshold=100 -mllvm -enable-partial-unswitch -mllvm -loop-unswitch-threshold=200000 -O3 -funroll-loops -Mrecursive -Kieee -fno-finite-math-only -lmvec -lamdlibm -ljemalloc -lflang

The flags files that were used to format this result can be browsed at
http://www.spec.org/cpu2017/flags/aocc200-flags-B1-1.html,
http://www.spec.org/cpu2017/flags/Lenovo-Platform-SPECcpu2017-Flags-V1.2-Rome2P-K.html.

You can also download the XML flags sources by saving the following links:
http://www.spec.org/cpu2017/flags/aocc200-flags-B1-1.xml,
http://www.spec.org/cpu2017/flags/Lenovo-Platform-SPECcpu2017-Flags-V1.2-Rome2P-K.xml.