OMP2012 Result Flag Description

Base Optimization Flags

C benchmarks

- -O3
- COPTIMIZE
- Enables O2 optimizations plus more aggressive optimizations, such as prefetching, scalar replacement, and loop and memory access transformations. Enables optimizations for maximum speed, such as:
  - Loop unrolling, including instruction scheduling
  - Code replication to eliminate branches
  - Padding the size of certain power-of-two arrays to allow more efficient cache use.
  On IA-32 and Intel EM64T processors, when O3 is used with options -ax or -x (Linux) or with options /Qax or /Qx (Windows), the compiler performs more aggressive data dependency analysis than for O2, which may result in longer compilation times.
  The O3 optimizations may not cause higher performance unless loop and memory access transformations take place. The optimizations may slow down code in some cases compared to O2 optimizations.
  The O3 option is recommended for applications that have loops that heavily use floating-point calculations and process large data sets. On IA-32 Windows platforms, -O3 sets the following:
  
  /GF (/Qvc7 and above), /Gf (/Qvc6 and below), and /Ob2
- Includes:
  - -GF
  - -Gf
  - -Ob_n
  - -O2
    - -Oi-
    - -Gs
    - -Oy
    - -Gy
    - -Os
    - -GF
    - -Gf
    - -Ob_n
    - -Og
    - -O1
      
      -unroll_n
      
      -Oi-
      
      -Op-
      
      -Oy
      
      -Gy
      
      -Os
      
      -GF
      
      -Gf
      
      -Ob_n
      
      -Og
- -ansi-alias
- COPTIMIZE
- Enable/disable(DEFAULT) use of ANSI aliasing rules in optimizations; user asserts that the program adheres to these rules.
- -no-prec-div
- COPTIMIZE
- (disable/enable[default] -[no-]prec-div)
  -prec-div improves precision of floating-point divides. It has a slight impact on speed. -no-prec-div disables this option and enables optimizations that give slightly less precise results than full IEEE division.
- -qopenmp
- COPTIMIZE
- Enables the parallelizer to generate multi-threaded code based on OpenMP* directives. Option -qopenmp is the replacement option for -openmp, which is deprecated.
- -ipo
- COPTIMIZE
- Multi-file ip optimizations that includes:
  - inline function expansion
  - interprocedural constant propogation
  - dead code elimination
  - propagation of function characteristics
  - passing arguments in registers
  - loop-invariant code motion
- -xMIC-AVX512
- COPTIMIZE
- May generate Intel(R) Advanced Vector Extensions 512 (Intel(R) AVX-512) Foundation instructions, Intel(R) AVX-512 Conflict Detection instructions, Intel(R) AVX-512 Exponential and Reciprocal instructions, Intel(R) AVX-512 Prefetch instructions for Intel(R) processors, and the instructions enabled with CORE-AVX2.
- -fp-model fast=2
- COPTIMIZE
- enable floating point model variation [no-]except - enable/disable floating point semantics fast[=1|2] - enables more aggressive floating point optimizations precise - allows value-safe optimizations source - enables intermediates in source precision strict - enables -fp-model precise -fp-model except, disables contractions and enables pragma stdc fenv_access double - rounds intermediates in 53-bit (double) precision extended - rounds intermediates in 64-bit (extended) precision

C++ benchmarks

- -O3
- CXXOPTIMIZE
- Enables O2 optimizations plus more aggressive optimizations, such as prefetching, scalar replacement, and loop and memory access transformations. Enables optimizations for maximum speed, such as:
  - Loop unrolling, including instruction scheduling
  - Code replication to eliminate branches
  - Padding the size of certain power-of-two arrays to allow more efficient cache use.
  On IA-32 and Intel EM64T processors, when O3 is used with options -ax or -x (Linux) or with options /Qax or /Qx (Windows), the compiler performs more aggressive data dependency analysis than for O2, which may result in longer compilation times.
  The O3 optimizations may not cause higher performance unless loop and memory access transformations take place. The optimizations may slow down code in some cases compared to O2 optimizations.
  The O3 option is recommended for applications that have loops that heavily use floating-point calculations and process large data sets. On IA-32 Windows platforms, -O3 sets the following:
  
  /GF (/Qvc7 and above), /Gf (/Qvc6 and below), and /Ob2
- Includes:
  - -GF
  - -Gf
  - -Ob_n
  - -O2
    - -Oi-
    - -Gs
    - -Oy
    - -Gy
    - -Os
    - -GF
    - -Gf
    - -Ob_n
    - -Og
    - -O1
      
      -unroll_n
      
      -Oi-
      
      -Op-
      
      -Oy
      
      -Gy
      
      -Os
      
      -GF
      
      -Gf
      
      -Ob_n
      
      -Og
- -ansi-alias
- CXXOPTIMIZE
- Enable/disable(DEFAULT) use of ANSI aliasing rules in optimizations; user asserts that the program adheres to these rules.
- -no-prec-div
- CXXOPTIMIZE
- (disable/enable[default] -[no-]prec-div)
  -prec-div improves precision of floating-point divides. It has a slight impact on speed. -no-prec-div disables this option and enables optimizations that give slightly less precise results than full IEEE division.
- -qopenmp
- CXXOPTIMIZE
- Enables the parallelizer to generate multi-threaded code based on OpenMP* directives. Option -qopenmp is the replacement option for -openmp, which is deprecated.
- -ipo
- CXXOPTIMIZE
- Multi-file ip optimizations that includes:
  - inline function expansion
  - interprocedural constant propogation
  - dead code elimination
  - propagation of function characteristics
  - passing arguments in registers
  - loop-invariant code motion
- -xMIC-AVX512
- CXXOPTIMIZE
- May generate Intel(R) Advanced Vector Extensions 512 (Intel(R) AVX-512) Foundation instructions, Intel(R) AVX-512 Conflict Detection instructions, Intel(R) AVX-512 Exponential and Reciprocal instructions, Intel(R) AVX-512 Prefetch instructions for Intel(R) processors, and the instructions enabled with CORE-AVX2.
- -fp-model fast=2
- CXXOPTIMIZE
- enable floating point model variation [no-]except - enable/disable floating point semantics fast[=1|2] - enables more aggressive floating point optimizations precise - allows value-safe optimizations source - enables intermediates in source precision strict - enables -fp-model precise -fp-model except, disables contractions and enables pragma stdc fenv_access double - rounds intermediates in 53-bit (double) precision extended - rounds intermediates in 64-bit (extended) precision

Fortran benchmarks

- -O3
- FOPTIMIZE
- Enables O2 optimizations plus more aggressive optimizations, such as prefetching, scalar replacement, and loop and memory access transformations. Enables optimizations for maximum speed, such as:
  - Loop unrolling, including instruction scheduling
  - Code replication to eliminate branches
  - Padding the size of certain power-of-two arrays to allow more efficient cache use.
  On IA-32 and Intel EM64T processors, when O3 is used with options -ax or -x (Linux) or with options /Qax or /Qx (Windows), the compiler performs more aggressive data dependency analysis than for O2, which may result in longer compilation times.
  The O3 optimizations may not cause higher performance unless loop and memory access transformations take place. The optimizations may slow down code in some cases compared to O2 optimizations.
  The O3 option is recommended for applications that have loops that heavily use floating-point calculations and process large data sets. On IA-32 Windows platforms, -O3 sets the following:
  
  /GF (/Qvc7 and above), /Gf (/Qvc6 and below), and /Ob2
- Includes:
  - -GF
  - -Gf
  - -Ob_n
  - -O2
    - -Oi-
    - -Gs
    - -Oy
    - -Gy
    - -Os
    - -GF
    - -Gf
    - -Ob_n
    - -Og
    - -O1
      
      -unroll_n
      
      -Oi-
      
      -Op-
      
      -Oy
      
      -Gy
      
      -Os
      
      -GF
      
      -Gf
      
      -Ob_n
      
      -Og
- -no-prec-div
- FOPTIMIZE
- (disable/enable[default] -[no-]prec-div)
  -prec-div improves precision of floating-point divides. It has a slight impact on speed. -no-prec-div disables this option and enables optimizations that give slightly less precise results than full IEEE division.
- -qopenmp
- FOPTIMIZE
- Enables the parallelizer to generate multi-threaded code based on OpenMP* directives. Option -qopenmp is the replacement option for -openmp, which is deprecated.
- -ipo
- FOPTIMIZE
- Multi-file ip optimizations that includes:
  - inline function expansion
  - interprocedural constant propogation
  - dead code elimination
  - propagation of function characteristics
  - passing arguments in registers
  - loop-invariant code motion
- -xMIC-AVX512
- FOPTIMIZE
- May generate Intel(R) Advanced Vector Extensions 512 (Intel(R) AVX-512) Foundation instructions, Intel(R) AVX-512 Conflict Detection instructions, Intel(R) AVX-512 Exponential and Reciprocal instructions, Intel(R) AVX-512 Prefetch instructions for Intel(R) processors, and the instructions enabled with CORE-AVX2.
- -fp-model fast=2
- FOPTIMIZE
- enable floating point model variation [no-]except - enable/disable floating point semantics fast[=1|2] - enables more aggressive floating point optimizations precise - allows value-safe optimizations source - enables intermediates in source precision strict - enables -fp-model precise -fp-model except, disables contractions and enables pragma stdc fenv_access double - rounds intermediates in 53-bit (double) precision extended - rounds intermediates in 64-bit (extended) precision

Implicitly Included Flags

This section contains descriptions of flags that were included implicitly by other flags, but which do not have a permanent home at SPEC.

Shell, Environment, and Other Software Settings

This result has been formatted using multiple flags files. The "sw environment" from each of them appears next.

Sw environment from Intel-ic16.0-linux64

SPEC OMP2012 Flag Description for the Intel(R) C/C++ Compiler for IA32 and Intel 64 applications and Intel(R) Fortran Compiler for IA32 and Intel 64 applications

Sw environment from colfax-knl

SPEC OMP2012 Platform Description

Only allocate memory from nodes. Allocation will fail when there is not enough memory available on these nodes. Nodes may be specified as noted above

Preferably allocate memory on node, but if memory cannot be allocated then fall back to other nodes. This option takes only a single node number. Relative notation may be used.

Firmware / BIOS / Microcode Settings

For questions about the meanings of these flags, please contact the tester.
For other inquiries, please contact webmaster@spec.org
Copyright 2012-2017 Standard Performance Evaluation Corporation
Tested with SPEC OMP2012 v1.1.
Report generated on Wed Jan 11 12:09:28 2017 by SPEC OMP2012 flags formatter v538.

OMP2012 Flag Description
Colfax International Intel Xeon Phi 7210, 1.30GHz, SMT off, Turbo off, flat DRAM+MCDRAM

Test sponsored by Indiana University

Base Compiler Invocation

C benchmarks

C++ benchmarks

Fortran benchmarks

Base Portability Flags

350.md

357.bt331

363.swim

367.imagick

Base Optimization Flags

C benchmarks

C++ benchmarks

Fortran benchmarks

Implicitly Included Flags

Shell, Environment, and Other Software Settings

Sw environment from Intel-ic16.0-linux64

SPEC OMP2012 Flag Description for the Intel(R) C/C++ Compiler for IA32 and Intel 64 applications and Intel(R) Fortran Compiler for IA32 and Intel 64 applications

Sw environment from colfax-knl

SPEC OMP2012 Platform Description

Firmware / BIOS / Microcode Settings

	Indicates that the flag description came from the user flags file.
	Indicates that the flag description came from the suite-wide flags file.
	Indicates that the flag description came from a per-benchmark flags file.

OMP2012 Flag DescriptionColfax International Intel Xeon Phi 7210, 1.30GHz, SMT off, Turbo off, flat DRAM+MCDRAM

Test sponsored by Indiana University

Base Compiler Invocation

Base Portability Flags

Base Optimization Flags

Implicitly Included Flags

Sw environment from Intel-ic16.0-linux64

SPEC OMP2012 Flag Description for the Intel(R) C/C++ Compiler for IA32 and Intel 64 applications and Intel(R) Fortran Compiler for IA32 and Intel 64 applications

Sw environment from colfax-knl

SPEC OMP2012 Platform Description

OMP2012 Flag Description
Colfax International Intel Xeon Phi 7210, 1.30GHz, SMT off, Turbo off, flat DRAM+MCDRAM