CPU2006 Result Flag Description

Base Portability Flags

400.perlbench

- -DSPEC_CPU_LP64
- CPORTABILITY
- This macro specifies that the target system uses the LP64 data model; specifically, that integers are 32 bits, while longs and pointers are 64 bits.
- Includes:
- -DSPEC_CPU_LINUX_X64
- CPORTABILITY
- This macro indicates that the benchmark is being compiled on an AMD64-compatible system running the Linux operating system.
- Includes:

401.bzip2

- -DSPEC_CPU_LP64
- CPORTABILITY
- This option is used to indicate that the host system's integers are 32-bits wide, and longs and pointers are 64-bits wide. Not all benchmarks recognize this macro, but the preferred practice for data model selection applies the flags to all benchmarks; this flag description is a placeholder for those benchmarks that do not recognize this macro.

403.gcc

- -DSPEC_CPU_HAVE_BOOL
- CPORTABILITY
- SPEC_CPU_HAVE_BOOL allows the compiler to indicate that it has the bool type.
- -DSPEC_CPU_LP64
- CPORTABILITY
- This option is used to indicate that the host system's integers are 32-bits wide, and longs and pointers are 64-bits wide. Not all benchmarks recognize this macro, but the preferred practice for data model selection applies the flags to all benchmarks; this flag description is a placeholder for those benchmarks that do not recognize this macro.

429.mcf

- -DSPEC_CPU_LP64
- CPORTABILITY
- This option is used to indicate that the host system's integers are 32-bits wide, and longs and pointers are 64-bits wide. Not all benchmarks recognize this macro, but the preferred practice for data model selection applies the flags to all benchmarks; this flag description is a placeholder for those benchmarks that do not recognize this macro.

445.gobmk

- -DSPEC_CPU_LP64
- CPORTABILITY
- This option is used to indicate that the host system's integers are 32-bits wide, and longs and pointers are 64-bits wide. Not all benchmarks recognize this macro, but the preferred practice for data model selection applies the flags to all benchmarks; this flag description is a placeholder for those benchmarks that do not recognize this macro.

456.hmmer

- -DSPEC_CPU_LP64
- CPORTABILITY
- This option is used to indicate that the host system's integers are 32-bits wide, and longs and pointers are 64-bits wide. Not all benchmarks recognize this macro, but the preferred practice for data model selection applies the flags to all benchmarks; this flag description is a placeholder for those benchmarks that do not recognize this macro.

458.sjeng

- -DSPEC_CPU_LP64
- CPORTABILITY
- This option is used to indicate that the host system's integers are 32-bits wide, and longs and pointers are 64-bits wide. Not all benchmarks recognize this macro, but the preferred practice for data model selection applies the flags to all benchmarks; this flag description is a placeholder for those benchmarks that do not recognize this macro.

462.libquantum

- -DSPEC_CPU_LP64
- CPORTABILITY
- This option is used to indicate that the host system's integers are 32-bits wide, and longs and pointers are 64-bits wide. Not all benchmarks recognize this macro, but the preferred practice for data model selection applies the flags to all benchmarks; this flag description is a placeholder for those benchmarks that do not recognize this macro.
- -DSPEC_CPU_LINUX
- CPORTABILITY
- Portability changes for Linux

464.h264ref

- -DSPEC_CPU_LP64
- CPORTABILITY
- This option is used to indicate that the host system's integers are 32-bits wide, and longs and pointers are 64-bits wide. Not all benchmarks recognize this macro, but the preferred practice for data model selection applies the flags to all benchmarks; this flag description is a placeholder for those benchmarks that do not recognize this macro.

483.xalancbmk

- -DSPEC_CPU_LINUX
- CXXPORTABILITY
- This flag can be set for SPEC compilation for Linux using default compiler.

Base Optimization Flags

C benchmarks

- -fast
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Chooses generally optimal flags for the target platform. As of the PGI 7.0 release, the flags "-fast" and "-fastsse" are equivlent for 64-bit compilations. For 32-bit compilations, "-fastsse" uses the same optimizations as 64-bits. However "-fast" does not include "-Mscalasse", "-Mcache_align", or "-Mvect=sse"
- Includes:
  - -O2
    - -O1
  - -Munroll=c:1
  - -Mautoinline
  - -Msmart
  - -Mlre
  - -Mnoframe
  - -Mvect=sse
    - -Mvect
      
      -Mvect=assoc
      
      -Mvect=altcode
  - -Mcache_align
  - -Mflushz
  - -Mdaz
  - -Mscalarsse
- -Mipa=fast
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Instructs the compiler to perform interprocedural analysis. Equivalant to -Mipa=align,arg,const,f90ptr,shape,globals,libc,localarg,ptr,pure.
- Includes:
- -Mipa=inline
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Interprocedural Analysis option: Automatically determine which functions to inline. IPA-based function inlining is performed from leaf routines upward.
- -Mipa=noarg
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Interprocedural Analysis option: Do not remove arguments replaced by -Mipa=ptr,const
- Includes:
  - -Mipa=ptr
  - -Mipa=const
- -Mfprelaxed
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Instructs the compiler to use relaxed precision in the calculation of some intrinsic functions. Can result in improved performance at the expense of numerical accuracy. The default on an AMD system is "-Mfprelaxed=rsqrt,order". The default on an Intel system is "-Mfprelaxed=rsqrt,sqrt,div,order"
- Includes:
- -Msmartalloc=huge:64
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Link with the PGI huge page runtime library and allocate a maximum n huge pages.
- -tp k8-64
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Specify the type of the target processor as AMD64 Processor 64-bit mode.
- -Bstatic_pgi
- pgcc, pgcpp, pgf95
- EXTRA_LDFLAGS
- Staticily link with the PGI runtime libraries. System libraries may still be dynamically linked.

C++ benchmarks

- -fastsse
- pgcc, pgcpp, pgf95
- CXXOPTIMIZE
- Chooses generally optimal flags for the target platform. As of the PGI 7.0 release, the flags "-fast" and "-fastsse" are equivlent for 64-bit compilations. For 32-bit compilations, "-fastsse" uses the same optimizations as 64-bits. However "-fast" does not include "-Mscalasse", "-Mcache_align", or "-Mvect=sse"
- Includes:
  - -O2
    - -O1
  - -Munroll=c:1
  - -Mautoinline
  - -Msmart
  - -Mlre
  - -Mnoframe
  - -Mvect=sse
    - -Mvect
      
      -Mvect=assoc
      
      -Mvect=altcode
  - -Mcache_align
  - -Mflushz
  - -Mdaz
  - -Mscalarsse
- -Mipa=fast
- pgcc, pgcpp, pgf95
- CXXOPTIMIZE
- Instructs the compiler to perform interprocedural analysis. Equivalant to -Mipa=align,arg,const,f90ptr,shape,globals,libc,localarg,ptr,pure.
- Includes:
- -Mipa=inline
- pgcc, pgcpp, pgf95
- CXXOPTIMIZE
- Interprocedural Analysis option: Automatically determine which functions to inline. IPA-based function inlining is performed from leaf routines upward.
- -Mfprelaxed
- pgcc, pgcpp, pgf95
- CXXOPTIMIZE
- Instructs the compiler to use relaxed precision in the calculation of some intrinsic functions. Can result in improved performance at the expense of numerical accuracy. The default on an AMD system is "-Mfprelaxed=rsqrt,order". The default on an Intel system is "-Mfprelaxed=rsqrt,sqrt,div,order"
- Includes:
- -Msmartalloc
- pgcc, pgcpp, pgf95
- CXXOPTIMIZE
- Adds a call to the routine "mallopt" in the main routine. This option can have a dramatic impact on the performance of programs that dynamically allocate memory, especially for those which have a few large mallocs. To be effective, this switch must be specified when compiling the file containing the Fortran, C, or C++ main routine.
- --zc_eh
- pgcpp
- CXXOPTIMIZE
- Generate zero-cost C++ exception handlers.
- -tp k8-32
- pgcc, pgcpp, pgf95
- CXXOPTIMIZE
- Specify the type of the target processor as AMD64 Processor 32-bit mode.
- -Bstatic_pgi
- pgcc, pgcpp, pgf95
- EXTRA_LDFLAGS
- Staticily link with the PGI runtime libraries. System libraries may still be dynamically linked.
- -L/proj/qa/smartheap/SmartHeap_8/lib -lsmartheap
- pgcc, pgcpp, pgf95
- EXTRA_CXXLIBS
- Link using MicroQuill's SmartHeap 8 (32-bit) library for Linux. Description from Microquill:
  
  SmartHeap is a fast (3X-100X faster than compiler-supplied libraries), portable (Windows, Linux, Solaris, HP-UX, IBM-AIX, Dec OSF Tru64, SGI Irix), reliable, ANSI-compliant malloc/operator new library. SmartHeap supports multiple memory pools, includes a fixed-size allocator, and is thread-safe. SmartHeap also includes comprehensive memory debugging APIs to detect leakage, overwrites, double-frees, wild pointers, out of memory, references to previously freed memory, and other memory errors.

Peak Optimization Flags

C benchmarks

400.perlbench

- -Mpfi
- pgcc, pgcpp, pgf95
- PASS1_CFLAGS, PASS1_LDFLAGS
- Generate profile-feedback instrumentation (PFI); this includes extra code to collect run-time statistics and dump them to a trace file for use in a subsequent compilation. PFI gathers information about a program's execution and data values but does not gather information from hardware performance counters. PFI does gather data for optimizations which are unique to profile-feedback optimization.
- -Mpfo
- pgcc, pgcpp, pgf95
- PASS2_CFLAGS, PASS2_LDFLAGS
- Enable profile-feedback optimizations.
- -Mipa=inline
- pgcc, pgcpp, pgf95
- PASS2_CFLAGS, PASS2_LDFLAGS
- Interprocedural Analysis option: Automatically determine which functions to inline. IPA-based function inlining is performed from leaf routines upward.
- -fast
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Chooses generally optimal flags for the target platform. As of the PGI 7.0 release, the flags "-fast" and "-fastsse" are equivlent for 64-bit compilations. For 32-bit compilations, "-fastsse" uses the same optimizations as 64-bits. However "-fast" does not include "-Mscalasse", "-Mcache_align", or "-Mvect=sse"
- Includes:
  - -O2
    - -O1
  - -Munroll=c:1
  - -Mautoinline
  - -Msmart
  - -Mlre
  - -Mnoframe
  - -Mvect=sse
    - -Mvect
      
      -Mvect=assoc
      
      -Mvect=altcode
  - -Mcache_align
  - -Mflushz
  - -Mdaz
  - -Mscalarsse
- -O4
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Performs all level 1, 2, and 3 optimizations and enables hoisting of guarded invariant floating point expressions.
- Includes:
  - -O3
    - -O2
      
      -O1
- -Mfprelaxed
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Instructs the compiler to use relaxed precision in the calculation of some intrinsic functions. Can result in improved performance at the expense of numerical accuracy. The default on an AMD system is "-Mfprelaxed=rsqrt,order". The default on an Intel system is "-Mfprelaxed=rsqrt,sqrt,div,order"
- Includes:
- -Msmartalloc=huge:8
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Link with the PGI huge page runtime library and allocate a maximum n huge pages.
- -Mnounroll
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Disable loop unrolling.
- -tp k8-64
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Specify the type of the target processor as AMD64 Processor 64-bit mode.
- -Bstatic_pgi
- pgcc, pgcpp, pgf95
- EXTRA_LDFLAGS
- Staticily link with the PGI runtime libraries. System libraries may still be dynamically linked.

401.bzip2

- -Mpfi
- pgcc, pgcpp, pgf95
- PASS1_CFLAGS, PASS1_LDFLAGS
- Generate profile-feedback instrumentation (PFI); this includes extra code to collect run-time statistics and dump them to a trace file for use in a subsequent compilation. PFI gathers information about a program's execution and data values but does not gather information from hardware performance counters. PFI does gather data for optimizations which are unique to profile-feedback optimization.
- -Mpfo
- pgcc, pgcpp, pgf95
- PASS2_CFLAGS, PASS2_LDFLAGS
- Enable profile-feedback optimizations.
- -fast
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Chooses generally optimal flags for the target platform. As of the PGI 7.0 release, the flags "-fast" and "-fastsse" are equivlent for 64-bit compilations. For 32-bit compilations, "-fastsse" uses the same optimizations as 64-bits. However "-fast" does not include "-Mscalasse", "-Mcache_align", or "-Mvect=sse"
- Includes:
  - -O2
    - -O1
  - -Munroll=c:1
  - -Mautoinline
  - -Msmart
  - -Mlre
  - -Mnoframe
  - -Mvect=sse
    - -Mvect
      
      -Mvect=assoc
      
      -Mvect=altcode
  - -Mcache_align
  - -Mflushz
  - -Mdaz
  - -Mscalarsse
- -O4
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Performs all level 1, 2, and 3 optimizations and enables hoisting of guarded invariant floating point expressions.
- Includes:
  - -O3
    - -O2
      
      -O1
- -Msmartalloc=huge:64
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Link with the PGI huge page runtime library and allocate a maximum n huge pages.
- -tp k8-64
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Specify the type of the target processor as AMD64 Processor 64-bit mode.
- -Bstatic_pgi
- pgcc, pgcpp, pgf95
- EXTRA_LDFLAGS
- Staticily link with the PGI runtime libraries. System libraries may still be dynamically linked.

403.gcc

- -fastsse
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Chooses generally optimal flags for the target platform. As of the PGI 7.0 release, the flags "-fast" and "-fastsse" are equivlent for 64-bit compilations. For 32-bit compilations, "-fastsse" uses the same optimizations as 64-bits. However "-fast" does not include "-Mscalasse", "-Mcache_align", or "-Mvect=sse"
- Includes:
  - -O2
    - -O1
  - -Munroll=c:1
  - -Mautoinline
  - -Msmart
  - -Mlre
  - -Mnoframe
  - -Mvect=sse
    - -Mvect
      
      -Mvect=assoc
      
      -Mvect=altcode
  - -Mcache_align
  - -Mflushz
  - -Mdaz
  - -Mscalarsse
- -Mfprelaxed
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Instructs the compiler to use relaxed precision in the calculation of some intrinsic functions. Can result in improved performance at the expense of numerical accuracy. The default on an AMD system is "-Mfprelaxed=rsqrt,order". The default on an Intel system is "-Mfprelaxed=rsqrt,sqrt,div,order"
- Includes:
- -Msmartalloc=huge:32
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Link with the PGI huge page runtime library and allocate a maximum n huge pages.
- -Mipa=fast
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Instructs the compiler to perform interprocedural analysis. Equivalant to -Mipa=align,arg,const,f90ptr,shape,globals,libc,localarg,ptr,pure.
- Includes:
- -Mipa=inline
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Interprocedural Analysis option: Automatically determine which functions to inline. IPA-based function inlining is performed from leaf routines upward.
- -tp k8-32
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Specify the type of the target processor as AMD64 Processor 32-bit mode.
- -Bstatic_pgi
- pgcc, pgcpp, pgf95
- EXTRA_LDFLAGS
- Staticily link with the PGI runtime libraries. System libraries may still be dynamically linked.

429.mcf

- -fastsse
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Chooses generally optimal flags for the target platform. As of the PGI 7.0 release, the flags "-fast" and "-fastsse" are equivlent for 64-bit compilations. For 32-bit compilations, "-fastsse" uses the same optimizations as 64-bits. However "-fast" does not include "-Mscalasse", "-Mcache_align", or "-Mvect=sse"
- Includes:
  - -O2
    - -O1
  - -Munroll=c:1
  - -Mautoinline
  - -Msmart
  - -Mlre
  - -Mnoframe
  - -Mvect=sse
    - -Mvect
      
      -Mvect=assoc
      
      -Mvect=altcode
  - -Mcache_align
  - -Mflushz
  - -Mdaz
  - -Mscalarsse
- -Mipa=fast
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Instructs the compiler to perform interprocedural analysis. Equivalant to -Mipa=align,arg,const,f90ptr,shape,globals,libc,localarg,ptr,pure.
- Includes:
- -Mipa=inline
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Interprocedural Analysis option: Automatically determine which functions to inline. IPA-based function inlining is performed from leaf routines upward.
- -Msmartalloc=huge:16
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Link with the PGI huge page runtime library and allocate a maximum n huge pages.
- -tp k8-32
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Specify the type of the target processor as AMD64 Processor 32-bit mode.
- -Bstatic_pgi
- pgcc, pgcpp, pgf95
- EXTRA_LDFLAGS
- Staticily link with the PGI runtime libraries. System libraries may still be dynamically linked.

445.gobmk

- -Mpfi
- pgcc, pgcpp, pgf95
- PASS1_CFLAGS, PASS1_LDFLAGS
- Generate profile-feedback instrumentation (PFI); this includes extra code to collect run-time statistics and dump them to a trace file for use in a subsequent compilation. PFI gathers information about a program's execution and data values but does not gather information from hardware performance counters. PFI does gather data for optimizations which are unique to profile-feedback optimization.
- -Mpfo
- pgcc, pgcpp, pgf95
- PASS2_CFLAGS, PASS2_LDFLAGS
- Enable profile-feedback optimizations.
- -Mipa=fast
- pgcc, pgcpp, pgf95
- PASS2_CFLAGS, PASS2_LDFLAGS
- Instructs the compiler to perform interprocedural analysis. Equivalant to -Mipa=align,arg,const,f90ptr,shape,globals,libc,localarg,ptr,pure.
- Includes:
- -fast
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Chooses generally optimal flags for the target platform. As of the PGI 7.0 release, the flags "-fast" and "-fastsse" are equivlent for 64-bit compilations. For 32-bit compilations, "-fastsse" uses the same optimizations as 64-bits. However "-fast" does not include "-Mscalasse", "-Mcache_align", or "-Mvect=sse"
- Includes:
  - -O2
    - -O1
  - -Munroll=c:1
  - -Mautoinline
  - -Msmart
  - -Mlre
  - -Mnoframe
  - -Mvect=sse
    - -Mvect
      
      -Mvect=assoc
      
      -Mvect=altcode
  - -Mcache_align
  - -Mflushz
  - -Mdaz
  - -Mscalarsse
- -O4
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Performs all level 1, 2, and 3 optimizations and enables hoisting of guarded invariant floating point expressions.
- Includes:
  - -O3
    - -O2
      
      -O1
- -Msmartalloc=huge:32
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Link with the PGI huge page runtime library and allocate a maximum n huge pages.
- -Mfprelaxed
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Instructs the compiler to use relaxed precision in the calculation of some intrinsic functions. Can result in improved performance at the expense of numerical accuracy. The default on an AMD system is "-Mfprelaxed=rsqrt,order". The default on an Intel system is "-Mfprelaxed=rsqrt,sqrt,div,order"
- Includes:
- -Mnovect
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Disable automatic vector pipelining.
- -tp k8-64
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Specify the type of the target processor as AMD64 Processor 64-bit mode.
- -Bstatic_pgi
- pgcc, pgcpp, pgf95
- EXTRA_LDFLAGS
- Staticily link with the PGI runtime libraries. System libraries may still be dynamically linked.

456.hmmer

- -fast
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Chooses generally optimal flags for the target platform. As of the PGI 7.0 release, the flags "-fast" and "-fastsse" are equivlent for 64-bit compilations. For 32-bit compilations, "-fastsse" uses the same optimizations as 64-bits. However "-fast" does not include "-Mscalasse", "-Mcache_align", or "-Mvect=sse"
- Includes:
  - -O2
    - -O1
  - -Munroll=c:1
  - -Mautoinline
  - -Msmart
  - -Mlre
  - -Mnoframe
  - -Mvect=sse
    - -Mvect
      
      -Mvect=assoc
      
      -Mvect=altcode
  - -Mcache_align
  - -Mflushz
  - -Mdaz
  - -Mscalarsse
- -Msmartalloc=huge:32
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Link with the PGI huge page runtime library and allocate a maximum n huge pages.
- -Mfprelaxed
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Instructs the compiler to use relaxed precision in the calculation of some intrinsic functions. Can result in improved performance at the expense of numerical accuracy. The default on an AMD system is "-Mfprelaxed=rsqrt,order". The default on an Intel system is "-Mfprelaxed=rsqrt,sqrt,div,order"
- Includes:
- -Msafeptr
- pgcc, pgcpp
- COPTIMIZE
- Instructs the C/C++ compiler to override data dependencies between pointers of a given storage class.
- -Mipa=const
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Interprocedural Analysis option: Enable interprocedural constant propagation.
- -Mipa=ptr
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Interprocedural Analysis option: Enable pointer disambiguation across procedure calls.
- -Mipa=arg
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Interprocedural Analysis option: Remove arguments replaced by -Mipa=ptr,const
- Includes:
  - -Mipa=ptr
  - -Mipa=const
- -tp k8-64
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Specify the type of the target processor as AMD64 Processor 64-bit mode.
- -Bstatic_pgi
- pgcc, pgcpp, pgf95
- EXTRA_LDFLAGS
- Staticily link with the PGI runtime libraries. System libraries may still be dynamically linked.

458.sjeng

- -Mpfi
- pgcc, pgcpp, pgf95
- PASS1_CFLAGS, PASS1_LDFLAGS
- Generate profile-feedback instrumentation (PFI); this includes extra code to collect run-time statistics and dump them to a trace file for use in a subsequent compilation. PFI gathers information about a program's execution and data values but does not gather information from hardware performance counters. PFI does gather data for optimizations which are unique to profile-feedback optimization.
- -Mipa=fast
- pgcc, pgcpp, pgf95
- PASS2_CFLAGS, PASS2_LDFLAGS
- Instructs the compiler to perform interprocedural analysis. Equivalant to -Mipa=align,arg,const,f90ptr,shape,globals,libc,localarg,ptr,pure.
- Includes:
- -Mipa=inline
- pgcc, pgcpp, pgf95
- PASS2_CFLAGS, PASS2_LDFLAGS
- Interprocedural Analysis option: Automatically determine which functions to inline. IPA-based function inlining is performed from leaf routines upward.
- -Mipa=noarg
- pgcc, pgcpp, pgf95
- PASS2_CFLAGS, PASS2_LDFLAGS
- Interprocedural Analysis option: Do not remove arguments replaced by -Mipa=ptr,const
- Includes:
  - -Mipa=ptr
  - -Mipa=const
- -Mpfo
- pgcc, pgcpp, pgf95
- PASS2_CFLAGS, PASS2_LDFLAGS
- Enable profile-feedback optimizations.
- -fast
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Chooses generally optimal flags for the target platform. As of the PGI 7.0 release, the flags "-fast" and "-fastsse" are equivlent for 64-bit compilations. For 32-bit compilations, "-fastsse" uses the same optimizations as 64-bits. However "-fast" does not include "-Mscalasse", "-Mcache_align", or "-Mvect=sse"
- Includes:
  - -O2
    - -O1
  - -Munroll=c:1
  - -Mautoinline
  - -Msmart
  - -Mlre
  - -Mnoframe
  - -Mvect=sse
    - -Mvect
      
      -Mvect=assoc
      
      -Mvect=altcode
  - -Mcache_align
  - -Mflushz
  - -Mdaz
  - -Mscalarsse
- -Msmartalloc=huge:32
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Link with the PGI huge page runtime library and allocate a maximum n huge pages.
- -Mfprelaxed
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Instructs the compiler to use relaxed precision in the calculation of some intrinsic functions. Can result in improved performance at the expense of numerical accuracy. The default on an AMD system is "-Mfprelaxed=rsqrt,order". The default on an Intel system is "-Mfprelaxed=rsqrt,sqrt,div,order"
- Includes:
- -tp k8-64
- pgcc, pgcpp, pgf95
- COPTIMIZE
- Specify the type of the target processor as AMD64 Processor 64-bit mode.
- -Bstatic_pgi
- pgcc, pgcpp, pgf95
- EXTRA_LDFLAGS
- Staticily link with the PGI runtime libraries. System libraries may still be dynamically linked.

C++ benchmarks

483.xalancbmk

- -fastsse
- pgcc, pgcpp, pgf95
- CXXOPTIMIZE
- Chooses generally optimal flags for the target platform. As of the PGI 7.0 release, the flags "-fast" and "-fastsse" are equivlent for 64-bit compilations. For 32-bit compilations, "-fastsse" uses the same optimizations as 64-bits. However "-fast" does not include "-Mscalasse", "-Mcache_align", or "-Mvect=sse"
- Includes:
  - -O2
    - -O1
  - -Munroll=c:1
  - -Mautoinline
  - -Msmart
  - -Mlre
  - -Mnoframe
  - -Mvect=sse
    - -Mvect
      
      -Mvect=assoc
      
      -Mvect=altcode
  - -Mcache_align
  - -Mflushz
  - -Mdaz
  - -Mscalarsse
- -O4
- pgcc, pgcpp, pgf95
- CXXOPTIMIZE
- Performs all level 1, 2, and 3 optimizations and enables hoisting of guarded invariant floating point expressions.
- Includes:
  - -O3
    - -O2
      
      -O1
- -Mipa=fast
- pgcc, pgcpp, pgf95
- CXXOPTIMIZE
- Instructs the compiler to perform interprocedural analysis. Equivalant to -Mipa=align,arg,const,f90ptr,shape,globals,libc,localarg,ptr,pure.
- Includes:
- -Mipa=inline
- pgcc, pgcpp, pgf95
- CXXOPTIMIZE
- Interprocedural Analysis option: Automatically determine which functions to inline. IPA-based function inlining is performed from leaf routines upward.
- -Mfprelaxed
- pgcc, pgcpp, pgf95
- CXXOPTIMIZE
- Instructs the compiler to use relaxed precision in the calculation of some intrinsic functions. Can result in improved performance at the expense of numerical accuracy. The default on an AMD system is "-Mfprelaxed=rsqrt,order". The default on an Intel system is "-Mfprelaxed=rsqrt,sqrt,div,order"
- Includes:
- -Msmartalloc
- pgcc, pgcpp, pgf95
- CXXOPTIMIZE
- Adds a call to the routine "mallopt" in the main routine. This option can have a dramatic impact on the performance of programs that dynamically allocate memory, especially for those which have a few large mallocs. To be effective, this switch must be specified when compiling the file containing the Fortran, C, or C++ main routine.
- --zc_eh
- pgcpp
- CXXOPTIMIZE
- Generate zero-cost C++ exception handlers.
- -tp k8-32
- pgcc, pgcpp, pgf95
- CXXOPTIMIZE
- Specify the type of the target processor as AMD64 Processor 32-bit mode.
- -Bstatic_pgi
- pgcc, pgcpp, pgf95
- EXTRA_LDFLAGS
- Staticily link with the PGI runtime libraries. System libraries may still be dynamically linked.
- -L/proj/qa/smartheap/SmartHeap_8/lib -lsmartheap
- pgcc, pgcpp, pgf95
- EXTRA_CXXLIBS
- Link using MicroQuill's SmartHeap 8 (32-bit) library for Linux. Description from Microquill:
  
  SmartHeap is a fast (3X-100X faster than compiler-supplied libraries), portable (Windows, Linux, Solaris, HP-UX, IBM-AIX, Dec OSF Tru64, SGI Irix), reliable, ANSI-compliant malloc/operator new library. SmartHeap supports multiple memory pools, includes a fixed-size allocator, and is thread-safe. SmartHeap also includes comprehensive memory debugging APIs to detect leakage, overwrites, double-frees, wild pointers, out of memory, references to previously freed memory, and other memory errors.

Implicitly Included Flags

This section contains descriptions of flags that were included implicitly by other flags, but which do not have a permanent home at SPEC.

System and Other Tuning Information

In order to take full advantage of using PGI's huge page runtime library, your system must be configured to use huge pages. It is safe to run binaries compiled with "-Msmartalloc=huge" on systems not configured to use huge pages, however, you will not benefit from the performance improvements huge pages offer. To configure your system for huge pages perform the following steps:

Note that further information about huge pages may be found in your Linux documentation file: /usr/src/linux/Documentation/vm/hugetlbpage.txt

Sets the stack size to n kbytes, or unlimited to allow the stack size to grow without limit.

For questions about the meanings of these flags, please contact the tester.
For other inquiries, please contact webmaster@spec.org
Copyright 2006-2014 Standard Performance Evaluation Corporation
Tested with SPEC CPU2006 v1.0.
Report generated on Tue Jul 22 12:48:27 2014 by SPEC CPU2006 flags formatter v6906.

	Indicates that the flag description came from the user flags file.
	Indicates that the flag description came from the suite-wide flags file.
	Indicates that the flag description came from a per-benchmark flags file.

CPU2006 Flag DescriptionTyan Thunder h2000M (S3992) AMD Opteron 2224 SE

Test sponsored by Advanced Micro Devices

Compilers: PGI 7.1-0

Operating systems: Linux

Base Compiler Invocation

Peak Compiler Invocation

Base Portability Flags

Peak Portability Flags

Base Optimization Flags

Peak Optimization Flags

Base Other Flags

Peak Other Flags

Implicitly Included Flags

CPU2006 Flag Description
Tyan Thunder h2000M (S3992) AMD Opteron 2224 SE