SPEC® CFP2006 Result

Copyright 2006-2014 Standard Performance Evaluation Corporation

IBM Corporation

IBM System x3755 (AMD Opteron 8380)

SPECfp®2006 = 20.6

CPU2006 license: 11 Test date: Jan-2009
Test sponsor: IBM Corporation Hardware Availability: Mar-2009
Tested by: Advanced Micro Devices Software Availability: May-2008
Benchmark results graph
Hardware
CPU Name: AMD Opteron 8380
CPU Characteristics:
CPU MHz: 2500
FPU: Integrated
CPU(s) enabled: 16 cores, 4 chips, 4 cores/chip
CPU(s) orderable: 1,2,3,4 chips
Primary Cache: 64 KB I + 64 KB D on chip per core
Secondary Cache: 512 KB I+D on chip per core
L3 Cache: 6 MB I+D on chip per chip
Other Cache: None
Memory: 64 GB (16 x 4 GB, DDR2-667 CL5 Reg Dual Rank)
Disk Subsystem: 1 x 73.4 GB SAS, 15000 RPM
Other Hardware: None
Software
Operating System: SuSE Linux Enterprise Server 10 (x86_64) SP1,
Kernel 2.6.16.46-0.12-smp
Compiler: PGI Server Complete Version 7.2
Auto Parallel: Yes
File System: ReiserFS
System State: Run level 3 (Full multiuser with network)
Base Pointers: 32/64-bit
Peak Pointers: 64-bit
Other Software: binutils 2.18.50

Results Table

Benchmark Base Peak
Seconds Ratio Seconds Ratio Seconds Ratio Seconds Ratio Seconds Ratio Seconds Ratio
Results appear in the order in which they were run. Bold underlined text indicates a median measurement.
410.bwaves 271   50.2 262   51.8 293   46.4 271   50.2 262   51.8 293   46.4
416.gamess 1291   15.2 1288   15.2 1290   15.2 1223   16.0 1222   16.0 1222   16.0
433.milc 626   14.7 627   14.6 627   14.6 623   14.7 614   15.0 614   15.0
434.zeusmp 690   13.2 688   13.2 689   13.2 626   14.5 612   14.9 612   14.9
435.gromacs 501   14.2 502   14.2 502   14.2 416   17.2 416   17.2 416   17.1
436.cactusADM 92.8 129   95.5 125   92.0 130   96.1 124   93.2 128   93.9 127  
437.leslie3d 671   14.0 917   10.3 678   13.9 750   12.5 745   12.6 604   15.6
444.namd 663   12.1 663   12.1 663   12.1 576   13.9 576   13.9 576   13.9
447.dealII 638   17.9 641   17.9 639   17.9 575   19.9 575   19.9 575   19.9
450.soplex 777   10.7 773   10.8 775   10.8 777   10.7 773   10.8 775   10.8
453.povray 337   15.8 337   15.8 337   15.8 312   17.0 315   16.9 314   16.9
454.calculix 520   15.9 522   15.8 520   15.9 423   19.5 422   19.5 422   19.5
459.GemsFDTD 380   27.9 357   29.8 367   28.9 380   27.9 357   29.8 367   28.9
465.tonto 666   14.8 668   14.7 666   14.8 587   16.8 590   16.7 591   16.7
470.lbm 529   26.0 529   26.0 528   26.0 529   26.0 529   26.0 528   26.0
481.wrf 544   20.5 541   20.6 544   20.5 606   18.4 607   18.4 617   18.1
482.sphinx3 1140   17.1 1071   18.2 1069   18.2 971   20.1 971   20.1 970   20.1

Submit Notes

The config file option 'submit' was used.
 'numactl' was used to bind copies to the cores.

Operating System Notes

 Environment stack size set to 'unlimited'.
 The powersaved was disabled, set the CPU frequency to its maximum.
 Total number of huge pages available is 14336.
 'ulimit -l 2097152' was used to set environment locked pages in memory quantity.
 Set vm/nr_hugepages=14336 in /etc/sysctl.conf
 mount -t hugetlbfs nodev /mnt/hugepages

General Notes

Environment variables set by runspec before the start of the run:
LD_LIBRARY_PATH = "/root/work/cpu2006v1.1/pgi72/linux_lib64:/root/work/cpu2006v1.1/pgi72/linux_lib32"
NCPUS = "16"

Base Compiler Invocation

C benchmarks:

 pgcc 

C++ benchmarks:

 pgcpp 

Fortran benchmarks:

 pgf95 

Benchmarks using both Fortran and C:

 pgcc   pgf95 

Base Portability Flags

410.bwaves:  -DSPEC_CPU_LP64 
416.gamess:  -DSPEC_CPU_LP64 
433.milc:  -DSPEC_CPU_LP64 
434.zeusmp:  -DSPEC_CPU_LP64 
435.gromacs:  -DSPEC_CPU_LP64   -Mnomain 
436.cactusADM:  -DSPEC_CPU_LP64   -Mnomain 
437.leslie3d:  -DSPEC_CPU_LP64 
444.namd:  -DSPEC_CPU_LP64 
447.dealII:  -DSPEC_CPU_LP64 
450.soplex:  -DSPEC_CPU_LP64 
453.povray:  -DSPEC_CPU_LP64 
454.calculix:  -DSPEC_CPU_LP64   -Mnomain 
459.GemsFDTD:  -DSPEC_CPU_LP64 
465.tonto:  -DSPEC_CPU_LP64 
470.lbm:  -DSPEC_CPU_LP64 
481.wrf:  -DSPEC_CPU_LP64   -DSPEC_CPU_CASE_FLAG   -DSPEC_CPU_LINUX 
482.sphinx3:  -DSPEC_CPU_LP64 

Base Optimization Flags

C benchmarks:

 -Mvect=cachesize:6291456   -fastsse   -Msmartalloc=huge   -Mconcur   -Mfprelaxed   -Mipa=fast   -Mipa=inline   -tp barcelona-64   -Bstatic_pgi 

C++ benchmarks:

 -Mvect=cachesize:6291456   -fastsse   -Msmartalloc=huge   -Mfprelaxed   -Mconcur   --zc_eh   -Mipa=fast   -Mipa=inline   -tp barcelona-64   -Bstatic_pgi 

Fortran benchmarks:

 -Mvect=cachesize:6291456   -fastsse   -Mfprelaxed   -Msmartalloc=huge   -Mconcur   -Mipa=fast   -Mipa=inline   -tp barcelona-64   -Bstatic_pgi 

Benchmarks using both Fortran and C:

 -Mvect=cachesize:6291456   -fastsse   -Msmartalloc=huge   -Mconcur   -Mfprelaxed   -Mipa=fast   -Mipa=inline   -tp barcelona-64   -Bstatic_pgi 

Base Other Flags

C benchmarks:

 -Mipa=jobs:8 

C++ benchmarks:

 -Mipa=jobs:8 

Fortran benchmarks:

 -Mipa=jobs:8 

Benchmarks using both Fortran and C:

 -Mipa=jobs:8 

Peak Compiler Invocation

C benchmarks:

 pgcc 

C++ benchmarks:

 pgcpp 

Fortran benchmarks:

 pgf95 

Benchmarks using both Fortran and C:

 pgcc   pgf95 

Peak Portability Flags

410.bwaves:  -DSPEC_CPU_LP64 
416.gamess:  -DSPEC_CPU_LP64 
433.milc:  -DSPEC_CPU_LP64 
434.zeusmp:  -DSPEC_CPU_LP64 
435.gromacs:  -DSPEC_CPU_LP64   -Mnomain 
436.cactusADM:  -DSPEC_CPU_LP64   -Mnomain 
437.leslie3d:  -DSPEC_CPU_LP64 
444.namd:  -DSPEC_CPU_LP64 
450.soplex:  -DSPEC_CPU_LP64 
453.povray:  -DSPEC_CPU_LP64 
454.calculix:  -DSPEC_CPU_LP64   -Mnomain 
459.GemsFDTD:  -DSPEC_CPU_LP64 
465.tonto:  -DSPEC_CPU_LP64 
470.lbm:  -DSPEC_CPU_LP64 
481.wrf:  -DSPEC_CPU_LP64   -DSPEC_CPU_CASE_FLAG   -DSPEC_CPU_LINUX 
482.sphinx3:  -DSPEC_CPU_LP64 

Peak Optimization Flags

C benchmarks:

433.milc:  -Mvect=cachesize:6291456   -fastsse   -Msmartalloc=huge   -Msafeptr   -Mconcur   -Mfprelaxed   -Mipa=inline   -Mipa=arg   -Mipa=const   -Mipa=ptr   -Mipa=shape   -tp barcelona-64   -Bstatic_pgi 
470.lbm:  basepeak = yes 
482.sphinx3:  -Mpfi(pass 1)   -Mpfo(pass 2)   -Mipa=fast(pass 2)   -Mipa=inline(pass 2)   -Mvect=cachesize:6291456   -fastsse   -Mfprelaxed   -Msmartalloc   -tp barcelona-64   -Bstatic_pgi 

C++ benchmarks:

444.namd:  -Mpfi(pass 1)   -Mpfo(pass 2)   -Mipa=fast(pass 2)   -Mipa=inline(pass 2)   -Mvect=cachesize:6291456   -fastsse   -Munroll=n:4   -Munroll=m:8   -Msmartalloc=huge   -Mnodepchk   -Mfprelaxed   --zc_eh   -tp barcelona-64   -Bstatic_pgi 
447.dealII:  -Mvect=cachesize:6291456   -fastsse   -alias=ansi   -Msmartalloc=huge   -Mprefetch=t0   -Mnovect   -Mfprelaxed   --zc_eh   -Mipa=fast   -Mipa=inline   -tp barcelona-32   -Bstatic_pgi 
450.soplex:  basepeak = yes 
453.povray:  -Mpfi=indirect(pass 1)   -Mpfo=indirect(pass 2)   -Mipa=fast(pass 2)   -Mipa=inlinenopfo:3(pass 2)   -Mipa=staticfunc(pass 2)   -Mvect=cachesize:6291456   -fastsse   -Msmartalloc=huge   -Mprefetch=t0   -Mfprelaxed   -tp barcelona-64   -Bstatic_pgi 

Fortran benchmarks:

410.bwaves:  basepeak = yes 
416.gamess:  -Mpfi(pass 1)   -Mpfo(pass 2)   -Mipa=fast(pass 2)   -Mipa=inline(pass 2)   -Mvect=cachesize:6291456   -fastsse   -Msmartalloc=huge   -Mvect=noaltcode   -Mprefetch=t0   -Mfprelaxed   -tp barcelona-64   -Bstatic_pgi 
434.zeusmp:  -Mvect=cachesize:6291456   -fastsse   -Mfprelaxed   -Mconcur   -Mprefetch=distance:8   -Mprefetch=t0   -Msmartalloc=huge   -Msmartalloc=hugebss   -Mipa=fast   -Mipa=inline   -tp barcelona-64   -Bstatic_pgi 
437.leslie3d:  -Mpfi=indirect(pass 1)   -Mpfo=indirect(pass 2)   -Mconcur=noaltcode(pass 2)   -Mipa=fast(pass 2)   -Mipa=inline(pass 2)   -Mvect=cachesize:6291456   -fastsse   -Mvect=fuse   -Msmartalloc=huge   -Mprefetch=distance:8   -Mprefetch=t0   -Mfprelaxed   -tp barcelona-64   -Bstatic_pgi 
459.GemsFDTD:  basepeak = yes 
465.tonto:  -Mvect=cachesize:6291456   -fastsse   -O4   -Mvect=noaltcode   -Msmartalloc=huge   -Mprefetch=distance:8   -Mprefetch=t0   -Mfprelaxed   -Mipa=fast   -Mipa=inline   -tp barcelona-64   -Bstatic_pgi 

Benchmarks using both Fortran and C:

435.gromacs:  -Mvect=cachesize:6291456   -fastsse   -Msmartalloc=huge   -Mfprelaxed   -Mconcur   -Mfpapprox=rsqrt   -Mipa=fast   -Mipa=inline   -tp barcelona-64   -Bstatic_pgi 
436.cactusADM:  -Mvect=cachesize:6291456   -fastsse   -Msmartalloc=huge   -Mfprelaxed   -Mconcur   -Mdse   -Mipa=fast   -Mipa=inline   -tp barcelona-64   -Bstatic_pgi 
454.calculix:  -Mpfi=indirect(pass 1)   -Mpfo=indirect(pass 2)   -Mipa=fast(pass 2)   -Mipa=inline(pass 2)   -Mvect=cachesize:6291456   -fastsse   -Msmartalloc=huge   -Mloop32   -Mprefetch=t0   -Mpre   -Mfprelaxed   -tp barcelona-64   -Bstatic_pgi 
481.wrf:  -Mvect=cachesize:6291456   -fastsse   -Mvect=noaltcode   -Msmartalloc=huge   -Mprefetch=distance:8   -Mconcur=noaltcode   -Mfprelaxed   -tp barcelona-64   -Bstatic_pgi 

Peak Other Flags

C benchmarks:

 -Mipa=jobs:8(pass 2) 

C++ benchmarks:

 -Mipa=jobs:8(pass 2) 

Fortran benchmarks:

 -Mipa=jobs:8 

Benchmarks using both Fortran and C (except as noted below):

 -Mipa=jobs:8(pass 2) 
481.wrf:  No flags used 

The flags file that was used to format this result can be browsed at
http://www.spec.org/cpu2006/flags/pgi72_linux_flags.20090713.html.

You can also download the XML flags source by saving the following link:
http://www.spec.org/cpu2006/flags/pgi72_linux_flags.20090713.xml.