SPEC® CFP2006 Result

Copyright 2006-2014 Standard Performance Evaluation Corporation

Dell Inc.


PowerEdge R715 (AMD Opteron 6220, 3.00 GHz)

SPECfp®2006 = 45.0

CPU2006 license: 55 Test date: Jul-2012
Test sponsor: Dell Inc. Hardware Availability: Nov-2011
Tested by: Dell Inc. Software Availability: Jul-2011
Benchmark results graph
Hardware
CPU Name: AMD Opteron 6220
CPU Characteristics: AMD Turbo CORE technology up to 3.60 GHz
CPU MHz: 3000
FPU: Integrated
CPU(s) enabled: 16 cores, 2 chips, 8 cores/chip
CPU(s) orderable: 1,2 chips
Primary Cache: 256 KB I on chip per chip,
64 KB I shared / 2 cores;
16 KB D on chip per core
Secondary Cache: 8 MB I+D on chip per chip, 2 MB shared / 2 cores
L3 Cache: 16 MB I+D on chip per chip, 8 MB shared / 4 cores
Other Cache: None
Memory: 64 GB (8 x 8 GB 2Rx4 PC3-12800R-11, ECC)
Disk Subsystem: 3 x 146 GB SAS, 15000 RPM
Other Hardware: None
Software
Operating System: SUSE Linux Enterprise Server 11 SP2 (x86_64)
3.0.13-0.27-default
Compiler: C/C++/Fortran: Version 4.2.5.2 of x86 Open64
Compiler Suite (from AMD)
Auto Parallel: Yes
File System: ext3
System State: Run level 3 (Full multiuser with network)
Base Pointers: 64-bit
Peak Pointers: 32/64-bit
Other Software: None

Results Table

Benchmark Base Peak
Seconds Ratio Seconds Ratio Seconds Ratio Seconds Ratio Seconds Ratio Seconds Ratio
Results appear in the order in which they were run. Bold underlined text indicates a median measurement.
410.bwaves 62.1 219   62.1 219   62.2 218   43.1 315   43.8 311   43.4 313  
416.gamess 1012   19.4 1016   19.3 1010   19.4 955   20.5 960   20.4 962   20.4
433.milc 279   32.9 279   32.9 279   32.9 235   39.1 234   39.2 233   39.3
434.zeusmp 138   65.8 136   66.7 137   66.6 128   71.3 128   70.9 128   71.2
435.gromacs 335   21.3 335   21.3 335   21.3 329   21.7 329   21.7 329   21.7
436.cactusADM 83.4 143   82.9 144   83.4 143   65.0 184   64.6 185   64.5 185  
437.leslie3d 347   27.1 347   27.1 347   27.1 347   27.1 347   27.1 347   27.1
444.namd 475   16.9 475   16.9 474   16.9 461   17.4 461   17.4 461   17.4
447.dealII 290   39.5 290   39.5 290   39.5 262   43.7 262   43.7 262   43.7
450.soplex 361   23.1 360   23.2 360   23.2 334   25.0 333   25.1 334   25.0
453.povray 247   21.6 247   21.5 248   21.5 231   23.1 230   23.1 231   23.1
454.calculix 261   31.6 261   31.7 261   31.6 259   31.9 258   31.9 259   31.9
459.GemsFDTD 281   37.8 281   37.8 280   37.8 251   42.3 251   42.3 251   42.2
465.tonto 399   24.7 398   24.7 398   24.7 399   24.7 398   24.7 398   24.7
470.lbm 146   93.8 147   93.6 146   93.9 43.5 316   43.6 315   44.1 312  
481.wrf 245   45.5 246   45.4 246   45.4 245   45.5 246   45.4 246   45.4
482.sphinx3 815   23.9 817   23.9 816   23.9 631   30.9 631   30.9 629   31.0

Submit Notes

The config file option 'submit' was used.
'numactl' was used to bind copies to the cores.
See the configuration file for details.

Operating System Notes

'ulimit -s unlimited' was used to set environment stack size
'ulimit -l 2097152'  was used to set environment locked pages in memory limit

Set transparent_hugepage=never as a boot parameter in /boot/grub/menu.lst
Set kernel/randomize_va_space=0 in /etc/sysctl.conf
cpuspeed stop was used to set the CPU frequency to its maximum.

Set vm/nr_hugepages=4000 in /etc/sysctl.conf
mount -t hugetlbfs nodev /mnt/hugepages

General Notes

Environment variables set by runspec before the start of the run:
HUGETLB_LIMIT = "4000"
LD_LIBRARY_PATH = "/root/cpu2006/amd1104-speed-libs-revA/32:/root/cpu2006/amd1104-speed-libs-revA/64"
O64_OMP_AFFINITY_MAP = "0,1,2,3,4,5,6,7,8,9,10,11,12,13,14,15"
O64_OMP_SPIN_COUNT = "800000"
O64_OMP_SPIN_USER_LOCK = "true"

The x86 Open64 Compiler Suite is only available from (and supported by) AMD at
http://developer.amd.com/cpu/open64

Binaries were compiled on a system with 2x AMD Opteron 6220 chips + 64GB Memory using RHEL 6.1

Base Compiler Invocation

C benchmarks:

 opencc 

C++ benchmarks:

 openCC 

Fortran benchmarks:

 openf95 

Benchmarks using both Fortran and C:

 opencc   openf95 

Base Portability Flags

410.bwaves:  -DSPEC_CPU_LP64 
416.gamess:  -DSPEC_CPU_LP64 
433.milc:  -DSPEC_CPU_LP64 
434.zeusmp:  -DSPEC_CPU_LP64 
435.gromacs:  -DSPEC_CPU_LP64 
436.cactusADM:  -DSPEC_CPU_LP64   -fno-second-underscore 
437.leslie3d:  -DSPEC_CPU_LP64 
444.namd:  -DSPEC_CPU_LP64 
447.dealII:  -DSPEC_CPU_LP64 
450.soplex:  -DSPEC_CPU_LP64 
453.povray:  -DSPEC_CPU_LP64 
454.calculix:  -DSPEC_CPU_LP64 
459.GemsFDTD:  -DSPEC_CPU_LP64 
465.tonto:  -DSPEC_CPU_LP64 
470.lbm:  -DSPEC_CPU_LP64 
481.wrf:  -DSPEC_CPU_LP64   -DSPEC_CPU_LINUX   -DSPEC_CPU_CASE_FLAG   -fno-second-underscore 
482.sphinx3:  -DSPEC_CPU_LP64 

Base Optimization Flags

C benchmarks:

 -march=bdver1   -Ofast   -HP:bdt=2m:heap=2m   -apo   -mso   -OPT:alias=restricted   -OPT:malloc_alg=2   -LNO:parallel_overhead=10000 

C++ benchmarks:

 -march=bdver1   -Ofast   -static   -CG:load_exe=0   -CG:p2align=0   -INLINE:aggressive=on   -HP:bdt=2m:heap=2m   -D__OPEN64_FAST_SET 

Fortran benchmarks:

 -march=bdver1   -Ofast   -LNO:blocking=off   -LNO:fusion_peeling_limit=0   -LNO:parallel_overhead=10000   -OPT:rsqrt=2   -OPT:unroll_size=256   -HP:bdt=2m:heap=2m   -apo 

Benchmarks using both Fortran and C:

 -march=bdver1   -Ofast   -HP:bdt=2m:heap=2m   -apo   -mso   -OPT:alias=restricted   -OPT:malloc_alg=2   -LNO:parallel_overhead=10000   -LNO:blocking=off   -LNO:fusion_peeling_limit=0   -OPT:rsqrt=2   -OPT:unroll_size=256 

Peak Compiler Invocation

C benchmarks:

 opencc 

C++ benchmarks:

 openCC 

Fortran benchmarks:

 openf95 

Benchmarks using both Fortran and C:

 opencc   openf95 

Peak Portability Flags

410.bwaves:  -DSPEC_CPU_LP64 
416.gamess:  -DSPEC_CPU_LP64 
433.milc:  -DSPEC_CPU_LP64 
434.zeusmp:  -DSPEC_CPU_LP64 
435.gromacs:  -DSPEC_CPU_LP64 
436.cactusADM:  -DSPEC_CPU_LP64   -fno-second-underscore 
437.leslie3d:  -DSPEC_CPU_LP64 
444.namd:  -DSPEC_CPU_LP64 
447.dealII:  -DSPEC_CPU_LP64 
453.povray:  -DSPEC_CPU_LP64 
454.calculix:  -DSPEC_CPU_LP64 
459.GemsFDTD:  -DSPEC_CPU_LP64 
465.tonto:  -DSPEC_CPU_LP64 
470.lbm:  -DSPEC_CPU_LP64 
481.wrf:  -DSPEC_CPU_LP64   -DSPEC_CPU_LINUX   -DSPEC_CPU_CASE_FLAG   -fno-second-underscore 
482.sphinx3:  -DSPEC_CPU_LP64 

Peak Optimization Flags

C benchmarks:

433.milc:  -march=bdver1   -Ofast   -CG:movnti=1   -CG:locs_best=on   -HP:bdt=2m:heap=2m   -IPA:plimit=7000   -IPA:callee_limit=1200   -OPT:struct_array_copy=2   -OPT:alias=field_sensitive 
470.lbm:  -march=bdver1   -Ofast   -mso   -apo   -CG:sse_cse_regs=0   -LNO:prefetch_ahead=4   -CG:locs_shallow_depth=1   -CG:cmp_peep=on   -CG:compute_to=on   -OPT:unroll_times_max=8   -OPT:unroll_size=256   -OPT:unroll_level=2   -OPT:keep_ext=on   -OPT:alias=restricted   -m3dnow   -IPA:inline=off 
482.sphinx3:  -march=bdver1   -fb_create fbdata(pass 1)   -fb_opt fbdata(pass 2)   -Ofast   -LNO:loop_model_simd=on   -LNO:simd_rm_unity_remainder=on   -OPT:malloc_alg=2   -CG:cmp_peep=on   -CG:local_sched_alg=2   -CG:use_incdec=off   -INLINE:aggressive=on   -WOPT:sib=on   -HP 

C++ benchmarks:

444.namd:  -march=bdver1   -fb_create fbdata(pass 1)   -fb_opt fbdata(pass 2)   -Ofast   -LNO:ignore_feedback=off   -CG:local_sched_alg=2   -CG:load_exe=0   -OPT:unroll_size=256   -fno-exceptions   -HP:bdt=2m:heap=2m 
447.dealII:  -march=bdver1   -Ofast   -LNO:simd=0   -D__OPEN64_FAST_SET   -static   -INLINE:aggressive=on   -OPT:alias=disjoint   -OPT:unroll_times_max=8   -OPT:unroll_size=256   -OPT:unroll_level=2   -HP:bdt=2m:heap=2m 
450.soplex:  -march=bdver1   -fb_create fbdata(pass 1)   -fb_opt fbdata(pass 2)   -O3   -INLINE:aggressive=on   -OPT:RO=1   -OPT:IEEE_arith=3   -OPT:IEEE_NaN_Inf=off   -OPT:fold_unsigned_relops=on   -fno-exceptions   -CG:p2align=0   -m32   -HP:bdt=2m:heap=2m   -WOPT:sib=on 
453.povray:  -march=bdver1   -fb_create fbdata(pass 1)   -fb_opt fbdata(pass 2)   -Ofast   -CG:pre_local_sched=off   -INLINE:aggressive=on   -HP:bdt=2m:heap=2m   -OPT:transform=2   -OPT:alias=disjoint   -WOPT:aggcm=0 

Fortran benchmarks:

410.bwaves:  -march=bdver1   -fb_create fbdata(pass 1)   -fb_opt fbdata(pass 2)   -Ofast   -apo   -OPT:Ofast   -OPT:treeheight=on   -LNO:blocking=off   -LNO:prefetch=2   -LNO:pf2=0   -LNO:prefetch_ahead=3   -LNO:ignore_feedback=off   -LNO:fu=4   -LNO:loop_model_simd=on   -LNO:simd_rm_unity_remainder=on   -WOPT:aggstr=0   -HP:bdt=2m:heap=2m   -CG:cmp_peep=on   -CG:p2align=0 
416.gamess:  -march=bdver1   -fb_create fbdata(pass 1)   -fb_opt fbdata(pass 2)   -O3   -LNO:fu=6   -LNO:blocking=0   -LNO:simd=0   -OPT:Ofast   -OPT:ro=3   -OPT:unroll_size=256   -OPT:unroll_times_max=2   -CG:local_sched_alg=1   -HP:bdt=2m:heap=2m   -WOPT:sib=on 
434.zeusmp:  -march=bdver1   -Ofast   -apo   -LNO:blocking=off   -LNO:interchange=off   -LNO:fusion_peeling_limit=0   -OPT:treeheight=on   -OPT:unroll_size=256   -CG:cmp_peep=on   -CG:compute_to=on   -GRA:prioritize_by_density=on   -HP:bdt=2m:heap=2m 
437.leslie3d:  basepeak = yes 
459.GemsFDTD:  -march=bdver1   -Ofast   -OPT:unroll_size=0   -LNO:fission=2   -CG:load_exe=0   -CG:local_sched_alg=2   -HP   -apo 
465.tonto:  basepeak = yes 

Benchmarks using both Fortran and C:

435.gromacs:  -march=bdver1   -fb_create fbdata(pass 1)   -fb_opt fbdata(pass 2)   -Ofast   -OPT:rsqrt=2   -HP:bdt=2m:heap=2m 
436.cactusADM:  -march=bdver1   -fb_create fbdata(pass 1)   -fb_opt fbdata(pass 2)   -Ofast   -LNO:blocking=off   -LNO:prefetch=2   -HP:bdt=2m:heap=2m   -CG:locs_shallow_depth=1   -CG:load_exe=0   -WOPT:sib=on   -apo 
454.calculix:  -march=bdver1   -Ofast   -OPT:unroll_size=256   -GRA:optimize_boundary=on   -HP:bdt=2m:heap=2m 
481.wrf:  basepeak = yes 

The flags files that were used to format this result can be browsed at
http://www.spec.org/cpu2006/flags/x86-open64-425-flags-speed-revA-I.html,
http://www.spec.org/cpu2006/flags/amd-platform-speed-revA-I.html.

You can also download the XML flags sources by saving the following links:
http://www.spec.org/cpu2006/flags/x86-open64-425-flags-speed-revA-I.xml,
http://www.spec.org/cpu2006/flags/amd-platform-speed-revA-I.xml.