mirror of
https://github.com/ClusterCockpit/cc-metric-collector.git
synced 2024-12-28 08:09:06 +01:00
34 lines
1.3 KiB
Plaintext
34 lines
1.3 KiB
Plaintext
|
SHORT Single Precision MFLOP/s
|
||
|
|
||
|
EVENTSET
|
||
|
FIXC0 INSTR_RETIRED_ANY
|
||
|
FIXC1 CPU_CLK_UNHALTED_CORE
|
||
|
FIXC2 CPU_CLK_UNHALTED_REF
|
||
|
PMC0 FP_COMP_OPS_EXE_SSE_FP_PACKED_SINGLE
|
||
|
PMC1 FP_COMP_OPS_EXE_SSE_FP_SCALAR_SINGLE
|
||
|
PMC2 SIMD_FP_256_PACKED_SINGLE
|
||
|
|
||
|
METRICS
|
||
|
Runtime (RDTSC) [s] time
|
||
|
Runtime unhalted [s] FIXC1*inverseClock
|
||
|
Clock [MHz] 1.E-06*(FIXC1/FIXC2)/inverseClock
|
||
|
CPI FIXC1/FIXC0
|
||
|
SP [MFLOP/s] 1.0E-06*(PMC0*4.0+PMC1+PMC2*8.0)/time
|
||
|
AVX SP [MFLOP/s] 1.0E-06*(PMC2*8.0)/time
|
||
|
Packed [MUOPS/s] 1.0E-06*(PMC0+PMC2)/time
|
||
|
Scalar [MUOPS/s] 1.0E-06*PMC1/time
|
||
|
Vectorization ratio 100*(PMC0+PMC2)/(PMC0+PMC1+PMC2)
|
||
|
|
||
|
LONG
|
||
|
Formulas:
|
||
|
SP [MFLOP/s] = 1.0E-06*(FP_COMP_OPS_EXE_SSE_FP_PACKED*4+FP_COMP_OPS_EXE_SSE_FP_SCALAR+SIMD_FP_256_PACKED_SINGLE*8)/runtime
|
||
|
AVX SP [MFLOP/s] = 1.0E-06*(SIMD_FP_256_PACKED_SINGLE*8)/runtime
|
||
|
Packed [MUOPS/s] = 1.0E-06*(FP_COMP_OPS_EXE_SSE_FP_PACKED_SINGLE+SIMD_FP_256_PACKED_SINGLE)/runtime
|
||
|
Scalar [MUOPS/s] = 1.0E-06*FP_COMP_OPS_EXE_SSE_FP_SCALAR_SINGLE/runtime
|
||
|
Vectorization ratio = 100*(FP_COMP_OPS_EXE_SSE_FP_PACKED_SINGLE+SIMD_FP_256_PACKED_SINGLE)/(FP_COMP_OPS_EXE_SSE_FP_SCALAR_SINGLE+FP_COMP_OPS_EXE_SSE_FP_PACKED_SINGLE+SIMD_FP_256_PACKED_SINGLE)
|
||
|
-
|
||
|
SSE scalar and packed single precision FLOP rates. Please note that the current
|
||
|
FLOP measurements on IvyBridge are potentially wrong.
|
||
|
So you cannot trust these counters at the moment!
|
||
|
|