【问题标题】:How to compute the theoretical peak performance of CPU如何计算 CPU 的理论峰值性能
【发布时间】:2011-06-09 07:57:23
【问题描述】:

这是我的cat /proc/cpuinfo 输出:

...

processor           : 15
vendor_id           : GenuineIntel
cpu family          : 6
model               : 26
model name          : Intel(R) Xeon(R) CPU           E5520  @ 2.27GHz
stepping            : 5
cpu MHz             : 1600.000
cache size          : 8192 KB
physical id         : 1
siblings            : 8
core id             : 3
cpu cores           : 4
apicid              : 23
fpu                 : yes
fpu_exception       : yes
cpuid level         : 11
wp                  : yes
flags               : fpu vme de pse tsc msr pae mce cx8 apic ...
bogomips            : 4533.56
clflush size        : 64
cache_alignment     : 64
address sizes       : 40 bits physical, 48 bits virtual
power management    :

这台机器有两个CPU,每个都有4核,具有超线程能力,所以处理器总数为16(2 CPU * 4核* 2超线程)。这些处理器具有相同的输出,为了保持整洁,我只显示最后一个的信息并省略标志行中的部分标志。

那么我该如何计算这台机器的 GFlops 峰值性能呢? 如果需要提供更多信息,请告诉我。

谢谢。

【问题讨论】:

  • 很抱歉,很奇怪,Hi, 无法显示。
  • 自动删除问候语。

标签: performance cpu cpu-speed


【解决方案1】:

您可以查看Intel export spec。 图表中的 GFLOP 通常被称为单个芯片的峰值。 它显示 E5520 为 36.256 Gflop/s。

这个单芯片有 4 个带 SSE 的物理内核。 所以这个 GFLOP 也可以计算为: 2.26GHz*2(mul,add)*2(SIMD 双精度)*4(物理核心) = 36.2。

您的系统有两个 CPU,因此您的峰值为 36.2*2 = 72.4 GFLOP/S。

【讨论】:

  • 谁能解释一下“(mul, add)”?
  • mul:浮点乘法,add:浮点加法。这些是在 CPU 内核上执行的指令,我们假设这两条指令可以同时发生,因为 CPU 内核具有分离的乘法器和加法器。
【解决方案2】:

你可以在这个网站找到一个公式:

http://www.novatte.com/our-blog/197-how-to-calculate-peak-theoretical-performance-of-a-cpu-based-hpc-system

这里是公式:

以 GFlops 为单位的性能 =(以 GHz 为单位的 CPU 速度)x(CPU 内核数)x(每个周期的 CPU 指令)x(每个节点的 CPU 数)。

所以在你的情况下:2.27x4x4x2=72.64 GFLOP/s CPU配置见这里http://ark.intel.com/products/40200

【讨论】:

    猜你喜欢
    • 1970-01-01
    • 2014-07-28
    • 1970-01-01
    • 1970-01-01
    • 2018-10-09
    • 1970-01-01
    • 2014-06-28
    • 1970-01-01
    • 1970-01-01
    相关资源
    最近更新 更多