| Cell | HPC Cell | Intel Clovertown | AMD Winsor+ | ATI R580 X1950 | nVidia G80 GTX | |
|---|---|---|---|---|---|---|
| Memory on Chip | 512KB(L2)+ 256KB(LS)x8 | 512KB(L2)+ 256KB(LS)x8 | 4MB(L2)x2 | 1MB(L2)x2 | (TC) | 16KBx16(SHM)+ (TC) |
| Memory off Chip | 25.6 GB/s | 25.6 GB/s | 10.41 GB/s(FSB) | 17.0 GB/s (DDR2-1066x2) | 64.0 GB/s | 86.4 GB/s |
| IO | 25.6 GB/s | 25.6 GB/s | 10.41 GB/s (FSB) | 22.4 GB/s (HT) | 4.0 GB/s (PCIex16) | 4.0 GB/s (PCIex16) |
Showing posts with label gpu. Show all posts
Showing posts with label gpu. Show all posts
Wednesday, July 4, 2007
Cell vs. GPU vs. CPU
A slide presents transistors and chip area for Cell, GPUs, and CPUs. The following are other resource comparisons I am interested.
Wednesday, May 16, 2007
G80 more details from R600
R600 came out finally. But, only a middle-end one can be ordered now. The reviews (1, 2, and etc.) didn't show much advantages over NVIDIA's G80. NVIDIA also came out a FUD presentation. It is quite interesting to read the presentation and it reveals more details of G80 and R600.
- G80 has one special function unit per ALU, which is never documented or showed in any presentation from NVIDIA and manuals of CUDA.
- R600 is a super-scalar VLIW architecture, which is different from R580 design and uses 5 scalar units instead of 1 vector unit and 1 scalar unit. That is an evolution step to the architectures used in GPU. The analysis on super-scalar VLIW architecture is fair to claim the efficiency issue. But, G80 is not a true scalar architecture, it's still kind of vector processor and also efficiency issue (and, yes, from somewhat different perspectives.) But, with compiler improvement, R600 will have more advantages if latency is critical to the computation.
- It is reported that R600 is a cache heavy design, most of die size is devoted to SRAM. Beside texture cache and vertex cache, R600 has additional read/write cache to virtualize registers. This will benefit GPGPU (depending on the configuration of that read/write cache.) If such a read/write cache is "real" cache, that will improve the programmability and performance of R600 compared to CUDA on G80. The ever increasing complexity of graphics workload and GPGPU popularity has created demands on memory hierarchy for GPU. R600 may be optimized for stream computing.
Labels:
architecture,
gpu
Thursday, March 22, 2007
new R600 photos
VR-Zone posted the new photos of R600-based X2900 XTX. It is very different from the previous photos. Definitely, the previous ones should be the engineering boards. The new one is much shorter and should be the same length as NVIDIA's G8800 GTX. More photos and schematic are also posted on the forum of VR-Zone.
The only question I want to know is when this monster will be available, the exact date.
The only question I want to know is when this monster will be available, the exact date.
Subscribe to:
Posts (Atom)