Guru3D.com
  • HOME
  • NEWS
    • Channels
    • Archive
    • Search
    • Submit
  • DOWNLOADS
    • New Downloads
    • Categories
    • Archive
    • Search
    • Submit
  • GAME REVIEWS
  • ARTICLES
    • Editorials
    • Guru3D VGA Charts
    • Rig of the Month
    • Join ROTM
    • PC Buyers Guide
    • Dated content
    • More Categories
  • HARDWARE REVIEWS
    • Videocards
    • Processors
    • Audio
    • Motherboards
    • Memory and Flash
    • SSD Storage
    • Chassis
    • Power Supply
    • Laptop and Mobile
    • Smartphone
    • Networking
    • Keyboard Mouse
    • Cooling
    • Knowledgebase
    • Search articles
    • More Categories
  • FORUMS
  • SEARCH
    • Search Articles
    • Search News
    • Search Files
  • NEWSLETTER
  • CONTACT

New Reviews
Gigabyte GeForce GTX 650 Ti Boost OC WindForce 2X review
MSI Radeon HD 7790 TurboDuo OC review
Metro Last Light VGA Graphics Benchmark performance test
Noctua NH-U12S and NH-U14S review
ASUS GeForce GTX 670 DirectCU Mini review
OCZ Vertex 3.20 SSD review
Gigabyte Radeon HD 7790 2GB OC review
Cooler Master Eisberg 240L Prestige review
Guru3D and OCZ Contest - PC Power 1200W PSU Giveaway
MSI GeForce GTX 650 Ti BOOST OC review

New Downloads
MSI Afterburner 3.0.0 Beta 10 Download
PhysX System Software 9.13.0325 Download
GPU-Z Download 0.7.1
HWiNFO32 4.18 Download
HWiNFO64 4.18 Download
GeForce 320.14 BETA Driver Download
Nvidia Lifelike Human Face Rendering Tech Demo Download
3DMark Download v1.1.0
XBMC Media Center Download 12.0 2
RTSS Rivatuner Statistics Server Download v5.1.1


New Forum Topics
by: dellon132 ATI Catalyst 12.11 Beta 11 Modded driver For Legacy GPUby: CPC_RedDawn Fine tuning my GPU overclock.by: Stone Gargoyle Xbox World reveals Next Gen Xbox?by: Hilbert Hagedoorn NVIDIA GeForce GTX 780, GTX 770 and GTX 760 Tiby: Stone Gargoyle FIFA 14 PC won't use new Ignite engineby: dekka nvidia drivers 320.22 whqlby: Pigchild Can't get past 4.4Ghz. Please review need advice.by: Hilbert Hagedoorn Fractal Design Node 304 gets all whiteby: Hilbert Hagedoorn Microsoft Xbox One console shownby: Hilbert Hagedoorn RadeonPro BETA (Automating 3D Settings) #2


Online Users
There are currently 2974 user(s) online:
ablev, ACMPraetorian, BahamutxD, BLEH!, DesGaizu, Dillinger, dsbig, Google, Huigie, kens30, Lane, leopr, Li4m79, Live Search, Memorian, MSN, scipio, Starfighter2, winny, Yahoo


Guru3D.com » Review » ASUS Radeon HD 7970 ROG MATRIX Platinum review » Page 3

ASUS Radeon HD 7970 ROG MATRIX Platinum review

Posted by Hilbert Hagedoorn on: 10/16/2012 06:21 AM [ 13 comment(s) ]

The Graphics engine architecture
Tweet

 

The Graphics engine architecture

So I kept the more complex stuff for last in the technology overview. If this seems a little too techy for you, skip this page please.

AMD is moving away from the VLIW5 and VLIW4 architecture we have seen in the last generation of products. If anything, VLIW4 has shown certain inefficiencies in the Radeon HD 6900 series and while VLIW designs are fine for graphics they are not so grand for computing.

The new graphics core architecture is now marketed as GCN, which is short for Graphics Core Next architecture and the architecture building block has changed significantly to remove certain inefficiencies seen in the VLIW architecture.

A GCN is in its essence the basis of a GPU that performs well at both graphical and computing tasks. For the compute side of things the new GCN Compute unit model has been introduced, it is designed for better utilization, high throughput and multi tasking. E.g. performance, performance, performance.

So your basic new Shader cluster is one called a (GCN) Compute Unit:

  • Non-VLIW Design
  • 16 wide SIMD Units
  • 64 KB registers / SIMD Unit

Now if we take 4 of these SIMD Units, that will form the basis of one Compute Unit (CU). Each SIMD unit is 16 wide, times four per compute unit means that each CU unit has 64 shader processors. The GPU has 32 Compute units meaning 64SIMDs x 32 CUs = 2048 Shader processors (for the R7970).

  • Engine has Dual Geometry engines / Asynchronous Compute engines
  • 8 render backends / 32 color ROPs per clock cycle / 128 Z/Stencil ROPs per clock
  • Engine ties to 768KB R/W L2 cache
  • Tahiti GPU has up-to 32 Compute Units

The Graphics Core Next Compute Unit (CU) has about the same floating point power per clock as the previous one (i.e. Cayman). It also has the same amount of register space (for the vector units). Each CU also has its own registers and local data share.

Again: one compute unit just as a Cayman SIMD is a collection of shader processors, four SIMDs form one compute unit. Cayman's (6900) problem was that it was not so efficient with multiple tasks at once.

Cayman had/has 16 4-wide VLIW processing elements for a total of 16x4=64 operations in parallel, while the new architecture has 4 16-wide vector processors, again for a total of 4x16=64 operations per clock. GCN also has a scalar processor that Cayman does not.

The distinction is in its bare essence that GCN does not need instruction level parallelism, each of the four 16-wide SIMD vector units execute a different wavefront being the whole 64-sized wavefront taking four cycles.

Radeon HD 7970

So the theoretical floating point power stays more or less the same per CU, but GCN will be more efficient since it does not require instruction level parallelism (we assume it costs some more area/transistors as well). The outcome, compiling also becomes much more uncomplicated and that means more efficiency and thus there it is again, better performance.

GCN is all about creating a GPU good for both graphics and computing purposes. Oh and all compute units ... combined with the other ASIC components form the GPU. See, easy peasy, right? :)





25 pages « 2 3 4 5 next »


Guru3D.com » Articles » ASUS Radeon HD 7970 ROG MATRIX Platinum review » Page 3

Related Articles
ASUS Radeon HD 7970 ROG MATRIX Platinum review
We review the ASUS Radeon HD 7970 ROG MATRIX PLATINUM graphics card. Designed to be one of the most tweak-able and desirable graphics cards.

ASUS Radeon HD 7970 DirectCU II review
We review the ASUS Radeon HD 7970 DirectCU II. A factory overclocked three slot wide beast with massive cooling capacity. The completely customized card comes with voltage measurement points, a 12-phase VRM circuitry with supper alloy caps and chokes as well as a special SAP capacitor added to maximize overclocking headroom, according to ASUS.

ASUS Radeon HD 7970 Crossfire review
We review the Radeon HD 7970 in Crossfire, a second board partner card has arrived. Let's take it to the next level -- multi-GPU gaming in Crossfire mode.

ASUS Radeon HD 6770 DirectCU Silent review
Today we look at the ASUS Radeon HD 6770 DirectCU Silent edition. An entry level product series for gamers based on a juniper graphics core that still can flex its muscle. The x-factor for this particular product however is that its completely passively cooled, what should be a relatively small card .. ends up like the USS enterprise in terms of size and design.

Follow Guru3D on Google+ - Facebook - YouTube - Twitter © 2013