NVIDIA's own documentation states the architectural bargain: a GPU devotes more transistors to data processing rather than to data caching and flow control, trading single-thread speed for parallel throughput and hiding memory latency with queued independent work.
ReportedSupported by the sources below, not yet editor-reviewed.
1 source1 retrieved & read
NVIDIA defined the GPU at the GeForce 256's launch on 11 October 1999 as a single-chip processor with integrated transform, lighting, triangle setup/clipping and rendering engines processing at least 10 million polygons per second — 17 million transistors on a TSMC 220 nm process, and the first consumer part to move geometry work off the CPU.
ReportedSupported by the sources below, not yet editor-reviewed.
1 source1 retrieved & read