Tagged articles

SIMT

6 articles · Page 1 of 1
Architects' Tech Alliance
Architects' Tech Alliance
Aug 6, 2026 · Fundamentals

Demystifying NVIDIA GPU Core Architecture: From Basics to AI Performance

This article breaks down NVIDIA GPU fundamentals—contrasting GPU with CPU design, tracing CUDA’s evolution, detailing the hardware hierarchy from chips to SM units, explaining memory tiers, and presenting a step‑by‑step performance‑optimization methodology for AI training and inference workloads.

CUDAGPU architecturePerformance Optimization
0 likes · 22 min read
Demystifying NVIDIA GPU Core Architecture: From Basics to AI Performance
Tencent Cloud Developer
Tencent Cloud Developer
Sep 26, 2025 · Fundamentals

Why GPUs Really Matter: From Architecture Basics to CUDA Programming

This article explains why GPUs have become the preferred platform for high‑performance computing, covering Dennard scaling, GPU speed advantages, theoretical FLOPS calculations, CUDA programming examples like SAXPY, the SIMT execution model, instruction pipelines, and modern techniques for handling branch divergence and register bank conflicts.

CUDA programmingGPU architectureGPU performance
0 likes · 38 min read
Why GPUs Really Matter: From Architecture Basics to CUDA Programming
Tencent Technical Engineering
Tencent Technical Engineering
Mar 21, 2025 · Fundamentals

Fundamentals of GPU Architecture and Programming

The article explains GPU fundamentals—from the end of Dennard scaling and why GPUs excel in parallel throughput, through CUDA programming basics like the SAXPY kernel and SIMT versus SIMD execution, to the evolution of the SIMT stack, modern scheduling, and a three‑step core architecture design.

CUDAGPUGPU programming
0 likes · 42 min read
Fundamentals of GPU Architecture and Programming