hgpu.org » GPGPU architecture
Sparsh Mittal
Tags: GPGPU, GPGPU architecture, GPU, Hardware, Hardware Architecture
November 18, 2014 by sparsh0mittal
Jaewoong Sim, Aniruddha Dasgupta, Hyesoon Kim, and Richard Vuduc
Tags: Analytical model, CUDA, GPGPU architecture, nVidia, Performance benefit prediction, Performance prediction, Tesla C2050
March 30, 2012 by Moaddeli
Recent source codes
* * *
Most viewed papers (last 30 days)
- UniCoder: Unified Visual-to-Code Generation via Symbolic Rewards and Reference-Guided Code Optimization
- Real FP4 Tensor-Core Code in Pure Rust on a Gaming GPU - with NVIDIA's Own Compiler
- CuFuzz: An API-Knowledge-Graph Coverage-Driven Fuzzing Framework for CUDA Libraries
- Enhancing the Performance Analysis of NCCL GPU Collectives
- Augmenting LLM Code Translation with Compiler Analysis for C to Triton Kernel Generation
- Compiler-Grounded Hierarchical Diagnosis for LLM-Based Triton Kernel Optimization
- Harness Engineering for LLM-Driven GPU Kernel Generation
- FlashPDE: A Drop-In Fused Triton Operator Library for Neural PDE Solvers
- NVIDIA-labs OO Agents: Native Python Object-Oriented Agents
- PortLBM: A Portable Lattice Boltzmann Tool Leveraging SYCL on AMD, NVIDIA, and Intel GPUs
* * *



