hgpu.org » Text mining
Yongpeng Zhang, Frank Mueller, Xiaohui Cui, Thomas Potok
March 11, 2011 by hgpu
M. D. Lieberman, J. Sankaranarayanan, H. Samet
December 12, 2010 by hgpu
Joseph M. Cavanagh, Thomas E. Potok, Xiaohui Cui
November 21, 2010 by hgpu
Recent source codes
* * *
Most viewed papers (last 30 days)
- DeepSeek-V4-Flash on AMD gfx90a: Correctness Recovery and Inference Performance Engineering
- Accelerating the Solving of Many Tiny General Linear Systems on GPUs: Application to Constitutive Laws
- AutoTuneBench: Trustworthy Measurement for Agent Auto-Tuning of LLM Serving Engines
- Hardware-Aware FP4 FlashAttention-4
- PrefixBench-H100: Characterizing Prefix Reuse and Time-to-First-Token in H100 LLM Serving
- Every Kernel Is a Join: Automatic Multi-GPU Parallelism for AI Computations in Einsummable
- pytest-gpu-proof: Enabling Cloud-CPU Continuous Integration for GPU Code with Local GPU Attestation
- Stencil Computation at the Intersection of AI and HPC
- Accelerating the Mitigation of LLM Inference Nondeterminism Across GPU Architectures
- MaxKernel: Agentic Kernel Generation for TPUs
* * *


