hgpu.org » nVidia B300
Zhuobin Huang, Kai Zhang, Weihao Cui, Hongshi Tan, Liang Luo, Christopher Dewan, Shen Li, Bingsheng He
Tags: AMD Radeon Instinct MI300X, ATI, Computer science, LLM, nVidia, nVidia B300, nVidia H100, Performance
September 28, 2026 by hgpu
Robert Hu
Tags: Computer science, CUDA, nVidia, nVidia B300, nVidia GB200, Package, PTX
September 14, 2026 by hgpu
Recent source codes
* * *
Most viewed papers (last 30 days)
- DeepSeek-V4-Flash on AMD gfx90a: Correctness Recovery and Inference Performance Engineering
- Accelerating the Solving of Many Tiny General Linear Systems on GPUs: Application to Constitutive Laws
- Hardware-Aware FP4 FlashAttention-4
- AutoTuneBench: Trustworthy Measurement for Agent Auto-Tuning of LLM Serving Engines
- PrefixBench-H100: Characterizing Prefix Reuse and Time-to-First-Token in H100 LLM Serving
- Every Kernel Is a Join: Automatic Multi-GPU Parallelism for AI Computations in Einsummable
- Stencil Computation at the Intersection of AI and HPC
- pytest-gpu-proof: Enabling Cloud-CPU Continuous Integration for GPU Code with Local GPU Attestation
- Accelerating the Mitigation of LLM Inference Nondeterminism Across GPU Architectures
- MaxKernel: Agentic Kernel Generation for TPUs
* * *



