hgpu.org » pyCUDA
Richard Schoonhoven, Ben van Werkhoven, Kees Joost Batenburg
Tags: AMD Radeon Instinct Mi50, ATI, Auto-Tuning, Benchmarking, Computer science, CUDA, nVidia, nVidia A100, nVidia GeForce GTX 1080 Ti, nVidia GeForce GTX Titan X, nVidia Titan RTX, OpenCL, Performance, pyCUDA, PyOpenCL, Tesla K20, Tesla P100, Tesla V100
October 9, 2022 by hgpu
Florencio Balboa Usabiaga, Blaise Delmotte, Aleksandar Donev
Tags: Condensed matter, CUDA, nVidia, Package, Physics, pyCUDA, Soft Condensed Matter
December 6, 2016 by hgpu
Recent source codes
* * *
Most viewed papers (last 30 days)
- Hand-Written PTX Tensor-Core GEMM Kernels: A Multi-Precision Study on NVIDIA L4
- Compiler-Grounded Hierarchical Diagnosis for LLM-Based Triton Kernel Optimization
- Spec Sheets Are Not Kernels: An ISA- and Source-Level Audit of INT8 Availability on NVIDIA Blackwell Ultra
- Harness Engineering for LLM-Driven GPU Kernel Generation
- Validation-Centric AI-Assisted GPU Porting of a 250,000+ Line Legacy Weather Simulation Code
- CAKE: Compiler-Agent Co-Design for Frontier Kernel Evolution
- FlashPDE: A Drop-In Fused Triton Operator Library for Neural PDE Solvers
- A Contract-Grade Verifier for LLM-Generated GPU Kernels, and a Native Blackwell Backward for the Gated-Linear-Recurrence Family
- NVIDIA-labs OO Agents: Native Python Object-Oriented Agents
- PortLBM: A Portable Lattice Boltzmann Tool Leveraging SYCL on AMD, NVIDIA, and Intel GPUs
* * *




