hgpu.org » Go
Derek L. Stinson
Tags: Computer science, CUDA, Deep learning, Go, nVidia, nVidia GeForce GTX 1080 Ti, Package, Thesis
May 17, 2020 by hgpu
Mirko Mariotti, Loriano Storchi, Daniele Spiga, Davide Salomoni, Tommaso Boccali, Daniele Bonacorsi
Tags: Computer science, FPGA, Go, Machine learning
December 8, 2019 by hgpu
Recent source codes
* * *
Most viewed papers (last 30 days)
- The Anatomy of a Triton Attention Kernel
- CudaForge: An Agent Framework with Hardware Feedback for CUDA Kernel Optimization
- Scalable GPU-Based Integrity Verification for Large Machine Learning Models
- INT v.s. FP: A Comprehensive Study of Fine-Grained Low-bit Quantization Formats
- An MLIR pipeline for offloading Fortran to FPGAs via OpenMP
- Enhancing Transformer Performance and Portability through Auto-tuning Frameworks
- KernelBand: Boosting LLM-based Kernel Optimization with a Hierarchical and Hardware-aware Multi-armed Bandit
- RDMA Point-to-Point Communication for LLM Systems
- A Study of Floating-Point Precision Tuning in Deep Learning Operators Implementations
- ProofWright: Towards Agentic Formal Verification of CUDA
* * *




