hgpu.org » nVidia GeForce GTX 620
Husheng Zhou
Tags: Computer science, CUDA, Deep learning, Heterogeneous systems, Neural networks, nVidia, nVidia GeForce GTX 480, nVidia GeForce GTX 620, nVidia GeForce GTX 660, nVidia Jetson TX2, nVidia Quadro 6000, Tesla K80, Thesis
May 12, 2019 by hgpu
Recent source codes
* * *
Most viewed papers (last 30 days)
- The Anatomy of a Triton Attention Kernel
- CudaForge: An Agent Framework with Hardware Feedback for CUDA Kernel Optimization
- Scalable GPU-Based Integrity Verification for Large Machine Learning Models
- INT v.s. FP: A Comprehensive Study of Fine-Grained Low-bit Quantization Formats
- An MLIR pipeline for offloading Fortran to FPGAs via OpenMP
- Enhancing Transformer Performance and Portability through Auto-tuning Frameworks
- KernelBand: Boosting LLM-based Kernel Optimization with a Hierarchical and Hardware-aware Multi-armed Bandit
- RDMA Point-to-Point Communication for LLM Systems
- A Study of Floating-Point Precision Tuning in Deep Learning Operators Implementations
- ProofWright: Towards Agentic Formal Verification of CUDA
* * *



