hgpu.org » nVidia GeForce RTX 5080
Zheng Li, Weiyan Wang, Ruiyuan Li, Chao Chen, Xianlei Long, Linjiang Zheng, Quanqing Xu, Chuanhui Yang
Tags: Algorithms, Compression, Computer science, CUDA, Heterogeneous systems, nVidia, nVidia GeForce RTX 5080, Package
November 16, 2025 by hgpu
Aaron Jarmusch, Nathan Graddon, Sunita Chandrasekaran
Tags: Benchmarking, Computer science, CUDA, HPC, nVidia, nVidia GeForce RTX 5080, nVidia H100, Performance, PTX
July 20, 2025 by hgpu
Recent source codes
* * *
Most viewed papers (last 30 days)
- Optimizing CUDA like a Human: Micro-Profiling Tools as Expert Surrogates for LLM-Based GPU Kernel Optimization
- AutoPass: Evidence-Guided LLM Agents for Compiler Performance Tuning
- daVinci-kernel: Co-Evolving Skill Selection, Summarization, and Utilization via RL for GPU Kernel Optimization
- Leveraging AI Ecosystem for Portable and Sustainable GPU Kernels in HPC
- Tangram: Hiding GPU Heterogeneity for Efficient LLM Parallelization
- Real FP4 Tensor-Core Code in Pure Rust on a Gaming GPU - with NVIDIA's Own Compiler
- UniCoder: Unified Visual-to-Code Generation via Symbolic Rewards and Reference-Guided Code Optimization
- Fearless Concurrency on the GPU
- SpecGen: Accelerating Agentic Kernel Optimization with Speculative Generation
- From Tokens to Regions: CUDA-Sensitive Instruction Tuning for GPU Kernel Generation
* * *




