hgpu.org » AMD Radeon R9 M370X
Cedric Nugteren
Tags: AMD Radeon R9 M370X, ARM, ATI, BLAS, Computer science, Intel HD 5100, Linear Algebra, Machine learning, nVidia, nVidia GeForce GTX 750 Ti, nVidia GeForce GTX Titan X, OpenCL, Package
May 18, 2017 by hgpu
Recent source codes
* * *
Most viewed papers (last 30 days)
- Optimizing CUDA like a Human: Micro-Profiling Tools as Expert Surrogates for LLM-Based GPU Kernel Optimization
- AutoPass: Evidence-Guided LLM Agents for Compiler Performance Tuning
- daVinci-kernel: Co-Evolving Skill Selection, Summarization, and Utilization via RL for GPU Kernel Optimization
- Leveraging AI Ecosystem for Portable and Sustainable GPU Kernels in HPC
- Tangram: Hiding GPU Heterogeneity for Efficient LLM Parallelization
- Real FP4 Tensor-Core Code in Pure Rust on a Gaming GPU - with NVIDIA's Own Compiler
- UniCoder: Unified Visual-to-Code Generation via Symbolic Rewards and Reference-Guided Code Optimization
- Fearless Concurrency on the GPU
- SpecGen: Accelerating Agentic Kernel Optimization with Speculative Generation
- From Tokens to Regions: CUDA-Sensitive Instruction Tuning for GPU Kernel Generation
* * *




