hgpu.org » nVidia GeForce 840 M
Dominik Größler
Tags: Benchmarking, Computer science, CUDA, nVidia, nVidia A100, nVidia GeForce 840 M, nVidia GeForce RTX 2080 Ti, nVidia Quadro P 6000, Package, Performance, PTX, Tesla K20, Tesla V100, Thesis
November 13, 2022 by hgpu
Paul Springer, Aravind Sankaran, Paolo Bientinesi
Tags: BLAS, Compilers, Computer science, CUDA, Intel Xeon Phi, Linear Algebra, Mathematical Software, nVidia, nVidia GeForce 840 M, Package, Performance, Tesla K40
July 8, 2016 by hgpu
Recent source codes
* * *
Most viewed papers (last 30 days)
- Optimizing CUDA like a Human: Micro-Profiling Tools as Expert Surrogates for LLM-Based GPU Kernel Optimization
- AutoPass: Evidence-Guided LLM Agents for Compiler Performance Tuning
- daVinci-kernel: Co-Evolving Skill Selection, Summarization, and Utilization via RL for GPU Kernel Optimization
- Leveraging AI Ecosystem for Portable and Sustainable GPU Kernels in HPC
- Tangram: Hiding GPU Heterogeneity for Efficient LLM Parallelization
- Real FP4 Tensor-Core Code in Pure Rust on a Gaming GPU - with NVIDIA's Own Compiler
- UniCoder: Unified Visual-to-Code Generation via Symbolic Rewards and Reference-Guided Code Optimization
- Fearless Concurrency on the GPU
- SpecGen: Accelerating Agentic Kernel Optimization with Speculative Generation
- From Tokens to Regions: CUDA-Sensitive Instruction Tuning for GPU Kernel Generation
* * *




