hgpu.org » Aparapi
Muhammad Adnan, Faisal Aslam, Zubair Nawaz, Syed Mansoor Sarwar
Tags: Aparapi, Benchmarking, Computer science, CUDA, Java, Matrix multiplication, nVidia, nVidia GeForce GT 630 M, OpenCL, OpenGL, Package
January 6, 2018 by hgpu
Gabriel O. Trisca
Tags: Aparapi, Computer science, Machine learning, Neural networks, nVidia, nVidia GeForce GTX Titan, OpenCL, RNN, Seismology, Sensing, Signal denoising, Tesla K20, Thesis
March 20, 2016 by hgpu
Recent source codes
* * *
Most viewed papers (last 30 days)
- Architecture-Aware LLM Inference Optimization on AMD Instinct GPUs: A Comprehensive Benchmark and Deployment Study
- LLMQ: Efficient Lower-Precision LLM Training for Consumer GPUs
- AutoKernel: Autonomous GPU Kernel Optimization via Iterative Agent-Driven Search
- An Efficient Heterogeneous Co-Design for Fine-Tuning on a Single GPU
- DRTriton: Large-Scale Synthetic Data Reinforcement Learning for Triton Kernel Generation
- KernelFoundry: Hardware-aware evolutionary GPU kernel optimization
- MobileKernelBench: Can LLMs Write Efficient Kernels for Mobile Devices?
- CuTeGen: An LLM-Based Agentic Framework for Generation and Optimization of High-Performance GPU Kernels using CuTe
- Mixed-precision numerics in scientific applications: survey and perspectives
- True 4-Bit Quantized Convolutional Neural Network Training on CPU: Achieving Full-Precision Parity
* * *




