hgpu.org » StarPU
Lucas Leandro Nesi, Samuel Thibault, Luka Stanisic, Lucas Mello Schnorr
Tags: Computer science, CUDA, Heterogeneous systems, nVidia, nVidia GeForce GTX 1080 Ti, Performance, StarPU
September 1, 2019 by hgpu
Samuel Thibault
Tags: Computer science, CUDA, Distributed computing, Heterogeneous systems, nVidia, nVidia Quadro FX 5800, OpenCL, Operating systems, StarPU, Task scheduling, Tesla C2050, Tesla K20, Tesla M2075, Thesis
December 23, 2018 by hgpu
Dalal Sukkari, Hatem Ltaief, Mathieu Faverge, David Keyes
Tags: Algorithms, Benchmarking, Computer science, Factorization, Intel Xeon Phi, nVidia, StarPU, Task scheduling, Tesla K80, Tesla P100
September 21, 2017 by hgpu
Recent source codes
* * *
Most viewed papers (last 30 days)
- PTXBench: Benchmark and Adapt LLMs for GPU Kernel Optimization with Architecture-specific PTX
- Hand-Written PTX Tensor-Core GEMM Kernels: A Multi-Precision Study on NVIDIA L4
- Compiler-Grounded Hierarchical Diagnosis for LLM-Based Triton Kernel Optimization
- Benchmarking Confidential Computing Performance on NVIDIA Blackwell GPUs
- RealisticTritonBench: A Benchmark for Triton-Kernel Generation in Real-World AI Frameworks
- Concurrency Response of Plain Global Loads on the NVIDIA H100
- Spec Sheets Are Not Kernels: An ISA- and Source-Level Audit of INT8 Availability on NVIDIA Blackwell Ultra
- Harness Engineering for LLM-Driven GPU Kernel Generation
- Validation-Centric AI-Assisted GPU Porting of a 250,000+ Line Legacy Weather Simulation Code
- CAKE: Compiler-Agent Co-Design for Frontier Kernel Evolution
* * *



