hgpu.org » StarPU
Lucas Leandro Nesi, Samuel Thibault, Luka Stanisic, Lucas Mello Schnorr
Tags: Computer science, CUDA, Heterogeneous systems, nVidia, nVidia GeForce GTX 1080 Ti, Performance, StarPU
September 1, 2019 by hgpu
Samuel Thibault
Tags: Computer science, CUDA, Distributed computing, Heterogeneous systems, nVidia, nVidia Quadro FX 5800, OpenCL, Operating systems, StarPU, Task scheduling, Tesla C2050, Tesla K20, Tesla M2075, Thesis
December 23, 2018 by hgpu
Dalal Sukkari, Hatem Ltaief, Mathieu Faverge, David Keyes
Tags: Algorithms, Benchmarking, Computer science, Factorization, Intel Xeon Phi, nVidia, StarPU, Task scheduling, Tesla K80, Tesla P100
September 21, 2017 by hgpu
Recent source codes
* * *
Most viewed papers (last 30 days)
- Optimizing CUDA like a Human: Micro-Profiling Tools as Expert Surrogates for LLM-Based GPU Kernel Optimization
- AutoPass: Evidence-Guided LLM Agents for Compiler Performance Tuning
- daVinci-kernel: Co-Evolving Skill Selection, Summarization, and Utilization via RL for GPU Kernel Optimization
- Leveraging AI Ecosystem for Portable and Sustainable GPU Kernels in HPC
- Tangram: Hiding GPU Heterogeneity for Efficient LLM Parallelization
- Real FP4 Tensor-Core Code in Pure Rust on a Gaming GPU - with NVIDIA's Own Compiler
- UniCoder: Unified Visual-to-Code Generation via Symbolic Rewards and Reference-Guided Code Optimization
- Fearless Concurrency on the GPU
- SpecGen: Accelerating Agentic Kernel Optimization with Speculative Generation
- From Tokens to Regions: CUDA-Sensitive Instruction Tuning for GPU Kernel Generation
* * *



