Tags: Computer science, CUDA, Memory, nVidia, Operating systems, Performance, Tesla V100
Tags: Computer science, CUDA, Heterogeneous systems, nVidia, OpenCL, Operating systems, PTX, SYCL, Thesis
Tags: Computer science, CUDA, Distributed computing, Heterogeneous systems, nVidia, nVidia Quadro FX 5800, OpenCL, Operating systems, StarPU, Task scheduling, Tesla C2050, Tesla K20, Tesla M2075, Thesis
Tags: Computer science, CUDA, nVidia, nVidia Jetson TK1, Operating systems, Performance, Security, SoC
Performance Evaluation of Container-based Virtualization for High Performance Computing Environments
Tags: Benchmarking, Computer science, CUDA, MPI, nVidia, Operating systems, Package, Performance, Tesla K20, Virtualization
Tags: Computer science, CUDA, Genetic programming, nVidia, nVidia GRID K520, Operating systems, Package
Tags: ATI, C++ AMP, Computer science, Operating systems
Tags: AMD Radeon R7 250, ATI, Cloud, Computer science, nVidia, nVidia GeForce GTX 750, OpenCL, Operating systems, Security
Recent source codes
Most viewed papers (last 30 days)
- Revealing NVIDIA Closed-Source Driver Command Streams for CPU-GPU Runtime Behavior Insight
- Evaluating CUDA Tile for AI Workloads on Hopper and Blackwell GPUs
- DITRON: Distributed Multi-level Tiling Compiler for Parallel Tensor Programs
- FACT: Compositional Kernel Synthesis with a Three-Stage Agentic Workflow
- CuBridge: An LLM-Based Framework for Understanding and Reconstructing High-Performance Attention Kernels
- CUDAHercules: Benchmarking Hardware-Aware Expert-level CUDA Optimization for LLMs
- KEET: Explaining Performance of GPU Kernels Using LLM Agents
- ARGUS: Agentic GPU Optimization Guided by Data-Flow Invariants
- Kerncap: Automated Kernel Extraction and Isolation for AMD GPUs
- A Human–Machine Collaborative Tuning Framework for Triton Kernel Optimization on SIMD Platforms




