Tags: Algorithms, Computer science, CUDA, Data parallelism, Metaheuristics, nVidia, nVidia GeForce GTX 680, Search
Tags: Algorithms, Computer science, CUDA, Metaheuristics, nVidia, nVidia GeForce GTX 285, Optimization, Overview, Particle swarm optimization, Tesla C1060, Tesla C2050
Tags: CUDA, Metaheuristics, nVidia, nVidia GeForce GTX 680
Tags: Algorithms, Computer science, Computer vision, CUDA, Image processing, Metaheuristics, nVidia, nVidia GeForce GTX 580
Tags: Algorithms, Computer science, CUDA, Metaheuristics, nVidia, Tesla C1060, Thesis
Tags: Algorithms, Computer science, CUDA, Metaheuristics, nVidia, Tesla C2050
Tags: Computer science, CUDA, Metaheuristics, nVidia, nVidia GeForce GTX 480, Package, Performance, Search
Tags: Artificial intelligence, Computer science, CUDA, Differential evolution, Evolutionary Computations, Metaheuristics, Neural and Evolutionary Computing, nVidia, nVidia GeForce GTS 450, Optimization, Package, Particle swarm optimization
Tags: Computer science, CUDA, Metaheuristics, nVidia, nVidia GeForce GT 420 M, OpenMP, Optimization, Path problems
Recent source codes
Most viewed papers (last 30 days)
- DeepSeek-V4-Flash on AMD gfx90a: Correctness Recovery and Inference Performance Engineering
- Accelerating the Solving of Many Tiny General Linear Systems on GPUs: Application to Constitutive Laws
- AutoTuneBench: Trustworthy Measurement for Agent Auto-Tuning of LLM Serving Engines
- Hardware-Aware FP4 FlashAttention-4
- PrefixBench-H100: Characterizing Prefix Reuse and Time-to-First-Token in H100 LLM Serving
- Every Kernel Is a Join: Automatic Multi-GPU Parallelism for AI Computations in Einsummable
- pytest-gpu-proof: Enabling Cloud-CPU Continuous Integration for GPU Code with Local GPU Attestation
- Stencil Computation at the Intersection of AI and HPC
- Accelerating the Mitigation of LLM Inference Nondeterminism Across GPU Architectures
- MaxKernel: Agentic Kernel Generation for TPUs



