hgpu.org » Particle swarm optimization
M. P. Wachowiak, A. E. Lambe Foster
Tags: Biological Physics, Computational Physics, CUDA, nVidia, Particle swarm optimization, Physics, Tesla S1070
October 25, 2012 by hgpu
Xiangzhen Li, Jingying Hu, Weiguo Lv, Guiqiang Wang, Xinrong Cai
Tags: ATI, ATI Radeon HD 6760, Computer science, DSP, Heterogeneous systems, OpenCL, Particle swarm optimization, Task scheduling
October 20, 2012 by hgpu
Daniel Leal Souza, Otavio Noura Teixeira, Dionne Cavalcante Monteiro, Roberto Celio Limao de Oliveira
Tags: Algorithms, Computer science, CUDA, nVidia, nVidia GeForce GT 330 M, Optimization, Particle swarm optimization
August 15, 2012 by hgpu
Recent source codes
* * *
Most viewed papers (last 30 days)
- DeepSeek-V4-Flash on AMD gfx90a: Correctness Recovery and Inference Performance Engineering
- AutoTuneBench: Trustworthy Measurement for Agent Auto-Tuning of LLM Serving Engines
- Accelerating the Solving of Many Tiny General Linear Systems on GPUs: Application to Constitutive Laws
- PrefixBench-H100: Characterizing Prefix Reuse and Time-to-First-Token in H100 LLM Serving
- Hardware-Aware FP4 FlashAttention-4
- Every Kernel Is a Join: Automatic Multi-GPU Parallelism for AI Computations in Einsummable
- pytest-gpu-proof: Enabling Cloud-CPU Continuous Integration for GPU Code with Local GPU Attestation
- Accelerating the Mitigation of LLM Inference Nondeterminism Across GPU Architectures
- Stencil Computation at the Intersection of AI and HPC
- Microarchitectural Memory Bandwidth Saturation, KV-Cache Paging Dynamics, and Time-to-First-Token Latency: A Comparative Benchmark of vLLM, TensorRT-LLM, and FlashAttention-3 on NVIDIA Hopper H100 versus AMD Instinct MI300X
* * *


