hgpu.org » TensorFlow
Martin Schrimpf
Tags: Computer science, CUDA, Deep learning, Machine learning, nVidia, Performance, Python, TensorFlow
December 3, 2016 by hgpu
Mohammad Babaeizadeh, Iuri Frosio, Stephen Tyree, Jason Clemons, Jan Kautz
Tags: Computer science, CUDA, Deep learning, Neural networks, nVidia, nVidia GeForce GTX Titan X, TensorFlow
November 23, 2016 by hgpu
Bingchen Gong, Brendan Jou, Felix Yu, Shih-Fu Chang
Tags: Caffe, Computer science, CUDA, Deep learning, Neural networks, nVidia, Package, TensorFlow
November 8, 2016 by hgpu
Alexander G. de G. Matthews, Mark van der Wilk, Tom Nickson, Keisuke Fujii, Alexis Boukouvalas, Pablo Leon-Villagra, Zoubin Ghahramani, James Hensman
Tags: Computer science, CUDA, Deep learning, Machine learning, nVidia, nVidia GeForce GTX Titan X, Package, Python, TensorFlow
October 29, 2016 by hgpu
Recent source codes
* * *
Most viewed papers (last 30 days)
- DeepSeek-V4-Flash on AMD gfx90a: Correctness Recovery and Inference Performance Engineering
- Accelerating the Solving of Many Tiny General Linear Systems on GPUs: Application to Constitutive Laws
- AutoTuneBench: Trustworthy Measurement for Agent Auto-Tuning of LLM Serving Engines
- Hardware-Aware FP4 FlashAttention-4
- PrefixBench-H100: Characterizing Prefix Reuse and Time-to-First-Token in H100 LLM Serving
- Every Kernel Is a Join: Automatic Multi-GPU Parallelism for AI Computations in Einsummable
- pytest-gpu-proof: Enabling Cloud-CPU Continuous Integration for GPU Code with Local GPU Attestation
- Stencil Computation at the Intersection of AI and HPC
- Accelerating the Mitigation of LLM Inference Nondeterminism Across GPU Architectures
- MaxKernel: Agentic Kernel Generation for TPUs
* * *



