Skip to content
#

gpu-benchmark

Here are 42 public repositories matching this topic...

Preregistered, DOI-archived llama.cpp study: Qwen3.8-27B on one RTX 3090 (sm_86). draft-mtp at --spec-draft-n-max 2 is +59.8% [+57.0, +62.8] server-reported decode over 25 purposive prompts; draft-dflash (PR #27342) +51.9%, not separated from it. Decode energy -37.1%, uncalibrated. 23-25 of 25 prompts diverge from serial greedy at 1600 tokens.

  • Updated Sep 2, 2026
  • Python

**Kernel-V8** is a high-performance GPU benchmarking engine built on the WebGL2 API. By rendering a complex 8th-order **Mandelbulb** fractal in real-time, it generates intense arithmetic workloads to evaluate the stability, thermal throttling, and peak compute throughput of modern graphics hardware.

  • Updated Jan 12, 2026
  • JavaScript

Add this topic to your repo

To associate your repository with the gpu-benchmark topic, visit your repo's landing page and select "manage topics."

Learn more