#
fp8
Here are 4 public repositories matching this topic...
An stress and benchmark utility for NVIDIA GPUs. Measures performance across various precisions (FP64, FP32, TF32, FP16, INT8) and monitors real-time vitals like power, temperature, and clock speeds.
benchmarking performance sparsity cpp hpc cuda nvidia stress-testing fp16 int8 gpu-benchmark fp32 fp64 fp8 fp4 tf32 tf16
-
Updated
Dec 12, 2025 - C++
Hardware-focused llama.cpp fork for Windows and dual AMD RDNA4 RX 9070 XT GPUs: ROCm/HIP + Vulkan backends, PyQt6 GUI (RDNA LLM Studio), MTP speculative decoding, FP8 attention, long-context Qwen3.8 benchmarks
windows benchmark amd vulkan rpc gpu-acceleration hip mtp rocm fp8 llama-cpp ggml llm-inference qwen agentic-ai rdna4 qwen3 gfx1201 rx-9070-xt
-
Updated
Sep 16, 2026 - C++
FP8 dtypes enumeration in python
-
Updated
Nov 16, 2023 - C++
Add this topic to your repo
To associate your repository with the fp8 topic, visit your repo's landing page and select "manage topics."