Low-latency NUMA-aware fork-join thread-pool with zero allocations, syscalls, CAS, or false-sharing on the hot path for C, C++, Rust, and Zig 🍴
arm x86-64 concurrency openmp mpi parallel-computing multithreading parallelism thread-pool numa memory-model risc-v threadpool rayon atomics compare-and-swap parallel-stl
-
Updated
Sep 20, 2026 - C++