RTX 5090 via NVIDIA OpenCL: 96/96 PASS, batch fingerprint identical to AMD
Co-Authored-By: Claude Fable 5.1 <noreply@anthropic.com>
This commit is contained in:
parent
5dc4414cc1
commit
731a73269e
1 changed files with 12 additions and 0 deletions
|
|
@ -197,3 +197,15 @@ Pack igneum-genesis-mh, memory-hard dataset, 1024 MiB, exchange via local memory
|
|||
| Hash rate | 4.38 Mhash/s on one compute unit, 1.82 GB/s useful |
|
||||
|
||||
Reading: the third GPU vendor. The same memory-hard program now produces identical hashes on Apple Metal, NVIDIA CUDA, Apple OpenCL and AMD OpenCL, cache and dataset included. The AMD number is from a two-CU integrated chip sharing system memory and is a correctness result only; the discrete AMD card is still to come. The local-memory exchange path, which wave64 cards will also use, is now proven on AMD silicon.
|
||||
|
||||
## 3 October 2026, RTX 5090 through NVIDIA OpenCL (fourth compiler path on the same card)
|
||||
|
||||
Pack igneum-genesis-mh, 1024 MiB, local-memory exchange (NVIDIA's OpenCL lists no sub-group shuffle extension).
|
||||
|
||||
| Check | Result |
|
||||
|---|---|
|
||||
| Cache check and dataset self-test | PASS |
|
||||
| Vectors | 96/96 PASS, batch fingerprint 98af644e993239e2, identical to the AMD gfx1036 run |
|
||||
| Hash rate | 219.6 Mhash/s via OpenCL against 229.0 via CUDA, about 4% apart, approximate |
|
||||
|
||||
Reading: NVIDIA's OpenCL compiler and NVIDIA's CUDA compiler agree with each other, with AMD's OpenCL, with Apple's Metal and OpenCL, and with the CPU reference. The batch fingerprint over 16.7 million consecutive nonces is identical on the AMD integrated chip and the 5090, which is a far stronger statement than the 96 vectors alone. The local-memory exchange costs about 4% against CUDA's warp shuffle on this card, approximate.
|
||||
|
|
|
|||
Loading…
Reference in a new issue