…ind -DONEBIT_GPU_PRIVATE
Our GPU kernel work (ggml-hrx fixes, the HRX prefill split, ROCmI4 on Vulkan) is
closed source, like the NPU kernels. The public build now pins:
- third_party/llama.cpp: AMD-Ecosystem/llama.cpp hrx-graph-develop-v2 f1a0aca and
third_party/hrx-system a351789, the pair ROCm/ggml-staging-automation tests, unchanged;
- third_party/llama.cpp-rocmfpx: charlie12345/ROCmFPX main fb08d7c.
-DONEBIT_GPU_PRIVATE=<gpu-kernels checkout> includes its addons/gpu/gpu.cmake, which
names the llama.cpp, hrx-system and ROCmFPX trees ONEBIT_HRX and ONEBIT_LEAN build
instead. Without it `1bit serve --prefill-device hrx` exits with a message, and the
serve_e2e_vulkan_prefill_hrx test is not added.
bump-hrx.yml and bump-rocmfpx.yml now only move the public pins (no fork sync or
rebase). Docs keep outcome-level statements of what each build can do.
Verified on Strix Halo (serve_e2e): public: Qwen3-0.6B hrx PASS, vulkan PASS,
--prefill-device hrx refused; Qwen3.6-35B-A3B Q8_0 hrx fails (compute error).
Private: Qwen3-0.6B hrx, vulkan, vulkan --prefill-device hrx PASS; Qwen3.6-35B-A3B
Q8_0 hrx PASS.
Co-Authored-By: Claude Opus 5.5 <noreply@anthropic.com>
Our GPU kernel work (ggml-hrx fixes, the HRX prefill split, ROCmI4 on Vulkan) becomes closed source, like the NPU kernels (#85). The engine stops pinning our public forks.
Public pins
third_party/llama.cpp1bit/hrx-vulkan-patchedc075cc1hrx-graph-develop-v2f1a0aca(AMD's pin, unchanged)third_party/hrx-system51b1739a351789(AMD's current pair)third_party/llama.cpp-rocmfpx1bit/vulkan-rocmi4b44bf74mainfb08d7cHook:
-DONEBIT_GPU_PRIVATE=<gpu-kernels checkout>includesaddons/gpu/gpu.cmakefrom the private repo, which names the llama.cpp / hrx-system / ROCmFPX treesONEBIT_HRXandONEBIT_LEANbuild instead. Without it,1bit serve --prefill-device hrxexits with a message andserve_e2e_vulkan_prefill_hrxis not added.Workflows:
bump-hrx.ymlandbump-rocmfpx.ymlnow only move the public pins (no fork sync, no rebase;HRX_BUMP_TOKENneeds only engine access). The private repo keeps its own pins; a bump flow there is still to do.Docs: hrx.md (public vs private table, outcome level only; the "Fixed"/"Our patches" internals removed), lean.md, serve.md, PORTING.md, README status, NOTICE.
Verified on Strix Halo (
tests/serve_e2e.sh, builds with-DONEBIT_HRX=ON -DONEBIT_LEAN=ON):-DONEBIT_GPU_PRIVATEAfter merge, the fork branches
1bit-MONSTER/llama.cpp1bit/hrx-vulkan-patched/1bit/hrx-vulkanand1bit-MONSTER/ROCmFPX1bit/vulkan-rocmi4are no longer used by the engine (private copies are in gpu-kernels).🤖 Generated with Claude Code