ggml-cpu(s390x): guard VXE-only repack helpers - #28775
Merged
taronaeo merged 1 commit intoSep 13, 2026
Merged
Conversation
taronaeo
approved these changes
Sep 11, 2026
This was referenced Sep 13, 2026
taronaeo
added a commit
to taronaeo/llama.cpp-s390x
that referenced
this pull request
Sep 13, 2026
Signed-off-by: Aaron Teo <aaron.teo1@ibm.com>
CISC
approved these changes
Sep 13, 2026
taronaeo
added a commit
to taronaeo/llama.cpp-s390x
that referenced
this pull request
Sep 13, 2026
This reverts commit d464525. Signed-off-by: Aaron Teo <aaron.teo1@ibm.com>
taronaeo
added a commit
to taronaeo/llama.cpp-s390x
that referenced
this pull request
Sep 13, 2026
Signed-off-by: Aaron Teo <aaron.teo1@ibm.com> ggml-cpu: add unused macro to fix ci Signed-off-by: Aaron Teo <aaron.teo1@ibm.com> Revert "ggml-cpu: temporarily add ggml-org#28775 patch until its merged" This reverts commit d464525. Signed-off-by: Aaron Teo <aaron.teo1@ibm.com>
5 tasks done
ggerganov
pushed a commit
that referenced
this pull request
Sep 14, 2026
* tests: add non-vxe build to tests Signed-off-by: Aaron Teo <aaron.teo1@ibm.com> ggml-cpu: add unused macro to fix ci Signed-off-by: Aaron Teo <aaron.teo1@ibm.com> Revert "ggml-cpu: temporarily add #28775 patch until its merged" This reverts commit d464525. Signed-off-by: Aaron Teo <aaron.teo1@ibm.com> * ggml-cpu: revert back to upstream/master Signed-off-by: Aaron Teo <aaron.teo1@ibm.com> --------- Signed-off-by: Aaron Teo <aaron.teo1@ibm.com>
1 task done
pl752
pushed a commit
to pl752/llama.cpp
that referenced
this pull request
Sep 15, 2026
pl752
pushed a commit
to pl752/llama.cpp
that referenced
this pull request
Sep 15, 2026
* tests: add non-vxe build to tests Signed-off-by: Aaron Teo <aaron.teo1@ibm.com> ggml-cpu: add unused macro to fix ci Signed-off-by: Aaron Teo <aaron.teo1@ibm.com> Revert "ggml-cpu: temporarily add ggml-org#28775 patch until its merged" This reverts commit d464525. Signed-off-by: Aaron Teo <aaron.teo1@ibm.com> * ggml-cpu: revert back to upstream/master Signed-off-by: Aaron Teo <aaron.teo1@ibm.com> --------- Signed-off-by: Aaron Teo <aaron.teo1@ibm.com>
5 of 6 tasks
vaiju1981
pushed a commit
to vaiju1981/java-llama.cpp
that referenced
this pull request
Sep 15, 2026
Second and final chunk, reaching the latest upstream release b10909. One commit (#28164, "metal : single-source fusion table + fusion debug rework"): 26 files, 111 KiB, and zero files on the priority-ordered API-compatibility list. The change is confined to the Metal backend plus upstream's own test and CI scaffolding; the `tests/*` files are applied but never compiled here, since a FetchContent subproject builds with `LLAMA_BUILD_TESTS=OFF`. All ten patches apply at pristine b10909 and all four standing drop-checks report "still required". `0012`'s `tests/CMakeLists.txt` hunk was checked against this chunk's edit to the same file and does not collide. `0013` was filed upstream during this bump as ggml-org/llama.cpp#28775 and is approved but not yet merged, so it is still required here. Co-Authored-By: Claude Opus 5 <noreply@anthropic.com> Claude-Session: https://claude.ai/code/session_01AnNYn8W1xuVxVJtyL34GyH
vaiju1981
pushed a commit
to vaiju1981/java-llama.cpp
that referenced
this pull request
Sep 15, 2026
Upstream merged this project's own PR ggml-org/llama.cpp#28775 ("ggml-cpu(s390x): guard VXE-only repack helpers", commit 6978052), first tagged at b10948, so patches/0013 is now redundant: pristine b10948 already wraps vxe_dot_acc / vxe_splat_granule / vxe_fold in the #if defined(__VXE__) || defined(__VXE2__) guard the patch added, and `git apply` of the patch fails "does not apply" exactly as the fail-loud applier is designed to signal. Dropped, not refreshed, per the 0009 precedent at b10280. The rationale that outlives the patch -- why build-linux-s390x is a scalar (non-VXE) cross build, and why -DGGML_VXE=ON is the wrong response to a future VXE compile error -- moves into a "0013 was dropped at the b10948 bump" note in CLAUDE.md; the s390x job's own comment in publish.yml now points there. Verified against pristine b10948: patches 0001-0012 apply clean, 0013 fails; the three standing drop-checks still say "still required" (0001 no common_params_parse_main in common/arg.h; 0010 vocab_type still uncast at server-context.cpp:4554; 0012 bare splits[i] /= split_sum at llama-model.cpp:1491). A fresh `cmake -B build -DBUILD_TESTING=ON` configures clean through the real FetchContent path, stamping nine patches at head 5f436dddb (= b10948), with extraction unchanged at 138 CLI / 57 request / 15 trainer names. With the real s390x-linux-gnu-g++ cross toolchain, upstream's guarded repack.cpp compiles clean both with the CI job's scalar flags and with -mvx -mzvector -march=z15. The rest of the range touches no file on the API-compatibility review list: no common/, include/, tools/server/ or tools/mtmd/ changes, so the three mechanical server-contract greps have no input to compare. Co-Authored-By: Claude Opus 5 <noreply@anthropic.com> Claude-Session: https://claude.ai/code/session_01SuJySZsHGCaxFq5Jo5cEtc
quimmedes
pushed a commit
to quimmedes/cafe-llama.cpp
that referenced
this pull request
Sep 16, 2026
quimmedes
pushed a commit
to quimmedes/cafe-llama.cpp
that referenced
this pull request
Sep 16, 2026
* tests: add non-vxe build to tests Signed-off-by: Aaron Teo <aaron.teo1@ibm.com> ggml-cpu: add unused macro to fix ci Signed-off-by: Aaron Teo <aaron.teo1@ibm.com> Revert "ggml-cpu: temporarily add ggml-org#28775 patch until its merged" This reverts commit d464525. Signed-off-by: Aaron Teo <aaron.teo1@ibm.com> * ggml-cpu: revert back to upstream/master Signed-off-by: Aaron Teo <aaron.teo1@ibm.com> --------- Signed-off-by: Aaron Teo <aaron.teo1@ibm.com>
zsogitbe
pushed a commit
to zsogitbe/llama.cpp
that referenced
this pull request
Sep 17, 2026
zsogitbe
pushed a commit
to zsogitbe/llama.cpp
that referenced
this pull request
Sep 17, 2026
* tests: add non-vxe build to tests Signed-off-by: Aaron Teo <aaron.teo1@ibm.com> ggml-cpu: add unused macro to fix ci Signed-off-by: Aaron Teo <aaron.teo1@ibm.com> Revert "ggml-cpu: temporarily add ggml-org#28775 patch until its merged" This reverts commit d464525. Signed-off-by: Aaron Teo <aaron.teo1@ibm.com> * ggml-cpu: revert back to upstream/master Signed-off-by: Aaron Teo <aaron.teo1@ibm.com> --------- Signed-off-by: Aaron Teo <aaron.teo1@ibm.com>
This file contains hidden or bidirectional Unicode text that may be interpreted or compiled differently than what appears below. To review, open the file in an editor that reveals hidden Unicode characters.
Learn more about bidirectional Unicode characters
Sign up for free
to join this conversation on GitHub.
Already have an account?
Sign in to comment
Add this suggestion to a batch that can be applied as a single commit.This suggestion is invalid because no changes were made to the code.Suggestions cannot be applied while the pull request is closed.Suggestions cannot be applied while viewing a subset of changes.Only one suggestion per line can be applied in a batch.Add this suggestion to a batch that can be applied as a single commit.Applying suggestions on deleted lines is not supported.You must change the existing code in this line in order to create a valid suggestion.Outdated suggestions cannot be applied.This suggestion has been applied or marked resolved.Suggestions cannot be applied from pending reviews.Suggestions cannot be applied on multi-line comments.Suggestions cannot be applied while the pull request is queued to merge.Suggestion cannot be applied right now. Please check back later.
Overview
On s390x without VXE/VXE2, the repack helpers
vxe_dot_acc,vxe_splat_granuleandvxe_foldare still compiled even though their vector intrinsics are unavailable, breaking the build.This wraps the three helpers in the same
#if defined(__VXE__) || defined(__VXE2__)guard as their call sites.Additional information
Reported in #28667, ; @taronaeo confirmed the missing guards there.
Requirements