Skip to content

ci : use hf-jobs-cpu-xl runner in server sanitize workflow - #29297

Merged
ggerganov merged 1 commit into
masterfrom
gg/server-sanitize-cpu-xl
Sep 23, 2026
Merged

ggerganov merged 1 commit into
masterfrom
gg/server-sanitize-cpu-xl

Conversation

@ggerganov

@ggerganov ggerganov commented Sep 23, 2026 •

Copy link
Copy Markdown
Member

Overview

Update the Server (sanitize) workflow to use the hf-jobs-cpu-xl runner instead of hf-jobs-cpu-upgrade.

Additional information

Should fix https://github.com/ggml-org/llama.cpp/actions/runs/35821670631/job/107054637988#step:9:7781

Requirements

  • I have read and agree with the contributing guidelines
  • AI usage disclosure: YES. pi:llama.cpp/DeepSeek-V4-Flash-Vision-Exp

Assisted-by: pi:llama.cpp/DeepSeek-V4-Flash-Vision-Exp
@github-actions github-actions Bot added the devops improvements to build systems and github actions label Sep 23, 2026
@ggerganov

Copy link
Copy Markdown
Member Author

@CISC Seems like hf-jobs-cpu-xl is not recognized?

@CISC

CISC commented Sep 23, 2026 •

Copy link
Copy Markdown
Member

@CISC Seems like hf-jobs-cpu-xl is not recognized?

I see it in the README, but not sure where that name originated from, it's not listed here:
https://github.com/huggingface/huggingface_hub/blob/main/src/huggingface_hub/_space_api.py#L68

@ggerganov

Copy link
Copy Markdown
Member Author

Maybe we can move these to the new AMD runners.

@CISC

CISC commented Sep 23, 2026

Copy link
Copy Markdown
Member

Maybe we can move these to the new AMD runners.

A little awkward as we have to set up a container, but doable.

@ggerganov

Copy link
Copy Markdown
Member Author

@CISC Seems like hf-jobs-cpu-xl is not recognized?

I see it in the README, but not sure where that name originated from, it's not listed here: https://github.com/huggingface/huggingface_hub/blob/main/src/huggingface_hub/_space_api.py#L68

@abidlabs Does it make sense to add the cpu-xl hardware preset to the hf-jobs?

@abidlabs

Copy link
Copy Markdown

Ah this is a bug in the dispatcher, because in flavors.py we get the labels from SpaceHardware (which doesn't include cpu-xl or cpu-performance) instead of JobHardware I'll switch the map to JobHardware and then if you redeploy the ggml-org dispatcher, then this PR should pick up a runner as is I believe

@abidlabs

Copy link
Copy Markdown

Done: huggingface/jobs-actions#20

@ggerganov
ggerganov marked this pull request as ready for review September 23, 2026 17:59
@ggerganov
ggerganov requested a review from a team as a code owner September 23, 2026 17:59
@ggerganov

Copy link
Copy Markdown
Member Author

Thanks, it works now!

@ggerganov
ggerganov merged commit 6e60f35 into master Sep 23, 2026
5 of 7 checks passed
@ggerganov
ggerganov deleted the gg/server-sanitize-cpu-xl branch September 23, 2026 18:30
@CISC

CISC commented Sep 24, 2026

Copy link
Copy Markdown
Member

The jobs seem to hang and time out...

@ggerganov

Copy link
Copy Markdown
Member Author

My guess is that running sanitized server builds in parallel starves the machine from resources. Let's try to disable the pytest workers for this workflow. I'll open a PR.

frostyautumnleaf pushed a commit to frostyautumnleaf/llama.cpp that referenced this pull request Oct 5, 2026
…29297)

Assisted-by: pi:llama.cpp/DeepSeek-V4-Flash-Vision-Exp
edwardyoon pushed a commit to edwardyoon/focus-llama that referenced this pull request Oct 7, 2026
…29297)

Assisted-by: pi:llama.cpp/DeepSeek-V4-Flash-Vision-Exp
(cherry picked from commit 6e60f35)
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

devops improvements to build systems and github actions

Projects

None yet

Development

Successfully merging this pull request may close these issues.

4 participants