Skip to content

cuda : Use byte strides for roll to allow non-contiguous ROLL operations - #29547

Merged
ggerganov merged 1 commit into
ggml-org:masterfrom
bertaye:cuda-roll-strides
Oct 8, 2026
Merged

ggerganov merged 1 commit into
ggml-org:masterfrom
bertaye:cuda-roll-strides

Conversation

@bertaye

@bertaye bertaye commented Sep 27, 2026 •

Copy link
Copy Markdown
Contributor

Overview

This PR allows cuda backend's roll operation to be performed on non-contiguous tensors by using byte strides.

Additional information

Before, the test-backends-ops ROLL testcase for CUDA backend was skipping permute case; after the fix it passes.

Requirements

@bertaye
bertaye requested a review from a team as a code owner September 27, 2026 23:49
@github-actions github-actions Bot added ggml changes relating to the ggml tensor library for machine learning CUDA Related to the CUDA backend labels Sep 27, 2026
@ggml-gh-bot

ggml-gh-bot Bot commented Sep 27, 2026

Copy link
Copy Markdown

Hi @bertaye, thanks for your contribution!

Per our contribution guidelines, the automated PR checker found the following issue(s) that need your attention:

  • PR Template not respected: Please respect the template when creating a new pull request. Make sure to fill out all required sections.

Please note that maintainers reserve the right to make final decisions on PRs. If you believe there is a mistake, please comment below.

@ggml-gh-bot ggml-gh-bot Bot added the draft PR will be changed to draft by github-actions bot label Sep 27, 2026
@github-actions
github-actions Bot marked this pull request as draft September 27, 2026 23:54
@github-actions github-actions Bot removed the draft PR will be changed to draft by github-actions bot label Sep 27, 2026
@bertaye
bertaye marked this pull request as ready for review September 28, 2026 05:37
@JohannesGaessler JohannesGaessler added the merge ready A maintainer can use this label to indicate that they consider the changes final and ready to merge. label Oct 7, 2026
@ggerganov
ggerganov merged commit d888016 into ggml-org:master Oct 8, 2026
14 checks passed
edwardyoon pushed a commit to edwardyoon/focus-llama that referenced this pull request Oct 8, 2026
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

CUDA Related to the CUDA backend ggml changes relating to the ggml tensor library for machine learning merge ready A maintainer can use this label to indicate that they consider the changes final and ready to merge.

Projects

None yet

Development

Successfully merging this pull request may close these issues.

3 participants