Skip to content

Fix cuda bundling - #13

Merged
Phqen1x merged 2 commits into
lemonadefrom
fix-cuda-bundling
Jun 5, 2026
Merged

Phqen1x merged 2 commits into
lemonadefrom
fix-cuda-bundling

Conversation

@Phqen1x

@Phqen1x Phqen1x commented Jun 5, 2026

Copy link
Copy Markdown
Owner

No description provided.

Phqen1x and others added 2 commits June 4, 2026 16:22
The ggml CUDA backend is built as a shared library plugin when
SD_BUILD_SHARED_LIBS=ON. Bundle it alongside the CUDA runtime libs so
that CUDA inference works out of the box without the plugin being
missing at runtime.

Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com>
Add -DGGML_CUDA=ON and -DCMAKE_CUDA_COMPILER explicitly to the Linux
CUDA cmake invocation so CUDA is actually compiled instead of silently
falling back to CPU when auto-detection fails.

Widen the ggml artifact collection from libggml-cuda.so only to
libggml*.so* (Linux) and ggml*.dll (Windows) to capture all ggml shared
libraries (core, base backend, cuda backend) regardless of where cmake
places them in the build tree.

Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com>
@Phqen1x
Phqen1x merged commit 2de7212 into lemonade Jun 5, 2026
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant