unsloth/studio
Daniel Han bb3676d2c8 llama.cpp CUDA detection: handle dlopen-ed backend (split build layout)
Current llama.cpp ships the CUDA backend as a dynamically-loaded plugin
(libggml-cuda.so* next to the binary), NOT a load-time dependency, so
ldd llama-server | grep libggml-cuda is a false negative: it reports no
CUDA on a perfectly good CUDA build. That made both is_cuda_server()
(provision_llama_cuda.sh) and _have_cuda_llama_server() (setup.sh) force a
needless full rebuild every run.

Fix both: keep the ldd check (old monolithic builds) and additionally treat
the presence of libggml-cuda.so* beside the binary as the CUDA signal. A
CPU-only build has no such backend, so this stays correct for the CPU case.

Verified on an N1X/sm_121 WSL build: llama-server --list-devices shows
CUDA0 JMJWOA-Generic-GPU and serves on the GPU, while ldd lists no
libggml-cuda; the new check correctly returns CUDA-present.

Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com>
2026-06-03 03:16:06 -07:00
..
backend Configurable upload Cap studio (for training) (#5808) 2026-06-02 08:52:19 -07:00
frontend Guard model-load success path against mid-refresh cancellation (#5944) 2026-06-03 10:18:57 +04:00
scripts llama.cpp CUDA detection: handle dlopen-ed backend (split build layout) 2026-06-03 03:16:06 -07:00
src-tauri Studio: persist Tauri window size and maximized state across launches (#5799) 2026-06-02 09:26:15 -07:00
__init__.py Final cleanup 2026-03-12 18:28:04 +00:00
install_llama_prebuilt.py Studio: cover B300 (sm_103) with the Linux prebuilt bundles (#5930) 2026-06-01 08:17:50 -07:00
install_python_stack.py studio: ROCm cleanups follow-up to #5301 (#5874) 2026-05-30 03:06:47 -07:00
LICENSE.AGPL-3.0 Add AGPL-3.0 license to studio folder 2026-03-09 19:36:25 +00:00
package-lock.json ci: advisory lockfile supply-chain audit (no install-script changes) (#5604) 2026-05-19 05:56:56 -07:00
package.json ci: advisory lockfile supply-chain audit (no install-script changes) (#5604) 2026-05-19 05:56:56 -07:00
setup.bat Final cleanup 2026-03-12 18:28:04 +00:00
setup.ps1 Studio: forward the resolved AMD gfx arch to the prebuilt installer (#5923) 2026-06-01 06:55:35 -07:00
setup.sh llama.cpp CUDA detection: handle dlopen-ed backend (split build layout) 2026-06-03 03:16:06 -07:00
Unsloth_Studio_Colab.ipynb Fix/studio colab proxy and iframe - Unsloth Studio not loading in Colab (iframe "refused to connect" and wrong URL) (#5844) 2026-05-28 23:54:48 -07:00