Commit graph

719 commits

Author SHA1 Message Date
Roland Tannous
5f98d232d0 fix: align llama-server binary discovery with upstream unsloth-zoo paths 2026-03-03 17:03:01 +00:00
Roland Tannous
7f339b5c97 Patch unsloth-zoo llama_cpp.py and unsloth save.py from windows-support branch 2026-03-02 10:45:09 +00:00
Roland Tannous
e7619a1291 Move llama.cpp clone/build from in-tree to ~/.unsloth/llama.cpp
- setup.sh: builds at ~/.unsloth/llama.cpp instead of ./llama.cpp
- setup.ps1: builds at %USERPROFILE%/.unsloth/llama.cpp
- inference llama_cpp.py: searches ~/.unsloth/ first, in-tree as legacy
- export.py: updated comments (unsloth-zoo handles path natively)
2026-03-02 04:04:41 +00:00
Roland Tannous
204123a212 Tighten Python bounds to >= 3.11, < 3.14 (matching setup.sh), only auto-install if missing 2026-03-01 13:05:10 +00:00
Roland Tannous
33be2ac093 Add Python 3.12 prerequisite check with auto-install via winget 2026-03-01 13:05:10 +00:00
Roland Tannous
148e93cb83 Auto-enable Windows Long Paths via UAC elevation during setup 2026-03-01 13:05:10 +00:00
Roland Tannous
77d1378d04 Remove unused CMP0194 cmake policy (eliminates cmake warning) 2026-03-01 13:05:10 +00:00
Roland Tannous
6d606cd18c Simplify: use winget OpenSSL.Dev instead of vcpkg for HTTPS support 2026-03-01 13:05:10 +00:00
Roland Tannous
2a9aba5017 Add vcpkg/curl[ssl] for HTTPS support in llama-server, enable LLAMA_CURL=ON 2026-03-01 13:05:10 +00:00
Roland Tannous
70d1567fe3 Download GGUF via huggingface_hub instead of llama-server -hf (fixes HTTPS not supported on Windows) 2026-03-01 13:05:10 +00:00
Roland Tannous
2a5e03945f Add .venv/Scripts to User PATH so unsloth-studio works without activation 2026-03-01 13:05:10 +00:00
Roland Tannous
9b6a928a78 Simplify completion banner: no venv activation needed 2026-03-01 13:05:10 +00:00
Roland Tannous
0430d22cc2 Auto-add CUDA DLLs to PATH when launching llama-server on Windows 2026-03-01 13:05:10 +00:00
Roland Tannous
e956fb4106 Warn user to uninstall incompatible CUDA toolkit instead of failed side-by-side 2026-03-01 13:05:10 +00:00
Roland Tannous
c182b8c438 Fallback: try descending CUDA versions if exact driver-max install fails 2026-03-01 13:05:10 +00:00
Roland Tannous
d965a51b70 Always persist compatible CUDA_PATH to User registry (overwrite stale values) 2026-03-01 13:05:10 +00:00
Roland Tannous
f80cef3abe Fix: scan side-by-side CUDA installs, pick compatible toolkit version 2026-03-01 13:05:10 +00:00
Roland Tannous
afa1344452 Build llama.cpp in-tree, auto-detect driver CUDA version for compatible toolkit 2026-03-01 13:05:10 +00:00
Roland Tannous
7377c92a4c Auto-detect driver CUDA version, install compatible toolkit instead of latest 2026-03-01 13:05:10 +00:00
Roland Tannous
18134ee808 Fix non-ASCII chars in test script for Windows PS 5.1 2026-03-01 13:05:10 +00:00
Roland Tannous
4ee5ada609 Add llama-cpp Windows test script, fix binary lookup paths 2026-03-01 13:05:10 +00:00
Roland Tannous
af90c9c3d2 Fix llama-server binary lookup for Windows (.exe, Release dir, ~/.unsloth) 2026-03-01 13:05:10 +00:00
Roland Tannous
c297d7aa84 Force num_proc=1 on Windows to avoid slow spawn overhead 2026-03-01 13:05:10 +00:00
Roland Tannous
602ae716c3 Set short TORCHINDUCTOR_CACHE_DIR to fix Windows MAX_PATH crash 2026-03-01 13:05:10 +00:00
Roland Tannous
c7fed1ba83 Fix npm Invalid Version: delete package-lock.json, relax Node constraint 2026-03-01 13:05:10 +00:00
Roland Tannous
e6e97ad4c1 Enforce Node LTS (v20-v22), add npm error checking, clean node_modules 2026-03-01 13:05:10 +00:00
Roland Tannous
721088047c Fix npm stderr crash on Windows ErrorActionPreference 2026-03-01 13:05:10 +00:00
Roland Tannous
ccfc00944f Fix Windows frontend build, add setup.bat, ANSI colors, aliases 2026-03-01 13:05:10 +00:00
Roland Tannous
9dc7af08a0 add setup.bat 2026-03-01 13:05:10 +00:00
Roland Tannous
26996bc609 Extract shared install_python_stack.py for cross-platform setup 2026-03-01 13:05:10 +00:00
Roland Tannous
46641bb872 Merge pull request #295 from unslothai/fix/fix-local-vision-gguf-loading
fix: support mmproj for local vision GGUF models + fix Windows pipe d…
2026-03-01 17:03:24 +04:00
Roland Tannous
986bef4f99 fix: support mmproj for local vision GGUF models + fix Windows pipe deadlock 2026-03-01 12:58:38 +00:00
Shine1i
50385c1eae feat(markdown): fix Mermaid integration with error handling and copy button 2026-02-27 20:02:40 +01:00
Roland Tannous
ed97809873 Merge pull request #284 from unslothai/fix/drop-model-task-hard-filter
Apply HF task filtering only for empty model queries
2026-02-27 15:46:17 +04:00
imagineer99
7682760581 fix: only apply HF task filter for empty model search queries 2026-02-27 11:31:53 +00:00
Wasim Yousef Said
d90ada2af2 Merge pull request #268 from unslothai/fix/delete-custom-config
Updated the delete custom preset in chat tab (filters)
2026-02-27 01:38:01 -08:00
Shine1i
c24bd48deb keep checkpoint on preset apply 2026-02-27 10:36:07 +01:00
Wasim Yousef Said
673f951234 Merge pull request #269 from unslothai/fix/fine-tuned-model-tooltip
Added tooltip for the fine tuned models in chat page
2026-02-27 01:34:49 -08:00
Wasim Yousef Said
e9d4e2871d Merge pull request #262 from unslothai/fix/truncated-text
Reduced name truncation on the training page
2026-02-27 01:33:05 -08:00
Shine1i
e8f832710c merge nightly 2026-02-27 10:31:37 +01:00
Roland Tannous
d0231abfea Merge pull request #267 from unslothai/fix/config-switch
Preserving model name during configuration type switch in chat page
2026-02-27 13:23:29 +04:00
Roland Tannous
f6bb12ccc1 Merge pull request #266 from unslothai/fix/dataset-search-remove-size-download-badges
Remove dataset metadata badges from HF dataset dropdowns
2026-02-27 13:22:49 +04:00
Roland Tannous
d6922f5e83 Merge pull request #282 from unslothai/fix/inference-auth
Added auth to inference endpoints
2026-02-27 13:18:38 +04:00
Roland Tannous
e51ec8a30a Merge pull request #278 from unslothai/fix/reorder-model-type-cards-onboarding
Reorder model type cards in onboarding to show Text first
2026-02-27 13:17:00 +04:00
Roland Tannous
c45fd1b0a9 Merge pull request #275 from unslothai/fix/show-size-gguf
Passes metadata to get model size
2026-02-27 13:15:38 +04:00
Manan17
b4311cca82 Aggregating sharded models, showing fit/oom for quantizations 2026-02-27 08:23:15 +00:00
samit
6a9969d67b added auth to inference endpoints 2026-02-27 00:20:36 -08:00
Manan17
4bd5213c05 Passes metadata to get model size 2026-02-27 07:38:55 +00:00
Roland Tannous
bbf1b9ab59 Merge pull request #276 from unslothai/fix/rebuild-llamacpp-setup
rebuild llama cpp for setup
2026-02-27 10:20:32 +04:00
imagineer99
9befc653ac fix: reorder model type cards in onboarding to show Text first 2026-02-27 06:16:26 +00:00