Commit graph

720 commits

Author SHA1 Message Date
Roland Tannous
58b00db5cb chore: add cross-platform Python installer with updated unsloth patch URLs 2026-03-03 17:31:59 +00:00
Roland Tannous
a4d2853fbc fix: align llama-server binary discovery with upstream unsloth-zoo paths 2026-03-03 17:03:01 +00:00
Roland Tannous
c64e50b46f Patch unsloth-zoo llama_cpp.py and unsloth save.py from windows-support branch 2026-03-02 10:45:09 +00:00
Roland Tannous
e280e457d1 Move llama.cpp clone/build from in-tree to ~/.unsloth/llama.cpp
- setup.sh: builds at ~/.unsloth/llama.cpp instead of ./llama.cpp
- setup.ps1: builds at %USERPROFILE%/.unsloth/llama.cpp
- inference llama_cpp.py: searches ~/.unsloth/ first, in-tree as legacy
- export.py: updated comments (unsloth-zoo handles path natively)
2026-03-02 04:04:41 +00:00
Roland Tannous
674cc67d78 Tighten Python bounds to >= 3.11, < 3.14 (matching setup.sh), only auto-install if missing 2026-03-01 13:05:10 +00:00
Roland Tannous
d5644d2d0d Add Python 3.12 prerequisite check with auto-install via winget 2026-03-01 13:05:10 +00:00
Roland Tannous
0267ba0a18 Auto-enable Windows Long Paths via UAC elevation during setup 2026-03-01 13:05:10 +00:00
Roland Tannous
6536bfb33b Remove unused CMP0194 cmake policy (eliminates cmake warning) 2026-03-01 13:05:10 +00:00
Roland Tannous
453f423d22 Simplify: use winget OpenSSL.Dev instead of vcpkg for HTTPS support 2026-03-01 13:05:10 +00:00
Roland Tannous
2e102b683e Add vcpkg/curl[ssl] for HTTPS support in llama-server, enable LLAMA_CURL=ON 2026-03-01 13:05:10 +00:00
Roland Tannous
6e5a3d1744 Download GGUF via huggingface_hub instead of llama-server -hf (fixes HTTPS not supported on Windows) 2026-03-01 13:05:10 +00:00
Roland Tannous
9eb0ff074b Add .venv/Scripts to User PATH so unsloth-studio works without activation 2026-03-01 13:05:10 +00:00
Roland Tannous
8e22b16bd8 Simplify completion banner: no venv activation needed 2026-03-01 13:05:10 +00:00
Roland Tannous
12867f701b Auto-add CUDA DLLs to PATH when launching llama-server on Windows 2026-03-01 13:05:10 +00:00
Roland Tannous
d79fe439ed Warn user to uninstall incompatible CUDA toolkit instead of failed side-by-side 2026-03-01 13:05:10 +00:00
Roland Tannous
b22c5b6ed8 Fallback: try descending CUDA versions if exact driver-max install fails 2026-03-01 13:05:10 +00:00
Roland Tannous
7b4d074857 Always persist compatible CUDA_PATH to User registry (overwrite stale values) 2026-03-01 13:05:10 +00:00
Roland Tannous
7576552717 Fix: scan side-by-side CUDA installs, pick compatible toolkit version 2026-03-01 13:05:10 +00:00
Roland Tannous
3521de7040 Build llama.cpp in-tree, auto-detect driver CUDA version for compatible toolkit 2026-03-01 13:05:10 +00:00
Roland Tannous
8d272ff8d5 Auto-detect driver CUDA version, install compatible toolkit instead of latest 2026-03-01 13:05:10 +00:00
Roland Tannous
bccbd26f3a Fix non-ASCII chars in test script for Windows PS 5.1 2026-03-01 13:05:10 +00:00
Roland Tannous
1684e48b1e Add llama-cpp Windows test script, fix binary lookup paths 2026-03-01 13:05:10 +00:00
Roland Tannous
f036a70681 Fix llama-server binary lookup for Windows (.exe, Release dir, ~/.unsloth) 2026-03-01 13:05:10 +00:00
Roland Tannous
7e021886c8 Force num_proc=1 on Windows to avoid slow spawn overhead 2026-03-01 13:05:10 +00:00
Roland Tannous
bd7c17708b Set short TORCHINDUCTOR_CACHE_DIR to fix Windows MAX_PATH crash 2026-03-01 13:05:10 +00:00
Roland Tannous
e1cc5e61b1 Fix npm Invalid Version: delete package-lock.json, relax Node constraint 2026-03-01 13:05:10 +00:00
Roland Tannous
aba3d8e29b Enforce Node LTS (v20-v22), add npm error checking, clean node_modules 2026-03-01 13:05:10 +00:00
Roland Tannous
2dfe0abaa1 Fix npm stderr crash on Windows ErrorActionPreference 2026-03-01 13:05:10 +00:00
Roland Tannous
662a1eb9d5 Fix Windows frontend build, add setup.bat, ANSI colors, aliases 2026-03-01 13:05:10 +00:00
Roland Tannous
783f0caf5f add setup.bat 2026-03-01 13:05:10 +00:00
Roland Tannous
28bac1859a Extract shared install_python_stack.py for cross-platform setup 2026-03-01 13:05:10 +00:00
Roland Tannous
4aba18375b Merge pull request #295 from unslothai/fix/fix-local-vision-gguf-loading
fix: support mmproj for local vision GGUF models + fix Windows pipe d…
2026-03-01 17:03:24 +04:00
Roland Tannous
ff93c97024 fix: support mmproj for local vision GGUF models + fix Windows pipe deadlock 2026-03-01 12:58:38 +00:00
Shine1i
ce33d673b9 feat(markdown): fix Mermaid integration with error handling and copy button 2026-02-27 20:02:40 +01:00
Roland Tannous
89d0a98192 Merge pull request #284 from unslothai/fix/drop-model-task-hard-filter
Apply HF task filtering only for empty model queries
2026-02-27 15:46:17 +04:00
imagineer99
49c319c2f7 fix: only apply HF task filter for empty model search queries 2026-02-27 11:31:53 +00:00
Wasim Yousef Said
ad7c1ceb40 Merge pull request #268 from unslothai/fix/delete-custom-config
Updated the delete custom preset in chat tab (filters)
2026-02-27 01:38:01 -08:00
Shine1i
6d74943dea keep checkpoint on preset apply 2026-02-27 10:36:07 +01:00
Wasim Yousef Said
7c364b594c Merge pull request #269 from unslothai/fix/fine-tuned-model-tooltip
Added tooltip for the fine tuned models in chat page
2026-02-27 01:34:49 -08:00
Wasim Yousef Said
5cd3d85882 Merge pull request #262 from unslothai/fix/truncated-text
Reduced name truncation on the training page
2026-02-27 01:33:05 -08:00
Shine1i
3c728f5eb3 merge nightly 2026-02-27 10:31:37 +01:00
Roland Tannous
6e12e2536d Merge pull request #267 from unslothai/fix/config-switch
Preserving model name during configuration type switch in chat page
2026-02-27 13:23:29 +04:00
Roland Tannous
5be7ae925d Merge pull request #266 from unslothai/fix/dataset-search-remove-size-download-badges
Remove dataset metadata badges from HF dataset dropdowns
2026-02-27 13:22:49 +04:00
Roland Tannous
cef36ee8e8 Merge pull request #282 from unslothai/fix/inference-auth
Added auth to inference endpoints
2026-02-27 13:18:38 +04:00
Roland Tannous
32d68b99e3 Merge pull request #278 from unslothai/fix/reorder-model-type-cards-onboarding
Reorder model type cards in onboarding to show Text first
2026-02-27 13:17:00 +04:00
Roland Tannous
90b06063d6 Merge pull request #275 from unslothai/fix/show-size-gguf
Passes metadata to get model size
2026-02-27 13:15:38 +04:00
Manan17
168957a87a Aggregating sharded models, showing fit/oom for quantizations 2026-02-27 08:23:15 +00:00
samit
b18a14d369 added auth to inference endpoints 2026-02-27 00:20:36 -08:00
Manan17
2fea4cadd3 Passes metadata to get model size 2026-02-27 07:38:55 +00:00
Roland Tannous
24c931c374 Merge pull request #276 from unslothai/fix/rebuild-llamacpp-setup
rebuild llama cpp for setup
2026-02-27 10:20:32 +04:00