Commit graph

301 commits

Author SHA1 Message Date
Roland Tannous
d9434fee4a fix: use raw github URL for vision.py patch + add VLM processor diagnostic logging 2026-02-25 10:29:05 +00:00
Roland Tannous
9e280eb105 Fix GGUF export cwd confusion: remove os.chdir, use absolute paths
Remove os.chdir(save_directory) from export.py which was causing all of
unsloth-zoo's relative-path internals (check_llama_cpp, use_local_gguf,
_download_convert_hf_to_gguf) to resolve against the export directory
instead of the repo root. This caused llama.cpp to be cloned inside each
export dir and destroyed the repo root's llama-server build on cleanup.

Now passes absolute paths to save_pretrained_gguf so unsloth resolves
llama.cpp from the repo root where setup.sh already built it.

Also builds llama-quantize in setup.sh (needed by unsloth-zoo's export
pipeline) and symlinks it to llama.cpp root for check_llama_cpp().
2026-02-25 03:30:54 +04:00
Roland Tannous
2ebeba8588 Switch GGUF backend from /v1/completions to /v1/chat/completions
Fixes two bugs:
1. Chat template tags (<|im_start|>, <|im_end|>) leaking into output
   because /v1/completions treated them as literal text
2. Image hallucination because image_b64 was never passed to llama-server

Now llama-server handles chat templates natively and receives images
as OpenAI-format multimodal content parts for vision models.
2026-02-24 19:21:01 +04:00
Roland Tannous
3ee4f1359a Use llama-server -hf mode, add GGUF variant selector, fix vision detection
Replace Python-side GGUF download with llama-server's native -hf flag for
HuggingFace repos. Add frontend variant picker so users can choose
quantization (Q4_K_M, Q8_0, BF16, etc.) with file sizes. Fix vision
detection via mmproj files instead of hardcoding is_vision=False.
2026-02-24 19:03:06 +04:00
Roland Tannous
5b7555cd3f Fix llama-server: build in-tree, fix path resolution, add LD_LIBRARY_PATH 2026-02-24 18:19:29 +04:00
Roland Tannous
4a82e704aa Preflight llama-server check before downloading remote GGUF files 2026-02-24 18:02:43 +04:00
Roland Tannous
c635d4f49c Fix GGUF detection for HuggingFace repo IDs (not just local paths) 2026-02-24 17:49:09 +04:00
Roland Tannous
2f985ccbb5 Add GGUF model inference via llama-server backend 2026-02-24 17:40:05 +04:00
Roland Tannous
834013aae5 Cap dataset.map num_proc on multi-GPU machines to prevent fork deadlocks 2026-02-23 14:25:31 +00:00
Roland Tannous
d94f842158 fix: error on >30% sample drop after train_on_responses_only instead of silent DataLoader crash 2026-02-23 12:21:06 +00:00
Roland Tannous
198433363a feat: clear unsloth_compiled_cache on startup, shutdown, and between model loads 2026-02-23 07:26:22 +00:00
Roland Tannous
62c260a109 Add GLM, Qwen3 MoE, TinyQwen3 MoE, and Ministral 3 VL model defaults and GLM train_on_responses_only mapping 2026-02-23 05:51:43 +00:00
Roland Tannous
e866159e3b fix: correct vision LoRA defaults for VLMs and remove vision fields from text-only model configs 2026-02-22 15:09:10 +00:00
Roland Tannous
e666442b6e fix: pass full Processor as processing_class for VLM SFTTrainer 2026-02-22 14:11:12 +00:00
Roland Tannous
95c5c679bc feat(chat): persist model per thread and auto-load on thread switch 2026-02-22 13:35:45 +00:00
Roland Tannous
ef8cbdde93 feat(chat): eject current model and start fresh thread on model switch 2026-02-22 13:17:03 +00:00
Roland Tannous
1fa64ed2b9 Merge pull request #211 from unslothai/fix/param-count-display
Fix: Remove download count fallback when model param count is unavailable
2026-02-22 16:52:25 +04:00
Roland Tannous
cccf7d2558 Merge pull request #204 from unslothai/fix/image-preview-thumbnail
Fix: Resolve image preview thumbnail not rendering before send
2026-02-22 16:09:24 +04:00
Roland Tannous
0caab5846a Merge pull request #212 from unslothai/fix/stop-unclickable
Updated stop button to be unavailable during cancel training
2026-02-22 16:08:47 +04:00
Roland Tannous
94b4d8b4b3 Merge pull request #203 from unslothai/feat/sort-unsloth-models-first
Feat: Sort unsloth models first in HF search dropdowns
2026-02-22 12:21:42 +04:00
Roland Tannous
03f4337e48 feat: dual-query HF model search to surface all unsloth size variants first 2026-02-22 08:20:25 +00:00
imagineer99
27db6b287d fix: remove download count fallback when model param count is unavailable 2026-02-22 06:30:50 +00:00
Shine1i
5535e5783d feat: add utils for deep cloning content and attachments in thread messages 2026-02-22 07:08:18 +01:00
imagineer99
ac399b152d fix: resolve image preview thumbnail not rendering before send 2026-02-22 02:39:56 +00:00
imagineer99
5bdec309e6 feat: sort unsloth models first in HF search dropdowns 2026-02-22 01:45:51 +00:00
Roland Tannous
73fe288e35 renaming downloading model to loading model 2026-02-21 13:30:41 +00:00
Roland Tannous
7d1d816177 Merge pull request #200 from unslothai/fix/sloth-z-index-overlay
fix: enable navbar z-index by adding relative position
2026-02-21 17:15:05 +04:00
Roland Tannous
8469605b72 Merge pull request #193 from unslothai/feature/vram-fit-chat
Added vram fit indicator to models in chat
2026-02-21 17:09:31 +04:00
Roland Tannous
d170398697 Merge pull request #196 from unslothai/fix/compare-dictate-attachment-buttons
Added dictate and add attachments feature in chat page
2026-02-21 16:52:56 +04:00
Roland Tannous
df518aa7cc move microphone icon in compare page to be next to send button 2026-02-21 12:51:38 +00:00
samit
e25966c6dc added SpeechRecognition declarations and missing type packages 2026-02-21 01:00:08 -08:00
imagineer99
9a26f04703 fix: enable navbar z-index by adding relative position 2026-02-21 08:37:47 +00:00
samit
54c42bf9f5 Added VRAM fit indicator to recoomended models 2026-02-20 23:55:10 -08:00
Roland Tannous
92c1eb1b35 Merge pull request #192 from unslothai/feature/model-download-status
Updated to edit loading as "downloading model"
2026-02-21 11:07:08 +04:00
samit
662f1bfbc4 updated stop button to unavailable during cancel training 2026-02-20 22:45:41 -08:00
samit
80976f8c70 Added dictate and add attachments feature 2026-02-20 22:14:27 -08:00
Roland Tannous
ef118d0d05 fix: load proper vision processor from base model when FastVisionModel returns raw tokenizer, add tokenize=False to vision chat template 2026-02-21 04:40:29 +00:00
Manan17
e9710874e1 Mapping proper tokenizer for VLMs 2026-02-21 01:57:05 +00:00
samit
262ffa59af added vram fit indicator to models in chat 2026-02-20 17:25:03 -08:00
samit
862e171ff5 updated to edit loading as downloading model 2026-02-20 15:49:33 -08:00
Manan17
756aa56cd2 fixed the vlm's text only errors 2026-02-20 22:23:26 +00:00
Roland Tannous
0d0038b901 Merge pull request #186 from unslothai/fix/colab-setup-fixes
Fix/colab setup fixes
2026-02-20 23:08:25 +04:00
Roland Tannous
e1a24c1ee9 add huggingface-hub==0.36.0 due to colab error 2026-02-20 18:26:20 +00:00
Roland Tannous
7c67228c40 add huggingface-hub==0.36.0 due to colab error 2026-02-20 18:24:48 +00:00
Roland Tannous
3bd0b8fb8b moved transformers4.57.1 to no-extra-deps 2026-02-20 18:08:58 +00:00
Roland Tannous
fffb2ea22e Merge branch 'nightly' into fix/dropdown-menu-prefill 2026-02-20 17:44:06 +00:00
imagineer99
2697b9ed74 fix: remove warmup text inference status 2026-02-20 15:35:06 +00:00
imagineer99
48ddbd96f9 Fix: model and dataset dropdowns selecting stale value on Enter 2026-02-20 11:33:20 +00:00
Roland Tannous
046419cee5 Merge pull request #179 from unslothai/fix/copy-mac
Added the copy feature on mac
2026-02-20 14:38:29 +04:00
Roland Tannous
66b64f0636 Merge pull request #166 from unslothai/feat/download-progress-indicator
feat: add download progress indicators for dataset preview and training overlay
2026-02-20 14:11:07 +04:00