Roland Tannous
d9434fee4a
fix: use raw github URL for vision.py patch + add VLM processor diagnostic logging
2026-02-25 10:29:05 +00:00
Roland Tannous
9e280eb105
Fix GGUF export cwd confusion: remove os.chdir, use absolute paths
...
Remove os.chdir(save_directory) from export.py which was causing all of
unsloth-zoo's relative-path internals (check_llama_cpp, use_local_gguf,
_download_convert_hf_to_gguf) to resolve against the export directory
instead of the repo root. This caused llama.cpp to be cloned inside each
export dir and destroyed the repo root's llama-server build on cleanup.
Now passes absolute paths to save_pretrained_gguf so unsloth resolves
llama.cpp from the repo root where setup.sh already built it.
Also builds llama-quantize in setup.sh (needed by unsloth-zoo's export
pipeline) and symlinks it to llama.cpp root for check_llama_cpp().
2026-02-25 03:30:54 +04:00
Roland Tannous
2ebeba8588
Switch GGUF backend from /v1/completions to /v1/chat/completions
...
Fixes two bugs:
1. Chat template tags (<|im_start|>, <|im_end|>) leaking into output
because /v1/completions treated them as literal text
2. Image hallucination because image_b64 was never passed to llama-server
Now llama-server handles chat templates natively and receives images
as OpenAI-format multimodal content parts for vision models.
2026-02-24 19:21:01 +04:00
Roland Tannous
3ee4f1359a
Use llama-server -hf mode, add GGUF variant selector, fix vision detection
...
Replace Python-side GGUF download with llama-server's native -hf flag for
HuggingFace repos. Add frontend variant picker so users can choose
quantization (Q4_K_M, Q8_0, BF16, etc.) with file sizes. Fix vision
detection via mmproj files instead of hardcoding is_vision=False.
2026-02-24 19:03:06 +04:00
Roland Tannous
5b7555cd3f
Fix llama-server: build in-tree, fix path resolution, add LD_LIBRARY_PATH
2026-02-24 18:19:29 +04:00
Roland Tannous
4a82e704aa
Preflight llama-server check before downloading remote GGUF files
2026-02-24 18:02:43 +04:00
Roland Tannous
c635d4f49c
Fix GGUF detection for HuggingFace repo IDs (not just local paths)
2026-02-24 17:49:09 +04:00
Roland Tannous
2f985ccbb5
Add GGUF model inference via llama-server backend
2026-02-24 17:40:05 +04:00
Roland Tannous
834013aae5
Cap dataset.map num_proc on multi-GPU machines to prevent fork deadlocks
2026-02-23 14:25:31 +00:00
Roland Tannous
d94f842158
fix: error on >30% sample drop after train_on_responses_only instead of silent DataLoader crash
2026-02-23 12:21:06 +00:00
Roland Tannous
198433363a
feat: clear unsloth_compiled_cache on startup, shutdown, and between model loads
2026-02-23 07:26:22 +00:00
Roland Tannous
62c260a109
Add GLM, Qwen3 MoE, TinyQwen3 MoE, and Ministral 3 VL model defaults and GLM train_on_responses_only mapping
2026-02-23 05:51:43 +00:00
Roland Tannous
e866159e3b
fix: correct vision LoRA defaults for VLMs and remove vision fields from text-only model configs
2026-02-22 15:09:10 +00:00
Roland Tannous
e666442b6e
fix: pass full Processor as processing_class for VLM SFTTrainer
2026-02-22 14:11:12 +00:00
Roland Tannous
95c5c679bc
feat(chat): persist model per thread and auto-load on thread switch
2026-02-22 13:35:45 +00:00
Roland Tannous
ef8cbdde93
feat(chat): eject current model and start fresh thread on model switch
2026-02-22 13:17:03 +00:00
Roland Tannous
1fa64ed2b9
Merge pull request #211 from unslothai/fix/param-count-display
...
Fix: Remove download count fallback when model param count is unavailable
2026-02-22 16:52:25 +04:00
Roland Tannous
cccf7d2558
Merge pull request #204 from unslothai/fix/image-preview-thumbnail
...
Fix: Resolve image preview thumbnail not rendering before send
2026-02-22 16:09:24 +04:00
Roland Tannous
0caab5846a
Merge pull request #212 from unslothai/fix/stop-unclickable
...
Updated stop button to be unavailable during cancel training
2026-02-22 16:08:47 +04:00
Roland Tannous
94b4d8b4b3
Merge pull request #203 from unslothai/feat/sort-unsloth-models-first
...
Feat: Sort unsloth models first in HF search dropdowns
2026-02-22 12:21:42 +04:00
Roland Tannous
03f4337e48
feat: dual-query HF model search to surface all unsloth size variants first
2026-02-22 08:20:25 +00:00
imagineer99
27db6b287d
fix: remove download count fallback when model param count is unavailable
2026-02-22 06:30:50 +00:00
Shine1i
5535e5783d
feat: add utils for deep cloning content and attachments in thread messages
2026-02-22 07:08:18 +01:00
imagineer99
ac399b152d
fix: resolve image preview thumbnail not rendering before send
2026-02-22 02:39:56 +00:00
imagineer99
5bdec309e6
feat: sort unsloth models first in HF search dropdowns
2026-02-22 01:45:51 +00:00
Roland Tannous
73fe288e35
renaming downloading model to loading model
2026-02-21 13:30:41 +00:00
Roland Tannous
7d1d816177
Merge pull request #200 from unslothai/fix/sloth-z-index-overlay
...
fix: enable navbar z-index by adding relative position
2026-02-21 17:15:05 +04:00
Roland Tannous
8469605b72
Merge pull request #193 from unslothai/feature/vram-fit-chat
...
Added vram fit indicator to models in chat
2026-02-21 17:09:31 +04:00
Roland Tannous
d170398697
Merge pull request #196 from unslothai/fix/compare-dictate-attachment-buttons
...
Added dictate and add attachments feature in chat page
2026-02-21 16:52:56 +04:00
Roland Tannous
df518aa7cc
move microphone icon in compare page to be next to send button
2026-02-21 12:51:38 +00:00
samit
e25966c6dc
added SpeechRecognition declarations and missing type packages
2026-02-21 01:00:08 -08:00
imagineer99
9a26f04703
fix: enable navbar z-index by adding relative position
2026-02-21 08:37:47 +00:00
samit
54c42bf9f5
Added VRAM fit indicator to recoomended models
2026-02-20 23:55:10 -08:00
Roland Tannous
92c1eb1b35
Merge pull request #192 from unslothai/feature/model-download-status
...
Updated to edit loading as "downloading model"
2026-02-21 11:07:08 +04:00
samit
662f1bfbc4
updated stop button to unavailable during cancel training
2026-02-20 22:45:41 -08:00
samit
80976f8c70
Added dictate and add attachments feature
2026-02-20 22:14:27 -08:00
Roland Tannous
ef118d0d05
fix: load proper vision processor from base model when FastVisionModel returns raw tokenizer, add tokenize=False to vision chat template
2026-02-21 04:40:29 +00:00
Manan17
e9710874e1
Mapping proper tokenizer for VLMs
2026-02-21 01:57:05 +00:00
samit
262ffa59af
added vram fit indicator to models in chat
2026-02-20 17:25:03 -08:00
samit
862e171ff5
updated to edit loading as downloading model
2026-02-20 15:49:33 -08:00
Manan17
756aa56cd2
fixed the vlm's text only errors
2026-02-20 22:23:26 +00:00
Roland Tannous
0d0038b901
Merge pull request #186 from unslothai/fix/colab-setup-fixes
...
Fix/colab setup fixes
2026-02-20 23:08:25 +04:00
Roland Tannous
e1a24c1ee9
add huggingface-hub==0.36.0 due to colab error
2026-02-20 18:26:20 +00:00
Roland Tannous
7c67228c40
add huggingface-hub==0.36.0 due to colab error
2026-02-20 18:24:48 +00:00
Roland Tannous
3bd0b8fb8b
moved transformers4.57.1 to no-extra-deps
2026-02-20 18:08:58 +00:00
Roland Tannous
fffb2ea22e
Merge branch 'nightly' into fix/dropdown-menu-prefill
2026-02-20 17:44:06 +00:00
imagineer99
2697b9ed74
fix: remove warmup text inference status
2026-02-20 15:35:06 +00:00
imagineer99
48ddbd96f9
Fix: model and dataset dropdowns selecting stale value on Enter
2026-02-20 11:33:20 +00:00
Roland Tannous
046419cee5
Merge pull request #179 from unslothai/fix/copy-mac
...
Added the copy feature on mac
2026-02-20 14:38:29 +04:00
Roland Tannous
66b64f0636
Merge pull request #166 from unslothai/feat/download-progress-indicator
...
feat: add download progress indicators for dataset preview and training overlay
2026-02-20 14:11:07 +04:00