Roland Tannous
0e7c8a2e5e
Switch GGUF backend from /v1/completions to /v1/chat/completions
...
Fixes two bugs:
1. Chat template tags (<|im_start|>, <|im_end|>) leaking into output
because /v1/completions treated them as literal text
2. Image hallucination because image_b64 was never passed to llama-server
Now llama-server handles chat templates natively and receives images
as OpenAI-format multimodal content parts for vision models.
2026-02-24 19:21:01 +04:00
Roland Tannous
ef1cd3ac98
Use llama-server -hf mode, add GGUF variant selector, fix vision detection
...
Replace Python-side GGUF download with llama-server's native -hf flag for
HuggingFace repos. Add frontend variant picker so users can choose
quantization (Q4_K_M, Q8_0, BF16, etc.) with file sizes. Fix vision
detection via mmproj files instead of hardcoding is_vision=False.
2026-02-24 19:03:06 +04:00
Roland Tannous
08aeeaee4b
Fix llama-server: build in-tree, fix path resolution, add LD_LIBRARY_PATH
2026-02-24 18:19:29 +04:00
Roland Tannous
4e88092452
Preflight llama-server check before downloading remote GGUF files
2026-02-24 18:02:43 +04:00
Roland Tannous
a900eb9ad7
Fix GGUF detection for HuggingFace repo IDs (not just local paths)
2026-02-24 17:49:09 +04:00
Roland Tannous
a40ebb1aab
Add GGUF model inference via llama-server backend
2026-02-24 17:40:05 +04:00
Roland Tannous
d74174f7f5
Cap dataset.map num_proc on multi-GPU machines to prevent fork deadlocks
2026-02-23 14:25:31 +00:00
Roland Tannous
3015916d26
fix: error on >30% sample drop after train_on_responses_only instead of silent DataLoader crash
2026-02-23 12:21:06 +00:00
Roland Tannous
dbbcdb4f09
feat: clear unsloth_compiled_cache on startup, shutdown, and between model loads
2026-02-23 07:26:22 +00:00
Roland Tannous
fb1c321ad3
Add GLM, Qwen3 MoE, TinyQwen3 MoE, and Ministral 3 VL model defaults and GLM train_on_responses_only mapping
2026-02-23 05:51:43 +00:00
Roland Tannous
132cdb0547
fix: correct vision LoRA defaults for VLMs and remove vision fields from text-only model configs
2026-02-22 15:09:10 +00:00
Roland Tannous
202b7cdfa7
fix: pass full Processor as processing_class for VLM SFTTrainer
2026-02-22 14:11:12 +00:00
Roland Tannous
761953b50e
feat(chat): persist model per thread and auto-load on thread switch
2026-02-22 13:35:45 +00:00
Roland Tannous
536a735acc
feat(chat): eject current model and start fresh thread on model switch
2026-02-22 13:17:03 +00:00
Roland Tannous
a490dca8a0
Merge pull request #211 from unslothai/fix/param-count-display
...
Fix: Remove download count fallback when model param count is unavailable
2026-02-22 16:52:25 +04:00
Roland Tannous
d3fbaa2256
Merge pull request #204 from unslothai/fix/image-preview-thumbnail
...
Fix: Resolve image preview thumbnail not rendering before send
2026-02-22 16:09:24 +04:00
Roland Tannous
36f10ea90e
Merge pull request #212 from unslothai/fix/stop-unclickable
...
Updated stop button to be unavailable during cancel training
2026-02-22 16:08:47 +04:00
Roland Tannous
df8e4bdc40
Merge pull request #203 from unslothai/feat/sort-unsloth-models-first
...
Feat: Sort unsloth models first in HF search dropdowns
2026-02-22 12:21:42 +04:00
Roland Tannous
75f3e5e2a1
feat: dual-query HF model search to surface all unsloth size variants first
2026-02-22 08:20:25 +00:00
imagineer99
4fe3772aa2
fix: remove download count fallback when model param count is unavailable
2026-02-22 06:30:50 +00:00
Shine1i
c726cff4c8
feat: add utils for deep cloning content and attachments in thread messages
2026-02-22 07:08:18 +01:00
imagineer99
52a2bfe016
fix: resolve image preview thumbnail not rendering before send
2026-02-22 02:39:56 +00:00
imagineer99
c815fc045d
feat: sort unsloth models first in HF search dropdowns
2026-02-22 01:45:51 +00:00
Roland Tannous
9bc32789a6
renaming downloading model to loading model
2026-02-21 13:30:41 +00:00
Roland Tannous
44828f582f
Merge pull request #200 from unslothai/fix/sloth-z-index-overlay
...
fix: enable navbar z-index by adding relative position
2026-02-21 17:15:05 +04:00
Roland Tannous
ea324f2806
Merge pull request #193 from unslothai/feature/vram-fit-chat
...
Added vram fit indicator to models in chat
2026-02-21 17:09:31 +04:00
Roland Tannous
3a4a576128
Merge pull request #196 from unslothai/fix/compare-dictate-attachment-buttons
...
Added dictate and add attachments feature in chat page
2026-02-21 16:52:56 +04:00
Roland Tannous
9b3cab6d5f
move microphone icon in compare page to be next to send button
2026-02-21 12:51:38 +00:00
samit
3fb9b4c056
added SpeechRecognition declarations and missing type packages
2026-02-21 01:00:08 -08:00
imagineer99
33b31c6be6
fix: enable navbar z-index by adding relative position
2026-02-21 08:37:47 +00:00
samit
f4a888cddb
Added VRAM fit indicator to recoomended models
2026-02-20 23:55:10 -08:00
Roland Tannous
ea0964de75
Merge pull request #192 from unslothai/feature/model-download-status
...
Updated to edit loading as "downloading model"
2026-02-21 11:07:08 +04:00
samit
e7d32a6461
updated stop button to unavailable during cancel training
2026-02-20 22:45:41 -08:00
samit
97f40bdc58
Added dictate and add attachments feature
2026-02-20 22:14:27 -08:00
Roland Tannous
c051e3d532
fix: load proper vision processor from base model when FastVisionModel returns raw tokenizer, add tokenize=False to vision chat template
2026-02-21 04:40:29 +00:00
Manan17
f6ebeb1d42
Mapping proper tokenizer for VLMs
2026-02-21 01:57:05 +00:00
samit
08c3c80d31
added vram fit indicator to models in chat
2026-02-20 17:25:03 -08:00
samit
0a3beade35
updated to edit loading as downloading model
2026-02-20 15:49:33 -08:00
Manan17
3fa9e773c2
fixed the vlm's text only errors
2026-02-20 22:23:26 +00:00
Roland Tannous
ed476534f7
Merge pull request #186 from unslothai/fix/colab-setup-fixes
...
Fix/colab setup fixes
2026-02-20 23:08:25 +04:00
Roland Tannous
08ff8de31d
add huggingface-hub==0.36.0 due to colab error
2026-02-20 18:26:20 +00:00
Roland Tannous
48e232b38c
add huggingface-hub==0.36.0 due to colab error
2026-02-20 18:24:48 +00:00
Roland Tannous
34131da9a4
moved transformers4.57.1 to no-extra-deps
2026-02-20 18:08:58 +00:00
Roland Tannous
e5f9ae5c9f
Merge branch 'nightly' into fix/dropdown-menu-prefill
2026-02-20 17:44:06 +00:00
imagineer99
77b7e8a9ba
fix: remove warmup text inference status
2026-02-20 15:35:06 +00:00
imagineer99
759ae059db
Fix: model and dataset dropdowns selecting stale value on Enter
2026-02-20 11:33:20 +00:00
Roland Tannous
811a9243b2
Merge pull request #179 from unslothai/fix/copy-mac
...
Added the copy feature on mac
2026-02-20 14:38:29 +04:00
Roland Tannous
cdd1f7fce2
Merge pull request #166 from unslothai/feat/download-progress-indicator
...
feat: add download progress indicators for dataset preview and training overlay
2026-02-20 14:11:07 +04:00
Roland Tannous
168055a267
Merge pull request #172 from unslothai/fix/training-param
...
Added optim and lr_scheduler_type in the frontend
2026-02-20 14:05:33 +04:00
samit
3d403c6c99
edited the font of the new parameters
2026-02-20 01:29:37 -08:00