Commit graph

147 commits

Author SHA1 Message Date
Leo Borcherding
a3daae1c40 fix: replace datetime.UTC with timezone.utc for Python 3.9+ compatibility
- Replace datetime.UTC with datetime.timezone.utc in authentication.py and storage.py
- Fixes ImportError on Python versions < 3.11
- timezone.utc works on Python 3.9+

Resolves #237
2026-02-24 14:37:00 -06:00
Roland Tannous
d38656139d Merge pull request #241 from unslothai/feature/adding-exported-models-for-chat
Adding exported model for chat
2026-02-24 13:55:45 +04:00
Roland Tannous
2149bc74ee Merge pull request #232 from unslothai/fix/disable-eval-by-default
# fix/disable eval by default
2026-02-24 13:35:11 +04:00
Roland Tannous
2be2933846 skip eval split and HF split detection when eval_steps is disabled 2026-02-24 09:26:54 +00:00
Manan17
aeb198f52d Fixing base model export issue for vlms 2026-02-24 01:34:11 +00:00
Manan17
4be677e45d Adding exported model for chat 2026-02-24 01:17:09 +00:00
Leo Borcherding
cdeed53a97 fix: disable eval by default, set eval_steps to 0.0
- Changed default eval_steps from 0.01 to 0.0 across backend and frontend
- Fixed UI to allow eval_steps=0 (removed min=0.001 constraint)
- Added conditional eval logic with helpful console messages
- Updated tooltip to explain how to disable evaluation
- Tested: confirmed eval disabled by default with eval_steps=0.0
2026-02-23 13:07:47 -06:00
Roland Tannous
d74174f7f5 Cap dataset.map num_proc on multi-GPU machines to prevent fork deadlocks 2026-02-23 14:25:31 +00:00
Roland Tannous
3015916d26 fix: error on >30% sample drop after train_on_responses_only instead of silent DataLoader crash 2026-02-23 12:21:06 +00:00
Roland Tannous
dbbcdb4f09 feat: clear unsloth_compiled_cache on startup, shutdown, and between model loads 2026-02-23 07:26:22 +00:00
Roland Tannous
fb1c321ad3 Add GLM, Qwen3 MoE, TinyQwen3 MoE, and Ministral 3 VL model defaults and GLM train_on_responses_only mapping 2026-02-23 05:51:43 +00:00
Roland Tannous
132cdb0547 fix: correct vision LoRA defaults for VLMs and remove vision fields from text-only model configs 2026-02-22 15:09:10 +00:00
Roland Tannous
202b7cdfa7 fix: pass full Processor as processing_class for VLM SFTTrainer 2026-02-22 14:11:12 +00:00
Roland Tannous
c051e3d532 fix: load proper vision processor from base model when FastVisionModel returns raw tokenizer, add tokenize=False to vision chat template 2026-02-21 04:40:29 +00:00
Manan17
f6ebeb1d42 Mapping proper tokenizer for VLMs 2026-02-21 01:57:05 +00:00
Manan17
3fa9e773c2 fixed the vlm's text only errors 2026-02-20 22:23:26 +00:00
Roland Tannous
ed476534f7 Merge pull request #186 from unslothai/fix/colab-setup-fixes
Fix/colab setup fixes
2026-02-20 23:08:25 +04:00
Roland Tannous
08ff8de31d add huggingface-hub==0.36.0 due to colab error 2026-02-20 18:26:20 +00:00
Roland Tannous
48e232b38c add huggingface-hub==0.36.0 due to colab error 2026-02-20 18:24:48 +00:00
Roland Tannous
34131da9a4 moved transformers4.57.1 to no-extra-deps 2026-02-20 18:08:58 +00:00
Manan17
798bfb8f6f Setting it to total cpu_count // 4 2026-02-20 06:32:01 +00:00
Manan17
fdeccec259 Fixing compare feature 2026-02-19 20:15:44 +00:00
Roland Tannous
18c41c2b08 Simplify dataset check to 2-tier, improve multimodal detection, auto-set trainOnCompletions, recheck dataset on reload 2026-02-19 11:25:54 +00:00
Roland Tannous
e2b7b4b54c change train_on_completions to true 2026-02-19 06:02:16 +00:00
Manan17
56869c63bd Passing use_auth = True and also having different checks which is missed by the is_vision function 2026-02-19 02:55:46 +00:00
Roland Tannous
adc0c78dbc reduce dataset_num_proc to 1/4 of cpu_count 2026-02-18 20:53:32 +00:00
Roland Tannous
6aceaec323 Merge pull request #155 from unslothai/fix/sft-tokenizer-unwrap-for-vlm-text
fix: Unwrap ProcessorMixin to raw tokenizer for text-only SFTTrainer on VLM-architecture models
2026-02-18 23:19:57 +04:00
Roland Tannous
e33920974b fix: unwrap ProcessorMixin to raw tokenizer for text-only SFTTrainer on VLM-architecture models 2026-02-18 19:13:20 +00:00
Roland Tannous
a6e2fa5b3a Merge remote-tracking branch 'origin/nightly' into fix/dataset-mapping-vlm-text-datasets 2026-02-18 18:11:45 +04:00
Roland Tannous
940328ce1e Merge remote-tracking branch 'origin/nightly' into feature/colab-notebook 2026-02-18 17:41:07 +04:00
Roland Tannous
d57b2742ab renamed UNSLOTH_FLEX_ATTENTION to UNSLOTH_ENABLE_FLEX_ATTENTION 2026-02-18 09:21:53 +00:00
Roland Tannous
5a02ed4f0f Disable flex attention on Blackwell+ GPUs (sm_120+) at startup 2026-02-18 08:58:25 +00:00
Roland Tannous
d69431fa57 Scale dataset num_proc dynamically to cpu_count//3 instead of hardcap 8 2026-02-18 08:38:53 +00:00
Manan17
76cd1dc24c fixing the hangup of training after multiple back to back training processes 2026-02-18 08:18:13 +00:00
Manan17
c37bf686a6 Dividing the total cpu_count // 3 2026-02-18 07:59:57 +00:00
Roland Tannous
14edb08cf5 Merge pull request #148 from unslothai/fix/linear
linear fix
2026-02-18 11:10:16 +04:00
Manan17
58116e7e7a fix the linear path on backend 2026-02-18 07:08:32 +00:00
Roland Tannous
d2f7eaf085 Merge pull request #145 from unslothai/fix/check-format-sample
fix: stream HF datasets in check-format endpoint to avoid full d…
2026-02-18 10:49:17 +04:00
Roland Tannous
d7853efd21 debug statements 2026-02-18 00:37:09 +00:00
Roland Tannous
0d8b67b706 fix: normalize target_modules [all-linear] list to string for Unsloth/PEFT compatibility 2026-02-18 00:32:21 +00:00
Roland Tannous
d2332622d1 fix: defensively rename VLM chat column to match model's forward() signature 2026-02-17 23:49:13 +00:00
Roland Tannous
3d0d1c7020 fix: cap dataset.map() num_proc to 8 to prevent CUDA fork deadlocks 2026-02-17 23:12:45 +00:00
Roland Tannous
5864dece26 fix: fix: stream HF datasets in check-format endpoint to avoid full downloads; add info logging to model config endpoints 2026-02-17 22:53:29 +00:00
Shine1i
dc0cec772d feat: enhance training stop and reset flow with detailed checks 2026-02-17 23:32:22 +01:00
Leo Borcherding
3e3c315b01 Merge nightly into feature/colab-notebook - resolved setup.sh conflicts 2026-02-17 15:51:14 -06:00
Wasim Yousef Said
765e1cfee2 Merge pull request #143 from unslothai/feature/local-models
feat: add schemas for local model discovery and listing
2026-02-17 13:10:04 -08:00
Shine1i
972cde7971 feat: add schemas for local model discovery and listing 2026-02-17 21:53:42 +01:00
Roland Tannous
c9fdce63e7 fix: skip sudo check on WSL during GGUF export to prevent password prompt hang 2026-02-17 19:30:02 +00:00
Roland Tannous
39b072d2ee move requirements/ to studio/backend/ and update paths in setup.sh 2026-02-17 19:02:25 +00:00
Roland Tannous
299bc65e36 chore: override model defaults to use max_steps=30, save_steps=30, num_epochs=0 for testing 2026-02-17 18:46:23 +00:00