Commit graph

134 commits

Author SHA1 Message Date
Roland Tannous
ef118d0d05 fix: load proper vision processor from base model when FastVisionModel returns raw tokenizer, add tokenize=False to vision chat template 2026-02-21 04:40:29 +00:00
Manan17
e9710874e1 Mapping proper tokenizer for VLMs 2026-02-21 01:57:05 +00:00
Manan17
756aa56cd2 fixed the vlm's text only errors 2026-02-20 22:23:26 +00:00
Roland Tannous
0d0038b901 Merge pull request #186 from unslothai/fix/colab-setup-fixes
Fix/colab setup fixes
2026-02-20 23:08:25 +04:00
Roland Tannous
e1a24c1ee9 add huggingface-hub==0.36.0 due to colab error 2026-02-20 18:26:20 +00:00
Roland Tannous
7c67228c40 add huggingface-hub==0.36.0 due to colab error 2026-02-20 18:24:48 +00:00
Roland Tannous
3bd0b8fb8b moved transformers4.57.1 to no-extra-deps 2026-02-20 18:08:58 +00:00
Manan17
bd0cee8c15 Setting it to total cpu_count // 4 2026-02-20 06:32:01 +00:00
Manan17
444ece6b07 Fixing compare feature 2026-02-19 20:15:44 +00:00
Roland Tannous
11b3029dc6 Simplify dataset check to 2-tier, improve multimodal detection, auto-set trainOnCompletions, recheck dataset on reload 2026-02-19 11:25:54 +00:00
Roland Tannous
58753151ef change train_on_completions to true 2026-02-19 06:02:16 +00:00
Manan17
2755cf922d Passing use_auth = True and also having different checks which is missed by the is_vision function 2026-02-19 02:55:46 +00:00
Roland Tannous
c876b38780 reduce dataset_num_proc to 1/4 of cpu_count 2026-02-18 20:53:32 +00:00
Roland Tannous
648b29ac9b Merge pull request #155 from unslothai/fix/sft-tokenizer-unwrap-for-vlm-text
fix: Unwrap ProcessorMixin to raw tokenizer for text-only SFTTrainer on VLM-architecture models
2026-02-18 23:19:57 +04:00
Roland Tannous
23214c41c0 fix: unwrap ProcessorMixin to raw tokenizer for text-only SFTTrainer on VLM-architecture models 2026-02-18 19:13:20 +00:00
Roland Tannous
9840864662 Merge remote-tracking branch 'origin/nightly' into fix/dataset-mapping-vlm-text-datasets 2026-02-18 18:11:45 +04:00
Roland Tannous
ad638118b9 Merge remote-tracking branch 'origin/nightly' into feature/colab-notebook 2026-02-18 17:41:07 +04:00
Roland Tannous
ae040cf681 renamed UNSLOTH_FLEX_ATTENTION to UNSLOTH_ENABLE_FLEX_ATTENTION 2026-02-18 09:21:53 +00:00
Roland Tannous
e3f4a9eb32 Disable flex attention on Blackwell+ GPUs (sm_120+) at startup 2026-02-18 08:58:25 +00:00
Roland Tannous
29b25169c0 Scale dataset num_proc dynamically to cpu_count//3 instead of hardcap 8 2026-02-18 08:38:53 +00:00
Manan17
949e57c334 fixing the hangup of training after multiple back to back training processes 2026-02-18 08:18:13 +00:00
Manan17
db0fa1a270 Dividing the total cpu_count // 3 2026-02-18 07:59:57 +00:00
Roland Tannous
c616697b22 Merge pull request #148 from unslothai/fix/linear
linear fix
2026-02-18 11:10:16 +04:00
Manan17
c832c903b4 fix the linear path on backend 2026-02-18 07:08:32 +00:00
Roland Tannous
38e577ef26 Merge pull request #145 from unslothai/fix/check-format-sample
fix: stream HF datasets in check-format endpoint to avoid full d…
2026-02-18 10:49:17 +04:00
Roland Tannous
9d737559a5 debug statements 2026-02-18 00:37:09 +00:00
Roland Tannous
d94a1e0289 fix: normalize target_modules [all-linear] list to string for Unsloth/PEFT compatibility 2026-02-18 00:32:21 +00:00
Roland Tannous
f0613f5d07 fix: defensively rename VLM chat column to match model's forward() signature 2026-02-17 23:49:13 +00:00
Roland Tannous
c0f210bc2a fix: cap dataset.map() num_proc to 8 to prevent CUDA fork deadlocks 2026-02-17 23:12:45 +00:00
Roland Tannous
4d0c6d20b3 fix: fix: stream HF datasets in check-format endpoint to avoid full downloads; add info logging to model config endpoints 2026-02-17 22:53:29 +00:00
Shine1i
b31461790f feat: enhance training stop and reset flow with detailed checks 2026-02-17 23:32:22 +01:00
Leo Borcherding
84b9a8aef6 Merge nightly into feature/colab-notebook - resolved setup.sh conflicts 2026-02-17 15:51:14 -06:00
Wasim Yousef Said
8fbc80f66a Merge pull request #143 from unslothai/feature/local-models
feat: add schemas for local model discovery and listing
2026-02-17 13:10:04 -08:00
Shine1i
a2cf89214e feat: add schemas for local model discovery and listing 2026-02-17 21:53:42 +01:00
Roland Tannous
c5558312c8 fix: skip sudo check on WSL during GGUF export to prevent password prompt hang 2026-02-17 19:30:02 +00:00
Roland Tannous
bcd9416ffb move requirements/ to studio/backend/ and update paths in setup.sh 2026-02-17 19:02:25 +00:00
Roland Tannous
028408e432 chore: override model defaults to use max_steps=30, save_steps=30, num_epochs=0 for testing 2026-02-17 18:46:23 +00:00
Roland Tannous
2cc3f03fbe Merge pull request #136 from unslothai/setup/update-setup-sh-dependencies
setup.sh: Replace inline pip installs with pinned requirements files
2026-02-17 22:04:03 +04:00
Shine1i
f47c424be3 feat: integrate gradient norm tracking in training runtime and metrics
- Enhanced chart logic to filter and visualize finite gradient norm values.
2026-02-17 18:26:59 +01:00
Roland Tannous
5dd93579c2 add full dependency chain for unsloth + unsloth-extras 2026-02-17 14:21:18 +00:00
Leo Borcherding
7fb720b4fc feat: Add simple 2-cell Colab notebook (no tunnel needed)
- Create studio/backend/colab.py using Colab's built-in proxy
- Uses google.colab.kernel.proxyPort() for URL (no cloudflare)
- Shows nice clickable link with IPython.display.HTML
- Notebook has just 2 cells: setup and start
- Much simpler than external tunneling approach
2026-02-17 04:57:30 -06:00
Manan17
8f1db03c15 Adding metadata for checkpoints 2026-02-16 23:46:17 +00:00
Wasim Yousef Said
16f79a73de Merge pull request #124 from unslothai/feature/bug-fixes
feat: support disabling top-k sampling with -1 and standardize normalization
2026-02-16 13:21:12 -08:00
Roland Tannous
108ec254cb Merge branch 'nightly' into feature/eval-split-auto-detection 2026-02-17 01:11:30 +04:00
Shine1i
4be6eefed3 feat: support disabling top-k sampling with -1 and standardize normalization logic
- Updated top-k parameter range to accept -1 in models and frontend.
- Added utility to normalize top-k for backend compatibility.
2026-02-16 21:33:24 +01:00
Roland Tannous
18879a521b feat: auto-detect model+dataset compatibility to select VLM vs LLM training path 2026-02-16 19:18:49 +00:00
Roland Tannous
ed6d4b2fb6 feat: apply default chat template for base models without tokenizer chat_template 2026-02-16 15:56:06 +00:00
Roland Tannous
3b117189c5 feat: add eval_enabled flag and format-first-then-split for eval dataset 2026-02-16 14:13:55 +00:00
Roland Tannous
5962bec41a feat: add eval split auto-detection, eval_steps hyperparam, and eval_loss chart integration 2026-02-16 13:51:10 +00:00
Roland Tannous
90c3561adb feat: add eval split auto-detection, eval_steps hyperparam, and eval_loss chart integration 2026-02-16 13:38:54 +00:00