Roland Tannous
|
08aeeaee4b
|
Fix llama-server: build in-tree, fix path resolution, add LD_LIBRARY_PATH
|
2026-02-24 18:19:29 +04:00 |
|
Roland Tannous
|
4e88092452
|
Preflight llama-server check before downloading remote GGUF files
|
2026-02-24 18:02:43 +04:00 |
|
Roland Tannous
|
a900eb9ad7
|
Fix GGUF detection for HuggingFace repo IDs (not just local paths)
|
2026-02-24 17:49:09 +04:00 |
|
Roland Tannous
|
a40ebb1aab
|
Add GGUF model inference via llama-server backend
|
2026-02-24 17:40:05 +04:00 |
|
Roland Tannous
|
d74174f7f5
|
Cap dataset.map num_proc on multi-GPU machines to prevent fork deadlocks
|
2026-02-23 14:25:31 +00:00 |
|
Roland Tannous
|
3015916d26
|
fix: error on >30% sample drop after train_on_responses_only instead of silent DataLoader crash
|
2026-02-23 12:21:06 +00:00 |
|
Roland Tannous
|
dbbcdb4f09
|
feat: clear unsloth_compiled_cache on startup, shutdown, and between model loads
|
2026-02-23 07:26:22 +00:00 |
|
Roland Tannous
|
fb1c321ad3
|
Add GLM, Qwen3 MoE, TinyQwen3 MoE, and Ministral 3 VL model defaults and GLM train_on_responses_only mapping
|
2026-02-23 05:51:43 +00:00 |
|
Roland Tannous
|
132cdb0547
|
fix: correct vision LoRA defaults for VLMs and remove vision fields from text-only model configs
|
2026-02-22 15:09:10 +00:00 |
|
Roland Tannous
|
202b7cdfa7
|
fix: pass full Processor as processing_class for VLM SFTTrainer
|
2026-02-22 14:11:12 +00:00 |
|
Roland Tannous
|
c051e3d532
|
fix: load proper vision processor from base model when FastVisionModel returns raw tokenizer, add tokenize=False to vision chat template
|
2026-02-21 04:40:29 +00:00 |
|
Manan17
|
f6ebeb1d42
|
Mapping proper tokenizer for VLMs
|
2026-02-21 01:57:05 +00:00 |
|
Manan17
|
3fa9e773c2
|
fixed the vlm's text only errors
|
2026-02-20 22:23:26 +00:00 |
|
Roland Tannous
|
ed476534f7
|
Merge pull request #186 from unslothai/fix/colab-setup-fixes
Fix/colab setup fixes
|
2026-02-20 23:08:25 +04:00 |
|
Roland Tannous
|
08ff8de31d
|
add huggingface-hub==0.36.0 due to colab error
|
2026-02-20 18:26:20 +00:00 |
|
Roland Tannous
|
48e232b38c
|
add huggingface-hub==0.36.0 due to colab error
|
2026-02-20 18:24:48 +00:00 |
|
Roland Tannous
|
34131da9a4
|
moved transformers4.57.1 to no-extra-deps
|
2026-02-20 18:08:58 +00:00 |
|
Manan17
|
798bfb8f6f
|
Setting it to total cpu_count // 4
|
2026-02-20 06:32:01 +00:00 |
|
Manan17
|
fdeccec259
|
Fixing compare feature
|
2026-02-19 20:15:44 +00:00 |
|
Roland Tannous
|
18c41c2b08
|
Simplify dataset check to 2-tier, improve multimodal detection, auto-set trainOnCompletions, recheck dataset on reload
|
2026-02-19 11:25:54 +00:00 |
|
Roland Tannous
|
e2b7b4b54c
|
change train_on_completions to true
|
2026-02-19 06:02:16 +00:00 |
|
Manan17
|
56869c63bd
|
Passing use_auth = True and also having different checks which is missed by the is_vision function
|
2026-02-19 02:55:46 +00:00 |
|
Roland Tannous
|
adc0c78dbc
|
reduce dataset_num_proc to 1/4 of cpu_count
|
2026-02-18 20:53:32 +00:00 |
|
Roland Tannous
|
6aceaec323
|
Merge pull request #155 from unslothai/fix/sft-tokenizer-unwrap-for-vlm-text
fix: Unwrap ProcessorMixin to raw tokenizer for text-only SFTTrainer on VLM-architecture models
|
2026-02-18 23:19:57 +04:00 |
|
Roland Tannous
|
e33920974b
|
fix: unwrap ProcessorMixin to raw tokenizer for text-only SFTTrainer on VLM-architecture models
|
2026-02-18 19:13:20 +00:00 |
|
Roland Tannous
|
a6e2fa5b3a
|
Merge remote-tracking branch 'origin/nightly' into fix/dataset-mapping-vlm-text-datasets
|
2026-02-18 18:11:45 +04:00 |
|
Roland Tannous
|
940328ce1e
|
Merge remote-tracking branch 'origin/nightly' into feature/colab-notebook
|
2026-02-18 17:41:07 +04:00 |
|
Roland Tannous
|
d57b2742ab
|
renamed UNSLOTH_FLEX_ATTENTION to UNSLOTH_ENABLE_FLEX_ATTENTION
|
2026-02-18 09:21:53 +00:00 |
|
Roland Tannous
|
5a02ed4f0f
|
Disable flex attention on Blackwell+ GPUs (sm_120+) at startup
|
2026-02-18 08:58:25 +00:00 |
|
Roland Tannous
|
d69431fa57
|
Scale dataset num_proc dynamically to cpu_count//3 instead of hardcap 8
|
2026-02-18 08:38:53 +00:00 |
|
Manan17
|
76cd1dc24c
|
fixing the hangup of training after multiple back to back training processes
|
2026-02-18 08:18:13 +00:00 |
|
Manan17
|
c37bf686a6
|
Dividing the total cpu_count // 3
|
2026-02-18 07:59:57 +00:00 |
|
Roland Tannous
|
14edb08cf5
|
Merge pull request #148 from unslothai/fix/linear
linear fix
|
2026-02-18 11:10:16 +04:00 |
|
Manan17
|
58116e7e7a
|
fix the linear path on backend
|
2026-02-18 07:08:32 +00:00 |
|
Roland Tannous
|
d2f7eaf085
|
Merge pull request #145 from unslothai/fix/check-format-sample
fix: stream HF datasets in check-format endpoint to avoid full d…
|
2026-02-18 10:49:17 +04:00 |
|
Roland Tannous
|
d7853efd21
|
debug statements
|
2026-02-18 00:37:09 +00:00 |
|
Roland Tannous
|
0d8b67b706
|
fix: normalize target_modules [all-linear] list to string for Unsloth/PEFT compatibility
|
2026-02-18 00:32:21 +00:00 |
|
Roland Tannous
|
d2332622d1
|
fix: defensively rename VLM chat column to match model's forward() signature
|
2026-02-17 23:49:13 +00:00 |
|
Roland Tannous
|
3d0d1c7020
|
fix: cap dataset.map() num_proc to 8 to prevent CUDA fork deadlocks
|
2026-02-17 23:12:45 +00:00 |
|
Roland Tannous
|
5864dece26
|
fix: fix: stream HF datasets in check-format endpoint to avoid full downloads; add info logging to model config endpoints
|
2026-02-17 22:53:29 +00:00 |
|
Shine1i
|
dc0cec772d
|
feat: enhance training stop and reset flow with detailed checks
|
2026-02-17 23:32:22 +01:00 |
|
Leo Borcherding
|
3e3c315b01
|
Merge nightly into feature/colab-notebook - resolved setup.sh conflicts
|
2026-02-17 15:51:14 -06:00 |
|
Wasim Yousef Said
|
765e1cfee2
|
Merge pull request #143 from unslothai/feature/local-models
feat: add schemas for local model discovery and listing
|
2026-02-17 13:10:04 -08:00 |
|
Shine1i
|
972cde7971
|
feat: add schemas for local model discovery and listing
|
2026-02-17 21:53:42 +01:00 |
|
Roland Tannous
|
c9fdce63e7
|
fix: skip sudo check on WSL during GGUF export to prevent password prompt hang
|
2026-02-17 19:30:02 +00:00 |
|
Roland Tannous
|
39b072d2ee
|
move requirements/ to studio/backend/ and update paths in setup.sh
|
2026-02-17 19:02:25 +00:00 |
|
Roland Tannous
|
299bc65e36
|
chore: override model defaults to use max_steps=30, save_steps=30, num_epochs=0 for testing
|
2026-02-17 18:46:23 +00:00 |
|
Roland Tannous
|
e818d97f24
|
Merge pull request #136 from unslothai/setup/update-setup-sh-dependencies
setup.sh: Replace inline pip installs with pinned requirements files
|
2026-02-17 22:04:03 +04:00 |
|
Shine1i
|
0be3e6f525
|
feat: integrate gradient norm tracking in training runtime and metrics
- Enhanced chart logic to filter and visualize finite gradient norm values.
|
2026-02-17 18:26:59 +01:00 |
|
Roland Tannous
|
e803e13d3e
|
add full dependency chain for unsloth + unsloth-extras
|
2026-02-17 14:21:18 +00:00 |
|