Roland Tannous
|
834013aae5
|
Cap dataset.map num_proc on multi-GPU machines to prevent fork deadlocks
|
2026-02-23 14:25:31 +00:00 |
|
Roland Tannous
|
d94f842158
|
fix: error on >30% sample drop after train_on_responses_only instead of silent DataLoader crash
|
2026-02-23 12:21:06 +00:00 |
|
Roland Tannous
|
198433363a
|
feat: clear unsloth_compiled_cache on startup, shutdown, and between model loads
|
2026-02-23 07:26:22 +00:00 |
|
Roland Tannous
|
e666442b6e
|
fix: pass full Processor as processing_class for VLM SFTTrainer
|
2026-02-22 14:11:12 +00:00 |
|
Roland Tannous
|
ef118d0d05
|
fix: load proper vision processor from base model when FastVisionModel returns raw tokenizer, add tokenize=False to vision chat template
|
2026-02-21 04:40:29 +00:00 |
|
Manan17
|
e9710874e1
|
Mapping proper tokenizer for VLMs
|
2026-02-21 01:57:05 +00:00 |
|
Manan17
|
756aa56cd2
|
fixed the vlm's text only errors
|
2026-02-20 22:23:26 +00:00 |
|
Manan17
|
bd0cee8c15
|
Setting it to total cpu_count // 4
|
2026-02-20 06:32:01 +00:00 |
|
Manan17
|
444ece6b07
|
Fixing compare feature
|
2026-02-19 20:15:44 +00:00 |
|
Roland Tannous
|
c876b38780
|
reduce dataset_num_proc to 1/4 of cpu_count
|
2026-02-18 20:53:32 +00:00 |
|
Roland Tannous
|
23214c41c0
|
fix: unwrap ProcessorMixin to raw tokenizer for text-only SFTTrainer on VLM-architecture models
|
2026-02-18 19:13:20 +00:00 |
|
Roland Tannous
|
9840864662
|
Merge remote-tracking branch 'origin/nightly' into fix/dataset-mapping-vlm-text-datasets
|
2026-02-18 18:11:45 +04:00 |
|
Manan17
|
949e57c334
|
fixing the hangup of training after multiple back to back training processes
|
2026-02-18 08:18:13 +00:00 |
|
Manan17
|
db0fa1a270
|
Dividing the total cpu_count // 3
|
2026-02-18 07:59:57 +00:00 |
|
Manan17
|
c832c903b4
|
fix the linear path on backend
|
2026-02-18 07:08:32 +00:00 |
|
Roland Tannous
|
9d737559a5
|
debug statements
|
2026-02-18 00:37:09 +00:00 |
|
Roland Tannous
|
d94a1e0289
|
fix: normalize target_modules [all-linear] list to string for Unsloth/PEFT compatibility
|
2026-02-18 00:32:21 +00:00 |
|
Roland Tannous
|
f0613f5d07
|
fix: defensively rename VLM chat column to match model's forward() signature
|
2026-02-17 23:49:13 +00:00 |
|
Shine1i
|
b31461790f
|
feat: enhance training stop and reset flow with detailed checks
|
2026-02-17 23:32:22 +01:00 |
|
Roland Tannous
|
c5558312c8
|
fix: skip sudo check on WSL during GGUF export to prevent password prompt hang
|
2026-02-17 19:30:02 +00:00 |
|
Shine1i
|
f47c424be3
|
feat: integrate gradient norm tracking in training runtime and metrics
- Enhanced chart logic to filter and visualize finite gradient norm values.
|
2026-02-17 18:26:59 +01:00 |
|
Manan17
|
8f1db03c15
|
Adding metadata for checkpoints
|
2026-02-16 23:46:17 +00:00 |
|
Wasim Yousef Said
|
16f79a73de
|
Merge pull request #124 from unslothai/feature/bug-fixes
feat: support disabling top-k sampling with -1 and standardize normalization
|
2026-02-16 13:21:12 -08:00 |
|
Roland Tannous
|
108ec254cb
|
Merge branch 'nightly' into feature/eval-split-auto-detection
|
2026-02-17 01:11:30 +04:00 |
|
Shine1i
|
4be6eefed3
|
feat: support disabling top-k sampling with -1 and standardize normalization logic
- Updated top-k parameter range to accept -1 in models and frontend.
- Added utility to normalize top-k for backend compatibility.
|
2026-02-16 21:33:24 +01:00 |
|
Roland Tannous
|
18879a521b
|
feat: auto-detect model+dataset compatibility to select VLM vs LLM training path
|
2026-02-16 19:18:49 +00:00 |
|
Roland Tannous
|
3b117189c5
|
feat: add eval_enabled flag and format-first-then-split for eval dataset
|
2026-02-16 14:13:55 +00:00 |
|
Roland Tannous
|
5962bec41a
|
feat: add eval split auto-detection, eval_steps hyperparam, and eval_loss chart integration
|
2026-02-16 13:51:10 +00:00 |
|
Roland Tannous
|
90c3561adb
|
feat: add eval split auto-detection, eval_steps hyperparam, and eval_loss chart integration
|
2026-02-16 13:38:54 +00:00 |
|
Roland Tannous
|
8a239dc83e
|
refactor: move checkpoint scanning to utils/models and /checkpoints endpoint to models router
|
2026-02-16 09:32:11 +00:00 |
|
Roland Tannous
|
be584ccfa7
|
Merge pull request #97 from unslothai/fix/progress-metics
Resolved the progress metrics
|
2026-02-16 11:55:01 +04:00 |
|
sshah229
|
5bf2472af9
|
modified the num_tokens logic
|
2026-02-16 00:38:40 -07:00 |
|
Roland Tannous
|
6b839a1481
|
feat: add min_p sampling parameter to /chat/completions generation pipeline
|
2026-02-16 06:33:17 +00:00 |
|
Manan17
|
eae183504e
|
Fixing the get checkpoint api
|
2026-02-16 04:47:28 +00:00 |
|
Roland Tannous
|
38cb5c9496
|
feat: thread dataset subset/split params from API routes through to load_dataset calls
|
2026-02-16 03:56:22 +00:00 |
|
Shine1i
|
f6397bf1ac
|
feat: add cancelation support for chat generation and streaming tasks
|
2026-02-15 18:23:27 +01:00 |
|
sshah229
|
7fd55ce14f
|
resolved the prgress metrics
|
2026-02-15 05:35:32 -07:00 |
|
Manan17
|
e672458821
|
Adding save-steps to the SFTConfig
|
2026-02-15 09:37:54 +00:00 |
|
Manan17
|
df5f45058f
|
Fixing stuck training processes
|
2026-02-15 05:38:06 +00:00 |
|
Manan17
|
354b7d0aca
|
feat: add cancel or save and stop training
|
2026-02-15 00:00:22 +00:00 |
|
Roland Tannous
|
4399687f93
|
strip extra debug statements
|
2026-02-14 19:23:51 +00:00 |
|
Roland Tannous
|
4d868e8d2b
|
replace model unloading and peft loading mechanism for compare feature
|
2026-02-14 19:18:49 +00:00 |
|
Roland Tannous
|
225b3f1750
|
del model.peft_config instead of using model.delete_adapter
|
2026-02-14 17:32:15 +00:00 |
|
Roland Tannous
|
f122154cf3
|
added print statements for activate_lora_adapter
|
2026-02-14 17:25:37 +00:00 |
|
Roland Tannous
|
754ccf1a67
|
swipped logger for print statements as logger isn't propagating
|
2026-02-14 17:21:26 +00:00 |
|
Roland Tannous
|
b930a17b1d
|
added logging
|
2026-02-14 17:09:07 +00:00 |
|
Roland Tannous
|
9bee0a3f63
|
exclude default from model.delete_adapter
|
2026-02-14 17:03:52 +00:00 |
|
Roland Tannous
|
0b305fd822
|
_apply_adapter_state now calls revert_to_base_model and activate_lora_adapter properly
|
2026-02-14 16:57:24 +00:00 |
|
Roland Tannous
|
8403bac48d
|
feat(inference): add use_adapter field for per-request adapter toggling in compare mode
|
2026-02-14 14:52:13 +00:00 |
|
Roland Tannous
|
4ab8f81780
|
migrate _generate_vision_response to use TextIteratorStreamer + background thread
|
2026-02-14 09:30:32 +00:00 |
|