Roland Tannous
|
d69431fa57
|
Scale dataset num_proc dynamically to cpu_count//3 instead of hardcap 8
|
2026-02-18 08:38:53 +00:00 |
|
Manan17
|
76cd1dc24c
|
fixing the hangup of training after multiple back to back training processes
|
2026-02-18 08:18:13 +00:00 |
|
Manan17
|
c37bf686a6
|
Dividing the total cpu_count // 3
|
2026-02-18 07:59:57 +00:00 |
|
Roland Tannous
|
14edb08cf5
|
Merge pull request #148 from unslothai/fix/linear
linear fix
|
2026-02-18 11:10:16 +04:00 |
|
Manan17
|
58116e7e7a
|
fix the linear path on backend
|
2026-02-18 07:08:32 +00:00 |
|
Roland Tannous
|
d2f7eaf085
|
Merge pull request #145 from unslothai/fix/check-format-sample
fix: stream HF datasets in check-format endpoint to avoid full d…
|
2026-02-18 10:49:17 +04:00 |
|
Roland Tannous
|
3d0d1c7020
|
fix: cap dataset.map() num_proc to 8 to prevent CUDA fork deadlocks
|
2026-02-17 23:12:45 +00:00 |
|
Roland Tannous
|
5864dece26
|
fix: fix: stream HF datasets in check-format endpoint to avoid full downloads; add info logging to model config endpoints
|
2026-02-17 22:53:29 +00:00 |
|
Shine1i
|
dc0cec772d
|
feat: enhance training stop and reset flow with detailed checks
|
2026-02-17 23:32:22 +01:00 |
|
Wasim Yousef Said
|
765e1cfee2
|
Merge pull request #143 from unslothai/feature/local-models
feat: add schemas for local model discovery and listing
|
2026-02-17 13:10:04 -08:00 |
|
Shine1i
|
972cde7971
|
feat: add schemas for local model discovery and listing
|
2026-02-17 21:53:42 +01:00 |
|
Roland Tannous
|
c9fdce63e7
|
fix: skip sudo check on WSL during GGUF export to prevent password prompt hang
|
2026-02-17 19:30:02 +00:00 |
|
Roland Tannous
|
39b072d2ee
|
move requirements/ to studio/backend/ and update paths in setup.sh
|
2026-02-17 19:02:25 +00:00 |
|
Roland Tannous
|
299bc65e36
|
chore: override model defaults to use max_steps=30, save_steps=30, num_epochs=0 for testing
|
2026-02-17 18:46:23 +00:00 |
|
Roland Tannous
|
e818d97f24
|
Merge pull request #136 from unslothai/setup/update-setup-sh-dependencies
setup.sh: Replace inline pip installs with pinned requirements files
|
2026-02-17 22:04:03 +04:00 |
|
Shine1i
|
0be3e6f525
|
feat: integrate gradient norm tracking in training runtime and metrics
- Enhanced chart logic to filter and visualize finite gradient norm values.
|
2026-02-17 18:26:59 +01:00 |
|
Roland Tannous
|
e803e13d3e
|
add full dependency chain for unsloth + unsloth-extras
|
2026-02-17 14:21:18 +00:00 |
|
Manan17
|
c7b7ecab4f
|
Adding metadata for checkpoints
|
2026-02-16 23:46:17 +00:00 |
|
Wasim Yousef Said
|
7b8220598e
|
Merge pull request #124 from unslothai/feature/bug-fixes
feat: support disabling top-k sampling with -1 and standardize normalization
|
2026-02-16 13:21:12 -08:00 |
|
Roland Tannous
|
ff0aec180a
|
Merge branch 'nightly' into feature/eval-split-auto-detection
|
2026-02-17 01:11:30 +04:00 |
|
Shine1i
|
0db7da96cc
|
feat: support disabling top-k sampling with -1 and standardize normalization logic
- Updated top-k parameter range to accept -1 in models and frontend.
- Added utility to normalize top-k for backend compatibility.
|
2026-02-16 21:33:24 +01:00 |
|
Roland Tannous
|
fa0ca59215
|
feat: auto-detect model+dataset compatibility to select VLM vs LLM training path
|
2026-02-16 19:18:49 +00:00 |
|
Roland Tannous
|
b32ad350c5
|
feat: apply default chat template for base models without tokenizer chat_template
|
2026-02-16 15:56:06 +00:00 |
|
Roland Tannous
|
5df3a0b250
|
feat: add eval_enabled flag and format-first-then-split for eval dataset
|
2026-02-16 14:13:55 +00:00 |
|
Roland Tannous
|
0aea3f149d
|
feat: add eval split auto-detection, eval_steps hyperparam, and eval_loss chart integration
|
2026-02-16 13:51:10 +00:00 |
|
Roland Tannous
|
37452d56cf
|
feat: add eval split auto-detection, eval_steps hyperparam, and eval_loss chart integration
|
2026-02-16 13:38:54 +00:00 |
|
Roland Tannous
|
a0ebd9183a
|
feat: add live GPU monitor with nvidia-smi polling during training
|
2026-02-16 11:47:43 +00:00 |
|
Roland Tannous
|
b20d50e8d0
|
feat: add GET /api/system/hardware endpoint for GPU info and package versions
|
2026-02-16 10:29:21 +00:00 |
|
Roland Tannous
|
fd49c56481
|
feat: include training loss per checkpoint in /api/models/checkpoints response
|
2026-02-16 09:50:28 +00:00 |
|
Roland Tannous
|
f0298edeb8
|
refactor: move checkpoint scanning to utils/models and /checkpoints endpoint to models router
|
2026-02-16 09:32:11 +00:00 |
|
Roland Tannous
|
6ecc03485d
|
Merge pull request #97 from unslothai/fix/progress-metics
Resolved the progress metrics
|
2026-02-16 11:55:01 +04:00 |
|
sshah229
|
63b34660ed
|
modified the num_tokens logic
|
2026-02-16 00:38:40 -07:00 |
|
Roland Tannous
|
909955767b
|
feat: add min_p sampling parameter to /chat/completions generation pipeline
|
2026-02-16 06:33:17 +00:00 |
|
Manan17
|
19276ae60b
|
Fixing the get checkpoint api
|
2026-02-16 04:47:28 +00:00 |
|
Roland Tannous
|
d0964652af
|
feat: thread dataset subset/split params from API routes through to load_dataset calls
|
2026-02-16 03:56:22 +00:00 |
|
Shine1i
|
571959e383
|
feat: add cancelation support for chat generation and streaming tasks
|
2026-02-15 18:23:27 +01:00 |
|
Shine1i
|
8529f89a75
|
fix lora: outputs path local
|
2026-02-15 16:58:24 +01:00 |
|
Roland Tannous
|
4dbd77786e
|
Merge pull request #96 from unslothai/feature/inference-yaml-ordered
Added the inference defaults for models
|
2026-02-15 17:43:31 +04:00 |
|
sshah229
|
0b1c635b43
|
resolved the prgress metrics
|
2026-02-15 05:35:32 -07:00 |
|
sshah229
|
9e50e167d9
|
added the inference fetching from model mappers
|
2026-02-15 02:48:53 -07:00 |
|
Manan17
|
6e4cde3bf8
|
Adding save-steps to the SFTConfig
|
2026-02-15 09:37:54 +00:00 |
|
sshah229
|
625bc1bbc6
|
added default inference config for default.yaml
|
2026-02-15 02:13:24 -07:00 |
|
sshah229
|
238fdc5c4a
|
added default inference config from unsloth notebooks
|
2026-02-15 02:13:24 -07:00 |
|
sshah229
|
b5d93adcf2
|
added configs from Ollama
|
2026-02-15 02:13:24 -07:00 |
|
sshah229
|
ac20103e54
|
added inference defaults from unsloth guides
|
2026-02-15 02:13:24 -07:00 |
|
Roland Tannous
|
f453791916
|
Merge pull request #29 from unslothai/feature/export
Added the export routes and pydantic models
|
2026-02-15 12:42:01 +04:00 |
|
Roland Tannous
|
2ed9408b50
|
Merge pull request #83 from unslothai/fix/training-stuck-multiprocessing-cuda
Fixing stuck training processes
|
2026-02-15 10:18:49 +04:00 |
|
Roland Tannous
|
2840efcc08
|
fix: add no-cache headers to index.html to prevent stale frontend after rebuild
|
2026-02-15 06:06:27 +00:00 |
|
Manan17
|
6ccbc4edce
|
Fixing stuck training processes
|
2026-02-15 05:38:06 +00:00 |
|
Manan17
|
97c6a09b84
|
feat: add cancel or save and stop training
|
2026-02-15 00:00:22 +00:00 |
|