Commit graph

4,672 commits

Author SHA1 Message Date
Manan17
6ccbc4edce Fixing stuck training processes 2026-02-15 05:38:06 +00:00
Roland Tannous
2c5a4359c6 Merge pull request #81 from unslothai/feature/early-stop-or-cancel-training-ui
feat: early stop or cancel training UI
2026-02-15 08:28:39 +04:00
Manan17
52cbf9b699 feat: UI for cancel or save and stop training 2026-02-15 00:22:28 +00:00
Manan17
97c6a09b84 feat: add cancel or save and stop training 2026-02-15 00:00:22 +00:00
Shine1i
070b9d41a4 feat: enhance sampler builders with datetime unit mapping, uuid format handling, and error reporting 2026-02-14 23:07:11 +01:00
Roland Tannous
1d362a36c3 Merge pull request #79 from unslothai/feat/compare-use-adapter
Adapter Toggling for chat compare feature
2026-02-15 00:25:53 +04:00
Roland Tannous
b334e49498 decouple reliance of backend on frontend for is_lora 2026-02-14 20:13:50 +00:00
Roland Tannous
be3934860f strip extra debug statements 2026-02-14 19:23:51 +00:00
Roland Tannous
3ff3def555 replace model unloading and peft loading mechanism for compare feature 2026-02-14 19:18:49 +00:00
Shine1i
ebc411e508 chore: ignore agent.md 2026-02-14 19:22:07 +01:00
Shine1i
964f7d1548 feat: refactor block definitions and utilities into modular components for enhanced maintainability 2026-02-14 19:13:19 +01:00
Shine1i
f2a00d6e44 feat: add seed dataset support with configuration, preview, and builder utilities 2026-02-14 18:44:38 +01:00
Roland Tannous
e7ae901737 del model.peft_config instead of using model.delete_adapter 2026-02-14 17:32:15 +00:00
Roland Tannous
7d8e991c1f added print statements for activate_lora_adapter 2026-02-14 17:25:37 +00:00
Roland Tannous
6fefbe9f0b swipped logger for print statements as logger isn't propagating 2026-02-14 17:21:26 +00:00
Roland Tannous
d0b94eae75 added logging 2026-02-14 17:09:07 +00:00
Roland Tannous
b5c8136957 exclude default from model.delete_adapter 2026-02-14 17:03:52 +00:00
Roland Tannous
35a6e40268 _apply_adapter_state now calls revert_to_base_model and activate_lora_adapter properly 2026-02-14 16:57:24 +00:00
Shine1i
2bd20d7d15 ignore tests 2026-02-14 16:42:58 +01:00
Shine1i
ca30e9f004 merge nightly 2026-02-14 16:42:20 +01:00
Shine1i
d28cc1670b lock 2026-02-14 16:37:38 +01:00
Shine1i
175fd0459c feat: add Jinja reference autocomplete components and enhance graph edges styling 2026-02-14 16:30:01 +01:00
Roland Tannous
f67ee58347 feat(inference): add use_adapter field for per-request adapter toggling in compare mode 2026-02-14 14:52:13 +00:00
Daniel Han
842099f2b0 Wrap models import with ROCm amdgpu ids fd2 filter (#4057)
Co-authored-by: Daniel Hanchen <danielhanchen@users.noreply.github.com>
2026-02-14 04:13:25 -08:00
Daniel Han
191cbe55ee Wrap unsloth_zoo import with HIP amdgpu.ids filter (#4056)
* Wrap unsloth_zoo import with HIP amdgpu.ids filter

* Refactor ROCm ids filter helpers for readability

* Rename ROCm ids filter helper and annotate call sites

* Remove obsolete amdgpu ids filter alias

* [pre-commit.ci] auto fixes from pre-commit.com hooks

for more information, see https://pre-commit.ci

---------

Co-authored-by: Daniel Hanchen <danielhanchen@users.noreply.github.com>
Co-authored-by: pre-commit-ci[bot] <66853113+pre-commit-ci[bot]@users.noreply.github.com>
2026-02-14 03:59:57 -08:00
Daniel Han
66db2a1417 Filter only amdgpu.ids fd2 noise during ROCm startup (#4054)
Co-authored-by: Daniel Hanchen <danielhanchen@users.noreply.github.com>
2026-02-14 03:35:41 -08:00
Daniel Han
66b09f2481 Make ROCm suppression detection robust for custom torch builds (#4053)
* Make ROCm suppression detection robust for custom torch builds

* Add ROCm detection debug logging behind UNSLOTH_ENABLE_LOGGING

---------

Co-authored-by: Daniel Hanchen <danielhanchen@users.noreply.github.com>
2026-02-14 02:59:49 -08:00
金黄色葡萄球君君
dd5ff9dcef ROCm: Add gfx950 (MI355X/CDNA4) to is_cdna() (#4051)
MI355X (gfx950) has the same 1024-thread workgroup limit as MI300X (gfx942),
but was missing from is_cdna(), causing all Triton kernels to use num_warps=32
(2048 threads) instead of 16 (1024 threads), resulting in OutOfResources crash.

Tested on: 8x AMD Instinct MI355X (gfx950), ROCm 7.1
2026-02-14 02:50:05 -08:00
Daniel Han
6ec46f49a6 Suppress HIP amdgpu.ids stderr noise during causal_conv1d check (#4052)
* Suppress HIP libdrm stderr noise in causal_conv1d probe

* Broaden HIP libdrm stderr suppression for early ROCm startup

---------

Co-authored-by: Daniel Hanchen <danielhanchen@users.noreply.github.com>
2026-02-14 02:44:34 -08:00
Daniel Han
1a929ce6f1 Simplify MI300X startup banner name (#4049)
* Improve HIP GPU name reporting in startup banner

* Drop MI300X arch suffix in banner name

* Normalize _utils.py file mode

* Simplify FA2 fallback text and filter AMD ids noise

* Strip trailing GPU arch suffix via regex

* [pre-commit.ci] auto fixes from pre-commit.com hooks

for more information, see https://pre-commit.ci

* Use gfx lookup default and normalize Ryzen AI naming

* Remove name-path Ryzen AI normalization

* Expand ROCm gfx map to full documented GPU name aliases

* Simplify HIP fallback naming to AMD gfx token

* Remove Ryzen Al torch_name normalization

---------

Co-authored-by: Daniel Hanchen <danielhanchen@users.noreply.github.com>
Co-authored-by: pre-commit-ci[bot] <66853113+pre-commit-ci[bot]@users.noreply.github.com>
2026-02-14 02:24:03 -08:00
Wasim Yousef Said
6ed179b459 Merge pull request #78 from unslothai/feature/vision-capabilities-chat
feat(chat): add vision image attachments for OpenAI-compatible chat
2026-02-14 02:01:17 -08:00
Shine1i
7611c7122c feat: add image handling support with Vision adapter and base64 serialization in chat runtime 2026-02-14 10:59:10 +01:00
Roland Tannous
c843a93797 Merge pull request #76 from unslothai/feature/inference-vision-openai-compatible
PR: OpenAI-compatible multimodal vision support + true vision streaming
2026-02-14 13:47:39 +04:00
Roland Tannous
418a374125 migrate _generate_vision_response to use TextIteratorStreamer + background thread 2026-02-14 09:30:32 +00:00
Roland Tannous
9de38cb773 feat(inference): accept OpenAI multimodal content parts (image_url) in /chat/completions 2026-02-14 09:06:25 +00:00
Roland Tannous
44d52b4103 Merge pull request #72 from unslothai/fix/sse-progress-timeout
Fix: SSE progress stream timeout during training
2026-02-14 09:58:14 +04:00
Roland Tannous
4f0fad2156 fix: increase SSE progress timeout to 30min and allow step-0 updates 2026-02-14 05:47:22 +00:00
Roland Tannous
957e39c50d Merge pull request #71 from unslothai/fix/frontend-default-path
Fix/frontend default path
2026-02-14 09:43:24 +04:00
Daniel Han
d3fcba134b Improve HIP GPU name detection in startup banner (#4048)
* Improve HIP GPU name reporting in startup banner

* [pre-commit.ci] auto fixes from pre-commit.com hooks

for more information, see https://pre-commit.ci

---------

Co-authored-by: Daniel Hanchen <danielhanchen@users.noreply.github.com>
Co-authored-by: pre-commit-ci[bot] <66853113+pre-commit-ci[bot]@users.noreply.github.com>
2026-02-13 21:32:34 -08:00
Roland Tannous
414d173624 replace function with alias 2026-02-14 05:28:47 +00:00
Roland Tannous
e09c5dfc7e fix alias command 2026-02-14 05:24:54 +00:00
Daniel Han
c14917b96e Handle broken causal_conv1d at import time (#4047)
* Handle broken causal_conv1d import at runtime

Add a startup import-time probe for causal_conv1d and disable the fast path when the shared library is ABI broken. This keeps Falcon H1/model loading resilient without requiring env flags.

- Add disable_broken_causal_conv1d in import_fixes.
- Invoke it early from unsloth/__init__ during package init.
- Make Falcon H1 optional imports in loader and models/__init__ soft-fail instead of failing hard.

* [pre-commit.ci] auto fixes from pre-commit.com hooks

for more information, see https://pre-commit.ci

* Enforce unavailable semantics for broken causal_conv1d

* [pre-commit.ci] auto fixes from pre-commit.com hooks

for more information, see https://pre-commit.ci

* Remove Falcon H1 import swallowing

* Restore optional Falcon H1 import guard

* [pre-commit.ci] auto fixes from pre-commit.com hooks

for more information, see https://pre-commit.ci

* Remove causal_conv1d regression tests

* Trim FA2 fallback messaging

---------

Co-authored-by: Daniel Hanchen <danielhanchen@users.noreply.github.com>
Co-authored-by: pre-commit-ci[bot] <66853113+pre-commit-ci[bot]@users.noreply.github.com>
2026-02-13 21:20:25 -08:00
Roland Tannous
7260c09268 change line arguments order in setup.sh 2026-02-14 05:13:49 +00:00
Roland Tannous
6b2a777f97 fix path in run_server 2026-02-14 05:06:17 +00:00
Roland Tannous
f07b919385 change default frontend path in run.py to studio/frontend/dist 2026-02-14 05:02:26 +00:00
Michael Han
2a7d098203 Update README with faster MoE.md
Adding MoE
2026-02-13 19:38:23 -08:00
Roland Tannous
933048d4f7 Merge pull request #70 from unslothai/feat/wire-custom-format-mapping-to-training
feat: wire `custom_format_mapping` through training pipeline
2026-02-14 01:09:35 +04:00
Roland Tannous
67edebfeb3 feat: wire custom_format_mapping through training pipeline to format_and_template_dataset 2026-02-13 21:07:36 +00:00
Roland Tannous
f12c5f61ef Merge pull request #69 from unslothai/fix/auto-detect-lora-in-model-config
fix: auto-detect LoRA adapters for both local and remote HF models in ModelConfig
2026-02-14 00:56:34 +04:00
Roland Tannous
8ce96df66f fix: auto-detect LoRA adapters for both local and remote HF models in ModelConfig 2026-02-13 20:54:40 +00:00