Commit graph

4,859 commits

Author SHA1 Message Date
Roland Tannous
7d8e991c1f added print statements for activate_lora_adapter 2026-02-14 17:25:37 +00:00
Roland Tannous
6fefbe9f0b swipped logger for print statements as logger isn't propagating 2026-02-14 17:21:26 +00:00
Roland Tannous
d0b94eae75 added logging 2026-02-14 17:09:07 +00:00
Roland Tannous
b5c8136957 exclude default from model.delete_adapter 2026-02-14 17:03:52 +00:00
Roland Tannous
35a6e40268 _apply_adapter_state now calls revert_to_base_model and activate_lora_adapter properly 2026-02-14 16:57:24 +00:00
Shine1i
2bd20d7d15 ignore tests 2026-02-14 16:42:58 +01:00
Shine1i
ca30e9f004 merge nightly 2026-02-14 16:42:20 +01:00
Shine1i
d28cc1670b lock 2026-02-14 16:37:38 +01:00
Shine1i
175fd0459c feat: add Jinja reference autocomplete components and enhance graph edges styling 2026-02-14 16:30:01 +01:00
Roland Tannous
f67ee58347 feat(inference): add use_adapter field for per-request adapter toggling in compare mode 2026-02-14 14:52:13 +00:00
Daniel Han
842099f2b0 Wrap models import with ROCm amdgpu ids fd2 filter (#4057)
Co-authored-by: Daniel Hanchen <danielhanchen@users.noreply.github.com>
2026-02-14 04:13:25 -08:00
Daniel Han
191cbe55ee Wrap unsloth_zoo import with HIP amdgpu.ids filter (#4056)
* Wrap unsloth_zoo import with HIP amdgpu.ids filter

* Refactor ROCm ids filter helpers for readability

* Rename ROCm ids filter helper and annotate call sites

* Remove obsolete amdgpu ids filter alias

* [pre-commit.ci] auto fixes from pre-commit.com hooks

for more information, see https://pre-commit.ci

---------

Co-authored-by: Daniel Hanchen <danielhanchen@users.noreply.github.com>
Co-authored-by: pre-commit-ci[bot] <66853113+pre-commit-ci[bot]@users.noreply.github.com>
2026-02-14 03:59:57 -08:00
Daniel Han
66db2a1417 Filter only amdgpu.ids fd2 noise during ROCm startup (#4054)
Co-authored-by: Daniel Hanchen <danielhanchen@users.noreply.github.com>
2026-02-14 03:35:41 -08:00
Daniel Han
66b09f2481 Make ROCm suppression detection robust for custom torch builds (#4053)
* Make ROCm suppression detection robust for custom torch builds

* Add ROCm detection debug logging behind UNSLOTH_ENABLE_LOGGING

---------

Co-authored-by: Daniel Hanchen <danielhanchen@users.noreply.github.com>
2026-02-14 02:59:49 -08:00
金黄色葡萄球君君
dd5ff9dcef ROCm: Add gfx950 (MI355X/CDNA4) to is_cdna() (#4051)
MI355X (gfx950) has the same 1024-thread workgroup limit as MI300X (gfx942),
but was missing from is_cdna(), causing all Triton kernels to use num_warps=32
(2048 threads) instead of 16 (1024 threads), resulting in OutOfResources crash.

Tested on: 8x AMD Instinct MI355X (gfx950), ROCm 7.1
2026-02-14 02:50:05 -08:00
Daniel Han
6ec46f49a6 Suppress HIP amdgpu.ids stderr noise during causal_conv1d check (#4052)
* Suppress HIP libdrm stderr noise in causal_conv1d probe

* Broaden HIP libdrm stderr suppression for early ROCm startup

---------

Co-authored-by: Daniel Hanchen <danielhanchen@users.noreply.github.com>
2026-02-14 02:44:34 -08:00
Daniel Han
1a929ce6f1 Simplify MI300X startup banner name (#4049)
* Improve HIP GPU name reporting in startup banner

* Drop MI300X arch suffix in banner name

* Normalize _utils.py file mode

* Simplify FA2 fallback text and filter AMD ids noise

* Strip trailing GPU arch suffix via regex

* [pre-commit.ci] auto fixes from pre-commit.com hooks

for more information, see https://pre-commit.ci

* Use gfx lookup default and normalize Ryzen AI naming

* Remove name-path Ryzen AI normalization

* Expand ROCm gfx map to full documented GPU name aliases

* Simplify HIP fallback naming to AMD gfx token

* Remove Ryzen Al torch_name normalization

---------

Co-authored-by: Daniel Hanchen <danielhanchen@users.noreply.github.com>
Co-authored-by: pre-commit-ci[bot] <66853113+pre-commit-ci[bot]@users.noreply.github.com>
2026-02-14 02:24:03 -08:00
Wasim Yousef Said
6ed179b459 Merge pull request #78 from unslothai/feature/vision-capabilities-chat
feat(chat): add vision image attachments for OpenAI-compatible chat
2026-02-14 02:01:17 -08:00
Shine1i
7611c7122c feat: add image handling support with Vision adapter and base64 serialization in chat runtime 2026-02-14 10:59:10 +01:00
Roland Tannous
c843a93797 Merge pull request #76 from unslothai/feature/inference-vision-openai-compatible
PR: OpenAI-compatible multimodal vision support + true vision streaming
2026-02-14 13:47:39 +04:00
Roland Tannous
418a374125 migrate _generate_vision_response to use TextIteratorStreamer + background thread 2026-02-14 09:30:32 +00:00
Roland Tannous
9de38cb773 feat(inference): accept OpenAI multimodal content parts (image_url) in /chat/completions 2026-02-14 09:06:25 +00:00
Roland Tannous
44d52b4103 Merge pull request #72 from unslothai/fix/sse-progress-timeout
Fix: SSE progress stream timeout during training
2026-02-14 09:58:14 +04:00
Roland Tannous
4f0fad2156 fix: increase SSE progress timeout to 30min and allow step-0 updates 2026-02-14 05:47:22 +00:00
Roland Tannous
957e39c50d Merge pull request #71 from unslothai/fix/frontend-default-path
Fix/frontend default path
2026-02-14 09:43:24 +04:00
Daniel Han
d3fcba134b Improve HIP GPU name detection in startup banner (#4048)
* Improve HIP GPU name reporting in startup banner

* [pre-commit.ci] auto fixes from pre-commit.com hooks

for more information, see https://pre-commit.ci

---------

Co-authored-by: Daniel Hanchen <danielhanchen@users.noreply.github.com>
Co-authored-by: pre-commit-ci[bot] <66853113+pre-commit-ci[bot]@users.noreply.github.com>
2026-02-13 21:32:34 -08:00
Roland Tannous
414d173624 replace function with alias 2026-02-14 05:28:47 +00:00
Roland Tannous
e09c5dfc7e fix alias command 2026-02-14 05:24:54 +00:00
Daniel Han
c14917b96e Handle broken causal_conv1d at import time (#4047)
* Handle broken causal_conv1d import at runtime

Add a startup import-time probe for causal_conv1d and disable the fast path when the shared library is ABI broken. This keeps Falcon H1/model loading resilient without requiring env flags.

- Add disable_broken_causal_conv1d in import_fixes.
- Invoke it early from unsloth/__init__ during package init.
- Make Falcon H1 optional imports in loader and models/__init__ soft-fail instead of failing hard.

* [pre-commit.ci] auto fixes from pre-commit.com hooks

for more information, see https://pre-commit.ci

* Enforce unavailable semantics for broken causal_conv1d

* [pre-commit.ci] auto fixes from pre-commit.com hooks

for more information, see https://pre-commit.ci

* Remove Falcon H1 import swallowing

* Restore optional Falcon H1 import guard

* [pre-commit.ci] auto fixes from pre-commit.com hooks

for more information, see https://pre-commit.ci

* Remove causal_conv1d regression tests

* Trim FA2 fallback messaging

---------

Co-authored-by: Daniel Hanchen <danielhanchen@users.noreply.github.com>
Co-authored-by: pre-commit-ci[bot] <66853113+pre-commit-ci[bot]@users.noreply.github.com>
2026-02-13 21:20:25 -08:00
Roland Tannous
7260c09268 change line arguments order in setup.sh 2026-02-14 05:13:49 +00:00
Roland Tannous
6b2a777f97 fix path in run_server 2026-02-14 05:06:17 +00:00
Roland Tannous
f07b919385 change default frontend path in run.py to studio/frontend/dist 2026-02-14 05:02:26 +00:00
Michael Han
2a7d098203 Update README with faster MoE.md
Adding MoE
2026-02-13 19:38:23 -08:00
Roland Tannous
933048d4f7 Merge pull request #70 from unslothai/feat/wire-custom-format-mapping-to-training
feat: wire `custom_format_mapping` through training pipeline
2026-02-14 01:09:35 +04:00
Roland Tannous
67edebfeb3 feat: wire custom_format_mapping through training pipeline to format_and_template_dataset 2026-02-13 21:07:36 +00:00
Roland Tannous
f12c5f61ef Merge pull request #69 from unslothai/fix/auto-detect-lora-in-model-config
fix: auto-detect LoRA adapters for both local and remote HF models in ModelConfig
2026-02-14 00:56:34 +04:00
Roland Tannous
8ce96df66f fix: auto-detect LoRA adapters for both local and remote HF models in ModelConfig 2026-02-13 20:54:40 +00:00
Roland Tannous
3d33753899 Merge pull request #68 from unslothai/fix/datasets-fix-vlm-detection
fix: auto-detect multimodal datasets in /check-format without requiri…
2026-02-14 00:04:51 +04:00
Wasim Yousef Said
bdcaf5518c Merge pull request #66 from unslothai/ui-fixes
UI fixes
2026-02-13 11:29:32 -08:00
Wasim Yousef Said
1f50b5b5f1 Merge pull request #67 from unslothai/style/polish-chat-sidebar-spacing
style: polish chat page spacing, typography adjustment, and panel alignment
2026-02-13 11:28:56 -08:00
Manan17
2210545493 fixing padding for titles 2026-02-13 19:22:39 +00:00
imagineer99
59acc087b6 style: polish chat page spacing, small typography, and panel alignment 2026-02-13 19:18:20 +00:00
Manan17
c0fbe7d4a5 Change of font space for title 2026-02-13 19:03:51 +00:00
Roland Tannous
a62a30bd9b Merge pull request #65 from unslothai/refactor/change-highlighted-text-color
Refactor/change highlighted text color
2026-02-13 21:36:08 +04:00
Roland Tannous
adf1ef5ea5 fix: auto-detect multimodal datasets in /check-format without requiring is_vlm flag 2026-02-13 17:29:39 +00:00
Manan17
acc00bd0a4 Changing the highlighted text color to be black while keeping the checkmark emerald 2026-02-13 17:24:57 +00:00
Wasim Yousef Said
6935fc3593 Merge pull request #63 from unslothai/feature/chat-openai-integration
feat(chat): integrate backend chat runtime + model load flow
2026-02-13 08:47:44 -08:00
Shine1i
ce7c9917ab feat: refactor suggestion handling and centralize defaults for thread UI 2026-02-13 17:42:45 +01:00
Shine1i
5fe4258401 feat: add warm-up indicator, new thread feature, and runtime improvements in chat UI 2026-02-13 17:28:01 +01:00
Roland Tannous
fe5a46ede3 Merge pull request #62 from unslothai/feature/add-easydict-addict
Feature/add easydict addict
2026-02-13 20:22:46 +04:00