Michael Han
f48e240bad
Merge pull request #2563 from davedgd/main
...
fix issue with qwen3 template double quote escapes
2025-05-16 22:38:39 -07:00
David Dobolyi
a063c4a41e
fix issue with qwen3 template double quote escapes
2025-05-16 23:26:03 -06:00
Etherll
99a2a36f73
Fix trust remote code ( #2357 )
...
* Update _utils.py
* Update loader.py
* Update loader.py
* Update vision.py
* Update unsloth/models/vision.py
* Update unsloth/models/vision.py
* Update unsloth/models/vision.py
* Update unsloth/models/vision.py
* Update unsloth/models/_utils.py
* Update unsloth/models/vision.py
---------
Co-authored-by: Daniel Han <danielhanchen@gmail.com>
2025-05-16 16:06:42 -07:00
Daniel Han
2524de493e
Update pyproject.toml
2025-05-16 15:38:19 -07:00
Daniel Han
15b6ac613a
Merge branch 'main' of https://github.com/unslothai/unsloth
2025-05-16 15:34:41 -07:00
Daniel Han
299c8a94a4
Update _utils.py
2025-05-16 15:33:49 -07:00
Michael Han
4937cd97f0
Merge pull request #2554 from Erland366/fix/generation_config
...
Quick fix on the CompileConfig error
2025-05-16 12:48:00 -07:00
Erland366
3cdbd879f7
Fix Nonetype on the compile_config
2025-05-16 13:16:34 +00:00
Michael Han
b22e654ef0
Update README.md
2025-05-16 01:56:40 -07:00
Michael Han
41e3701251
Update README.md
...
TTS support
2025-05-15 15:15:53 -07:00
Daniel Han
dc6c4dc385
TTS ( #2545 )
...
* Update rl_replacements.py
* Update vision.py
* Update rl_replacements.py
* Update vision.py
* Update vision.py
* Update vision.py
* Update vision.py
* Update vision.py
* Update vision.py
* Remove double generate patch
* Update vision.py
* Update vision.py
* Update vision.py
* Update vision.py
* Update vision.py
* Update mapper.py
* Update vision.py
* fix: config.torch_dtype in LlamaModel_fast_forward_inference (#2091 )
* fix: config.torch_dtype in LlamaModel_fast_forward_inference
* Update llama.py
* update for consistency
---------
Co-authored-by: Daniel Han <danielhanchen@gmail.com>
* versioning
* Update vision.py
* Update vision.py
* Update vision.py
* Update vision.py
* Update vision.py
* Update vision.py
* Update vision.py
* Update vision.py
* Update vision.py
* Update vision.py
* model_type_arch
* Update vision.py
* Update vision.py
* Update vision.py
* Update vision.py
* Update vision.py
* Update vision.py
* Update loader.py
* check
* Update _utils.py
* Update loader.py
* Update loader.py
* Remove prints
* Update README.md
typo
* Update _utils.py
* Update _utils.py
* versioning
* Update _utils.py
* Update _utils.py
* Update _utils.py
* Update llama.py
* Update llama.py
* Update llama.py
* Update llama.py
* Update llama.py
* Update llama.py
* Update llama.py
* Update llama.py
* Update llama.py
* Update llama.py
* Update llama.py
* Update llama.py
* Update llama.py
* Update llama.py
* Update llama.py
* Update llama.py
* Update llama.py
* Update llama.py
* Update llama.py
* Update llama.py
* Update llama.py
* Update llama.py
* Update llama.py
* Update llama.py
* Update llama.py
* Update vision.py
* HF Transfer
* fix(utils): add missing importlib import to fix NameError (#2134 )
This commit fixes a NameError that occurs when `importlib` is referenced in _utils.py
without being imported, especially when UNSLOTH_USE_MODELSCOPE=1 is enabled.
By adding the missing import statement, the code will no longer throw a NameError.
* Add QLoRA Train and Merge16bit Test (#2130 )
* add reference and unsloth lora merging tests
* add test / dataset printing to test scripts
* allow running tests from repo root
* add qlora test readme
* more readme edits
* ruff formatting
* additional readme comments
* forgot to add actual tests
* add apache license
* Update pyproject.toml
* Update vision.py
* Update vision.py
* Update vision.py
* Update vision.py
* Update loader.py
* Update loader.py
* Revert
* Update vision.py
* Update vision.py
* Update vision.py
* Update vision.py
* Update vision.py
* Bug fix
* Update mapper.py
* check SDPA for Mistral 3, Pixtral
* Update vision.py
* Versioning
* Update rl_replacements.py
* Update README.md
* add model registry
* move hf hub utils to unsloth/utils
* refactor global model info dicts to dataclasses
* fix dataclass init
* fix llama registration
* remove deprecated key function
* start registry reog
* add llama vision
* quant types -> Enum
* remap literal quant types to QuantType Enum
* add llama model registration
* fix quant tag mapping
* add qwen2.5 models to registry
* add option to include original model in registry
* handle quant types per model size
* separate registration of base and instruct llama3.2
* add QwenQVQ to registry
* add gemma3 to registry
* add phi
* add deepseek v3
* add deepseek r1 base
* add deepseek r1 zero
* add deepseek distill llama
* add deepseek distill models
* remove redundant code when constructing model names
* add mistral small to registry
* rename model registration methods
* rename deepseek registration methods
* refactor naming for mistral and phi
* add global register models
* refactor model registration tests for new registry apis
* add model search method
* remove deprecated registration api
* add quant type test
* add registry readme
* make llama registration more specific
* clear registry when executing individual model registration file
* more registry readme updates
* Update _auto_install.py
* Llama4
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Synthetic data
* Update mapper.py
* Xet and Synthetic
* Update synthetic.py
* Update loader.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update pyproject.toml
* Delete .gitignore
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update _utils.py
* Update pyproject.toml
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update synthetic.py
* Update chat_templates.py
* Seasame force float16 / float32
* Fix Seasame
* Update loader.py
* Update vision.py
* Update vision.py
* Update vision.py
* Update loader.py
* is_multimodal
* Update loader.py
* Update loader.py
* Update loader.py
* Update loader.py
* Update vision.py
* Update vision.py
* Update vision.py
* UNSLOTH_DISABLE_STATIC_GENERATION
* Update vision.py
* Auto vision detection
* Sesame
* Whisper
* Update loader.py
* Update loader.py
* Update loader.py
---------
Co-authored-by: lurf21 <93976703+lurf21@users.noreply.github.com>
Co-authored-by: Jack Shi Wei Lun <87535974+jackswl@users.noreply.github.com>
Co-authored-by: naliazheli <nalia0316@gmail.com>
Co-authored-by: jeromeku <jerome.ku@gmail.com>
Co-authored-by: Michael Han <107991372+shimmyshimmer@users.noreply.github.com>
2025-05-15 09:23:52 -07:00
Michael Han
b5ba71a3d3
Update README.md
2025-05-15 06:54:05 -07:00
Janusz
b781a7ad38
Add use_rslora reference to LoraConfig inititalisation ( #2539 )
...
Co-authored-by: jkumz <janusz.kumor01@gmail.com>
2025-05-15 04:24:18 -07:00
omahs
28304e4101
Fix typos ( #2540 )
2025-05-15 04:23:27 -07:00
Michael Han
b64c84ef33
Merge pull request #2537 from kiankyars/main
...
Add Qwen-3 chat template and Ollama template support
2025-05-14 20:45:53 -07:00
Kian Kyars
e147af330e
undo accident
2025-05-14 19:00:23 -06:00
Kian Kyars
059ccd8221
style: Place Qwen-3 template after Gemma-3, match style with other templates
2025-05-14 18:59:17 -06:00
Kian Kyars
48dc104728
Update Qwen-3 chat and Ollama templates to official full version, placed after Gemma-3
2025-05-14 18:42:55 -06:00
Kian Kyars
40ac241994
Add Qwen-3 chat template and Ollama template support
2025-05-14 18:35:02 -06:00
Daniel Han
e18a41d10f
Update pyproject.toml
2025-05-14 05:42:47 -07:00
Daniel Han
da41d4c21f
Update _utils.py
2025-05-14 05:42:11 -07:00
Daniel Han
0bbf131238
Update synthetic.py
2025-05-14 04:05:23 -07:00
Daniel Han
f971fce721
Update synthetic.py
2025-05-14 03:54:46 -07:00
Daniel Han
17d8517144
Update synthetic.py
2025-05-14 03:49:27 -07:00
Daniel Han
b47bbd3f55
Update synthetic.py
2025-05-14 03:47:29 -07:00
Daniel Han
e05735db0c
Update synthetic.py
2025-05-14 03:44:15 -07:00
Michael Han
074573a13b
Merge pull request #2527 from mmathew23/csm
...
Add Sesame CSM
2025-05-14 02:26:12 -07:00
DoubleMathew
e317fc222d
Merge branch 'unslothai:main' into csm
2025-05-13 17:58:11 -05:00
Daniel Han
cdb8eaaf42
Versioning
2025-05-13 09:10:24 -07:00
Michael Han
f4cbf303fe
Update README.md
2025-05-13 01:39:59 -07:00
Daniel Han
65710647b5
Update loader_utils.py
2025-05-12 21:06:30 -07:00
Daniel Han
e99b66e711
Update pyproject.toml
2025-05-12 16:28:50 -07:00
feng lui
48cb9c724c
vLLM Windows CUDA support [tested] ( #2158 )
...
* Update loader.py
change vllm installed check by transformers utils function
* Update llama.py
change vllm installed check by transformers utils function
* add sample notebook
* fix Indentation
* add global is_vLLM_available function
* Pythonic style
* Delete nb/Qwen2.5_(3B)-GRPO-windows.ipynb
Would be great to move it to https://github.com/unslothai/notebooks - appreciate it!
---------
Co-authored-by: Daniel Han <danielhanchen@gmail.com>
2025-05-12 05:33:42 -07:00
Daniel Han
67026e28ee
Update pyproject.toml
2025-05-12 04:24:18 -07:00
Daniel Han
cecdfc5a34
Fix Intel GPU
2025-05-12 03:10:21 -07:00
Lei Zhenyuan
fe6b83fd7e
[2/N] Enable intel GPU for unsloth ( #2388 )
...
* add DEVICE_TYPE and resolve device specific API
* reuse import torch
* move env under device typr
* resolve comments
* add more comments
* add more comments
2025-05-12 02:58:21 -07:00
Lei Zhenyuan
5bf77aabf5
first pr for intel GPU, resolve __init__.py and pyproject.toml ( #2350 )
...
add better comments
2025-05-12 02:38:40 -07:00
Daniel Han
c37380c63b
Fix GRPO eval
2025-05-12 02:35:31 -07:00
Michael Han
e3f6c5eff4
Merge pull request #2466 from mmathew23/fix_pop_token_type_ids
...
the pixtral vision notebook fails during inference
2025-05-09 14:57:31 -07:00
Mathew Mathew
3a56b3a24a
turn off compilation and fast generation for csm
2025-05-09 16:50:16 -05:00
Michael Han
1a99b4dc94
Merge pull request #2492 from yuanzhedong/yz/dev/fix-readme
...
Fix readme example
2025-05-07 23:29:59 -07:00
Yuanzhe Dong
75f3f8a7e5
Fix readme example
2025-05-06 19:26:35 -07:00
Michael Han
8821057420
Update README.md
...
Adding extra synthetic data notebook, cleaning repo
2025-05-05 20:56:01 -07:00
Daniel Han
6c0b8a57e4
Update __init__.py
2025-05-04 18:06:39 -07:00
Michael Han
9e2ef7c50c
Uploading HQ Unsloth Sticker
2025-05-04 05:31:57 -07:00
Michael Han
c4d0fd42be
Updating HQ logos
2025-05-04 05:25:06 -07:00
Daniel Han
84779ee11b
Better vllm deletion
2025-05-04 05:03:58 -07:00
Daniel Han
f9bf537130
Update pyproject.toml
2025-05-04 04:29:10 -07:00
Daniel Han
bad8069807
Update _utils.py
2025-05-04 03:03:50 -07:00
Michael Han
bb802c8a4a
Update README.md
2025-05-02 23:14:34 -07:00