Roland Tannous
|
ef118d0d05
|
fix: load proper vision processor from base model when FastVisionModel returns raw tokenizer, add tokenize=False to vision chat template
|
2026-02-21 04:40:29 +00:00 |
|
Manan17
|
e9710874e1
|
Mapping proper tokenizer for VLMs
|
2026-02-21 01:57:05 +00:00 |
|
Manan17
|
756aa56cd2
|
fixed the vlm's text only errors
|
2026-02-20 22:23:26 +00:00 |
|
Manan17
|
444ece6b07
|
Fixing compare feature
|
2026-02-19 20:15:44 +00:00 |
|
Shine1i
|
4be6eefed3
|
feat: support disabling top-k sampling with -1 and standardize normalization logic
- Updated top-k parameter range to accept -1 in models and frontend.
- Added utility to normalize top-k for backend compatibility.
|
2026-02-16 21:33:24 +01:00 |
|
Roland Tannous
|
6b839a1481
|
feat: add min_p sampling parameter to /chat/completions generation pipeline
|
2026-02-16 06:33:17 +00:00 |
|
Shine1i
|
f6397bf1ac
|
feat: add cancelation support for chat generation and streaming tasks
|
2026-02-15 18:23:27 +01:00 |
|
Roland Tannous
|
4399687f93
|
strip extra debug statements
|
2026-02-14 19:23:51 +00:00 |
|
Roland Tannous
|
4d868e8d2b
|
replace model unloading and peft loading mechanism for compare feature
|
2026-02-14 19:18:49 +00:00 |
|
Roland Tannous
|
225b3f1750
|
del model.peft_config instead of using model.delete_adapter
|
2026-02-14 17:32:15 +00:00 |
|
Roland Tannous
|
f122154cf3
|
added print statements for activate_lora_adapter
|
2026-02-14 17:25:37 +00:00 |
|
Roland Tannous
|
754ccf1a67
|
swipped logger for print statements as logger isn't propagating
|
2026-02-14 17:21:26 +00:00 |
|
Roland Tannous
|
b930a17b1d
|
added logging
|
2026-02-14 17:09:07 +00:00 |
|
Roland Tannous
|
9bee0a3f63
|
exclude default from model.delete_adapter
|
2026-02-14 17:03:52 +00:00 |
|
Roland Tannous
|
0b305fd822
|
_apply_adapter_state now calls revert_to_base_model and activate_lora_adapter properly
|
2026-02-14 16:57:24 +00:00 |
|
Roland Tannous
|
8403bac48d
|
feat(inference): add use_adapter field for per-request adapter toggling in compare mode
|
2026-02-14 14:52:13 +00:00 |
|
Roland Tannous
|
4ab8f81780
|
migrate _generate_vision_response to use TextIteratorStreamer + background thread
|
2026-02-14 09:30:32 +00:00 |
|
Roland Tannous
|
a6ee9ee957
|
use get_device() for device selection and clear_gpu_cache() for GPU memory cleanup in inference, trainer, and export
|
2026-02-11 16:56:52 +00:00 |
|
Roland Tannous
|
c17ba10f96
|
refactor/inference-api-routes-part-1
|
2026-02-03 16:57:57 +00:00 |
|
Roland Tannous
|
47ead076cf
|
Refactor [dataset_utils.py](cci:7://file:///home/support/new-ui-prototype/studio/backend/utils/datasets/dataset_utils.py:0:0-0:0) into focused modules
|
2026-02-03 14:38:02 +00:00 |
|
Roland Tannous
|
75d8dcc824
|
root studio folder
|
2026-02-02 09:13:49 +00:00 |
|