unsloth/studio/backend/core
Daniel Han 852755cc92 studio: fix stale GGUF metadata when switching models, update default helper model
Reset _supports_reasoning, _supports_tools, _chat_template, and
_context_length at the top of _read_gguf_metadata so flags from a
previously loaded model do not carry over. Without this, loading a
reasoning model (eg Qwen3.5-4B) then switching to a non-reasoning
model (eg Qwen3-4B-Instruct-2507) would keep supports_reasoning=True
from the first model, causing the UI to show "Thought for 0 seconds"
and passing --chat-template-kwargs enable_thinking to a model whose
chat template does not support it.

Also update the default helper GGUF from Qwen3-4B-Instruct-2507-GGUF
to Qwen3.5-4B-GGUF to match the frontend fallback auto-load model.
2026-03-17 07:46:21 +00:00
..
data_recipe Improve AI Assist: Update default model, model output parsing, logging, and dataset mapping UX (#4323) 2026-03-16 16:04:35 +04:00
export Final cleanup 2026-03-12 18:28:04 +00:00
inference studio: fix stale GGUF metadata when switching models, update default helper model 2026-03-17 07:46:21 +00:00
training studio: web search, KV cache dtype, training progress, inference fixes 2026-03-17 00:30:01 -07:00
__init__.py Final cleanup 2026-03-12 18:28:04 +00:00