unsloth/studio/backend
Roland Tannous d96e3a7096 fix(vram): use training-aware estimates and backend arch-based VRAM for model fitness
- Replace loading-only VRAM formula with full training estimate (weights +
  LoRA adapters + optimizer states + gradients + activations + overhead)
  for all three methods: QLoRA, LoRA, full fine-tuning
- Expose architecture-based VRAM estimates from backend /api/models/config,
  reusing already-loaded AutoConfig to avoid extra HF round-trip
- Store per-method estimates in training config state; selected model badge
  uses authoritative backend estimate (handles MoE like gpt-oss-20b correctly)
- Replace file-size heuristic in autoSelectTrainingMethod with backend estimates
- Use total VRAM (not free) since chat models are offloaded before training
2026-04-01 04:55:51 +00:00
..
assets fix(studio): correct default weight_decay and learning rate (#4695) 2026-03-31 13:50:25 +04:00
auth fix: remove old comments (#4292) 2026-03-14 16:50:13 +04:00
core Studio: simplify tool-call dedup and replace html2text with builtin converter (#4722) 2026-03-31 06:15:18 -07:00
loggers Final cleanup 2026-03-12 18:28:04 +00:00
models fix(vram): use training-aware estimates and backend arch-based VRAM for model fitness 2026-04-01 04:55:51 +00:00
plugins Bump Data Designer to 0.5.4 (removes litellm dependency) (#4569) 2026-03-25 02:01:43 -07:00
requirements fix: no-torch install deps without pulling torch transitively (#4650) 2026-03-27 05:19:26 -07:00
routes fix(vram): use training-aware estimates and backend arch-based VRAM for model fitness 2026-04-01 04:55:51 +00:00
state Final cleanup 2026-03-12 18:28:04 +00:00
storage feat: custom scan folders for GGUF model discovery (#4723) 2026-03-31 06:40:31 -07:00
tests Studio: simplify tool-call dedup and replace html2text with builtin converter (#4722) 2026-03-31 06:15:18 -07:00
utils [studio] multi gpu: revert to balanced for inference. (#4698) 2026-03-31 01:24:41 -07:00
__init__.py Final cleanup 2026-03-12 18:28:04 +00:00
_platform_compat.py Fix Studio crash on Anaconda/conda-forge Python (#4484) 2026-03-22 05:36:55 -07:00
colab.py Allow install_python_stack to run on Colab (#4633) 2026-03-27 00:29:27 +04:00
main.py [Studio] multi gpu finetuning/inference via "balanced_low0/sequential" device_map (#4602) 2026-03-30 02:33:15 -07:00
run.py fix(studio): avoid UnicodeEncodeError on Windows cp1252 consoles (#4699) 2026-03-30 06:40:47 -07:00
startup_banner.py studio: unify Windows installer/setup logging style, verbosity controls, and startup messaging (#4651) 2026-03-30 00:53:23 -07:00