Commit graph

76 commits

Author SHA1 Message Date
Manan17
19276ae60b Fixing the get checkpoint api 2026-02-16 04:47:28 +00:00
Roland Tannous
d0964652af feat: thread dataset subset/split params from API routes through to load_dataset calls 2026-02-16 03:56:22 +00:00
Shine1i
571959e383 feat: add cancelation support for chat generation and streaming tasks 2026-02-15 18:23:27 +01:00
Shine1i
8529f89a75 fix lora: outputs path local 2026-02-15 16:58:24 +01:00
Roland Tannous
4dbd77786e Merge pull request #96 from unslothai/feature/inference-yaml-ordered
Added the inference defaults for models
2026-02-15 17:43:31 +04:00
sshah229
9e50e167d9 added the inference fetching from model mappers 2026-02-15 02:48:53 -07:00
Manan17
6e4cde3bf8 Adding save-steps to the SFTConfig 2026-02-15 09:37:54 +00:00
sshah229
625bc1bbc6 added default inference config for default.yaml 2026-02-15 02:13:24 -07:00
sshah229
238fdc5c4a added default inference config from unsloth notebooks 2026-02-15 02:13:24 -07:00
sshah229
b5d93adcf2 added configs from Ollama 2026-02-15 02:13:24 -07:00
sshah229
ac20103e54 added inference defaults from unsloth guides 2026-02-15 02:13:24 -07:00
Roland Tannous
f453791916 Merge pull request #29 from unslothai/feature/export
Added the export routes and pydantic models
2026-02-15 12:42:01 +04:00
Roland Tannous
2ed9408b50 Merge pull request #83 from unslothai/fix/training-stuck-multiprocessing-cuda
Fixing stuck training processes
2026-02-15 10:18:49 +04:00
Roland Tannous
2840efcc08 fix: add no-cache headers to index.html to prevent stale frontend after rebuild 2026-02-15 06:06:27 +00:00
Manan17
6ccbc4edce Fixing stuck training processes 2026-02-15 05:38:06 +00:00
Manan17
97c6a09b84 feat: add cancel or save and stop training 2026-02-15 00:00:22 +00:00
Roland Tannous
b334e49498 decouple reliance of backend on frontend for is_lora 2026-02-14 20:13:50 +00:00
Roland Tannous
be3934860f strip extra debug statements 2026-02-14 19:23:51 +00:00
Roland Tannous
3ff3def555 replace model unloading and peft loading mechanism for compare feature 2026-02-14 19:18:49 +00:00
Roland Tannous
e7ae901737 del model.peft_config instead of using model.delete_adapter 2026-02-14 17:32:15 +00:00
Roland Tannous
7d8e991c1f added print statements for activate_lora_adapter 2026-02-14 17:25:37 +00:00
Roland Tannous
6fefbe9f0b swipped logger for print statements as logger isn't propagating 2026-02-14 17:21:26 +00:00
Roland Tannous
d0b94eae75 added logging 2026-02-14 17:09:07 +00:00
Roland Tannous
b5c8136957 exclude default from model.delete_adapter 2026-02-14 17:03:52 +00:00
Roland Tannous
35a6e40268 _apply_adapter_state now calls revert_to_base_model and activate_lora_adapter properly 2026-02-14 16:57:24 +00:00
Roland Tannous
f67ee58347 feat(inference): add use_adapter field for per-request adapter toggling in compare mode 2026-02-14 14:52:13 +00:00
Roland Tannous
418a374125 migrate _generate_vision_response to use TextIteratorStreamer + background thread 2026-02-14 09:30:32 +00:00
Roland Tannous
9de38cb773 feat(inference): accept OpenAI multimodal content parts (image_url) in /chat/completions 2026-02-14 09:06:25 +00:00
Roland Tannous
4f0fad2156 fix: increase SSE progress timeout to 30min and allow step-0 updates 2026-02-14 05:47:22 +00:00
Roland Tannous
6b2a777f97 fix path in run_server 2026-02-14 05:06:17 +00:00
Roland Tannous
f07b919385 change default frontend path in run.py to studio/frontend/dist 2026-02-14 05:02:26 +00:00
Roland Tannous
67edebfeb3 feat: wire custom_format_mapping through training pipeline to format_and_template_dataset 2026-02-13 21:07:36 +00:00
Roland Tannous
8ce96df66f fix: auto-detect LoRA adapters for both local and remote HF models in ModelConfig 2026-02-13 20:54:40 +00:00
Roland Tannous
adf1ef5ea5 fix: auto-detect multimodal datasets in /check-format without requiring is_vlm flag 2026-02-13 17:29:39 +00:00
Shine1i
d58fa17c81 feat: add support for serialized previews in dataset API and improve training initialization logging 2026-02-13 13:47:17 +01:00
Roland Tannous
5f155010f6 read external ips with fallback to standard notation 0.0.0.0 2026-02-13 10:28:20 +00:00
Roland Tannous
837596a9e7 feat: show external IP in startup banner 2026-02-13 10:23:12 +00:00
Roland Tannous
5602f7ccb4 fix: rollback auth.db user row if token generation fails during setup 2026-02-13 10:11:07 +00:00
Roland Tannous
f52bddc23f refactor: remove gradio dependency from training backend 2026-02-13 09:25:49 +00:00
Roland Tannous
75f775d088 fix: change epoch type from int to float to match TrainerState 2026-02-13 06:51:55 +00:00
sshah229
82be5b237f fixed the script directory 2026-02-12 21:55:36 -07:00
Roland Tannous
8403190cdd feat: add OpenAI-compatible POST /chat/completions endpoint with streaming and non-streaming support 2026-02-12 19:00:05 +00:00
Roland Tannous
509659ba97 feat: add SSE reconnection resilience with spec-compliant event fields, Last-Event-ID resume, and metric_history fallback in /status 2026-02-12 17:58:48 +00:00
Roland Tannous
dd71b0f18a return raw preview samples on format detection failure for manual column mapping 2026-02-12 15:39:56 +00:00
Roland Tannous
4c791bd5aa feat(datasets): check-format to return preview samples 2026-02-12 11:25:47 +00:00
sshah229
ee703dd6c6 added router in main 2026-02-11 18:51:53 -07:00
sshah229
40bfe42974 added the pydantic models and routes for export 2026-02-11 18:34:12 -07:00
Roland Tannous
da1cde971c use get_device() for device selection and clear_gpu_cache() for GPU memory cleanup in inference, trainer, and export 2026-02-11 16:56:52 +00:00
Roland Tannous
1a6bfe51b6 added @needs_torch to test_cuda_oom 2026-02-11 16:12:37 +00:00
Roland Tannous
95038d6129 add @needs_mlx decorator on tests 2026-02-11 16:10:22 +00:00