Roland Tannous
cc11f066b1
feat: add tqdm progress bar to VLM conversion and download benchmark test
2026-03-04 13:30:27 +00:00
Roland Tannous
7804a4db2e
fix: add early probe to fail fast on datasets with too many broken image URLs
2026-03-04 08:05:40 +00:00
Roland Tannous
9cbd3d44a7
fix: use fsspec for URL image downloads with per-sample error handling
2026-03-04 07:50:55 +00:00
Roland Tannous
3c4bf80cc2
fix: cast URL image columns to HF Image() type in VLM conversion
2026-03-04 06:42:37 +00:00
Roland Tannous
c297d7aa84
Force num_proc=1 on Windows to avoid slow spawn overhead
2026-03-01 13:05:10 +00:00
Roland Tannous
986bef4f99
fix: support mmproj for local vision GGUF models + fix Windows pipe deadlock
2026-03-01 12:58:38 +00:00
Manan17
b4311cca82
Aggregating sharded models, showing fit/oom for quantizations
2026-02-27 08:23:15 +00:00
Manan17
4bd5213c05
Passes metadata to get model size
2026-02-27 07:38:55 +00:00
Roland Tannous
c06adc3878
Flatten GGUF subdirs in export and fix metadata lookup in scanner
2026-02-26 11:35:04 +04:00
Roland Tannous
b91b979bf8
Add GGUF tag for exported models in chat page selector
2026-02-25 19:01:47 +04:00
Roland Tannous
efaa0bacfb
Merge branch 'nightly' into feat/gguf-llama-cpp-inference
2026-02-25 16:06:03 +04:00
Manan17
5ce88f9aa1
fixing the chatml None error
2026-02-25 10:23:13 +00:00
Manan17
202dd082b7
My changes for dataset
2026-02-25 08:15:44 +00:00
Manan17
611febd2e3
adding custom mapping according to the chat templates
2026-02-25 07:56:30 +00:00
Roland Tannous
3ee4f1359a
Use llama-server -hf mode, add GGUF variant selector, fix vision detection
...
Replace Python-side GGUF download with llama-server's native -hf flag for
HuggingFace repos. Add frontend variant picker so users can choose
quantization (Q4_K_M, Q8_0, BF16, etc.) with file sizes. Fix vision
detection via mmproj files instead of hardcoding is_vision=False.
2026-02-24 19:03:06 +04:00
Roland Tannous
4a82e704aa
Preflight llama-server check before downloading remote GGUF files
2026-02-24 18:02:43 +04:00
Roland Tannous
c635d4f49c
Fix GGUF detection for HuggingFace repo IDs (not just local paths)
2026-02-24 17:49:09 +04:00
Roland Tannous
2f985ccbb5
Add GGUF model inference via llama-server backend
2026-02-24 17:40:05 +04:00
Manan17
1071c137f4
Adding exported model for chat
2026-02-24 01:17:09 +00:00
Roland Tannous
834013aae5
Cap dataset.map num_proc on multi-GPU machines to prevent fork deadlocks
2026-02-23 14:25:31 +00:00
Roland Tannous
198433363a
feat: clear unsloth_compiled_cache on startup, shutdown, and between model loads
2026-02-23 07:26:22 +00:00
Roland Tannous
62c260a109
Add GLM, Qwen3 MoE, TinyQwen3 MoE, and Ministral 3 VL model defaults and GLM train_on_responses_only mapping
2026-02-23 05:51:43 +00:00
Roland Tannous
11b3029dc6
Simplify dataset check to 2-tier, improve multimodal detection, auto-set trainOnCompletions, recheck dataset on reload
2026-02-19 11:25:54 +00:00
Manan17
2755cf922d
Passing use_auth = True and also having different checks which is missed by the is_vision function
2026-02-19 02:55:46 +00:00
Roland Tannous
29b25169c0
Scale dataset num_proc dynamically to cpu_count//3 instead of hardcap 8
2026-02-18 08:38:53 +00:00
Manan17
949e57c334
fixing the hangup of training after multiple back to back training processes
2026-02-18 08:18:13 +00:00
Roland Tannous
c0f210bc2a
fix: cap dataset.map() num_proc to 8 to prevent CUDA fork deadlocks
2026-02-17 23:12:45 +00:00
Manan17
8f1db03c15
Adding metadata for checkpoints
2026-02-16 23:46:17 +00:00
Roland Tannous
ed6d4b2fb6
feat: apply default chat template for base models without tokenizer chat_template
2026-02-16 15:56:06 +00:00
Roland Tannous
d49506b7b1
feat: add live GPU monitor with nvidia-smi polling during training
2026-02-16 11:47:43 +00:00
Roland Tannous
f3aa353540
feat: add GET /api/system/hardware endpoint for GPU info and package versions
2026-02-16 10:29:21 +00:00
Roland Tannous
1109839d2c
feat: include training loss per checkpoint in /api/models/checkpoints response
2026-02-16 09:50:28 +00:00
Roland Tannous
8a239dc83e
refactor: move checkpoint scanning to utils/models and /checkpoints endpoint to models router
2026-02-16 09:32:11 +00:00
Shine1i
78c7b6d7ba
fix lora: outputs path local
2026-02-15 16:58:24 +01:00
sshah229
2483b98985
added the inference fetching from model mappers
2026-02-15 02:48:53 -07:00
Roland Tannous
5652591d03
fix: auto-detect LoRA adapters for both local and remote HF models in ModelConfig
2026-02-13 20:54:40 +00:00
Roland Tannous
4e4fc367b6
fix: auto-detect multimodal datasets in /check-format without requiring is_vlm flag
2026-02-13 17:29:39 +00:00
sshah229
ae0b809adf
fixed the script directory
2026-02-12 21:55:36 -07:00
Roland Tannous
e31d4a3c40
replace torch MPS with MLX
2026-02-11 16:04:35 +00:00
Roland Tannous
a579a3d4ea
integrate global hardware detection at lifespan entrypoint
2026-02-11 15:34:26 +00:00
Roland Tannous
432c99ee7a
feat: add Apple Silicon (MPS) compatibility to backend utils + tests
2026-02-11 14:00:39 +00:00
Roland Tannous
528d2e27f0
authentication refactor - added setup token and token refresh mechanism
2026-02-11 12:09:47 +00:00
Roland Tannous
996b16f9ee
Add datasets check-format endpoint
2026-02-03 20:42:25 +00:00
Roland Tannous
4d89dd302c
fix custom_format_mapping flow for manual column mapping
2026-02-03 18:42:07 +00:00
Roland Tannous
fa3724b82c
Add Flag for Dataset Detection
2026-02-03 18:21:05 +00:00
Roland Tannous
842b2b3186
remove duplicates from dataset_utils.py
2026-02-03 18:03:01 +00:00
Roland Tannous
c17ba10f96
refactor/inference-api-routes-part-1
2026-02-03 16:57:57 +00:00
Roland Tannous
47ead076cf
Refactor [dataset_utils.py](cci:7://file:///home/support/new-ui-prototype/studio/backend/utils/datasets/dataset_utils.py:0:0-0:0) into focused modules
2026-02-03 14:38:02 +00:00
Roland Tannous
e390ca1092
fix: add utils/models directory that was ignored by gitignore
2026-02-02 19:52:34 +00:00
Roland Tannous
75d8dcc824
root studio folder
2026-02-02 09:13:49 +00:00