unsloth/studio/backend/core/inference
2026-02-25 16:06:03 +04:00
..
__init__.py Add GGUF model inference via llama-server backend 2026-02-24 17:40:05 +04:00
inference.py Fixing base model export issue for vlms 2026-02-24 01:34:11 +00:00
llama_cpp.py Switch GGUF backend from /v1/completions to /v1/chat/completions 2026-02-24 19:21:01 +04:00