unsloth/studio/backend/core/inference
2026-03-01 08:04:38 +00:00
..
__init__.py Add GGUF model inference via llama-server backend 2026-02-24 17:40:05 +04:00
audio_codecs.py revamping up the code and adding inference 2026-03-01 02:30:31 +00:00
inference.py code cleanup 2026-03-01 08:04:38 +00:00
llama_cpp.py Aggregating sharded models, showing fit/oom for quantizations 2026-02-27 08:23:15 +00:00