Roland Tannous
|
909955767b
|
feat: add min_p sampling parameter to /chat/completions generation pipeline
|
2026-02-16 06:33:17 +00:00 |
|
Shine1i
|
571959e383
|
feat: add cancelation support for chat generation and streaming tasks
|
2026-02-15 18:23:27 +01:00 |
|
sshah229
|
9e50e167d9
|
added the inference fetching from model mappers
|
2026-02-15 02:48:53 -07:00 |
|
Roland Tannous
|
b334e49498
|
decouple reliance of backend on frontend for is_lora
|
2026-02-14 20:13:50 +00:00 |
|
Roland Tannous
|
f67ee58347
|
feat(inference): add use_adapter field for per-request adapter toggling in compare mode
|
2026-02-14 14:52:13 +00:00 |
|
Roland Tannous
|
9de38cb773
|
feat(inference): accept OpenAI multimodal content parts (image_url) in /chat/completions
|
2026-02-14 09:06:25 +00:00 |
|
Roland Tannous
|
8403190cdd
|
feat: add OpenAI-compatible POST /chat/completions endpoint with streaming and non-streaming support
|
2026-02-12 19:00:05 +00:00 |
|
Roland Tannous
|
7ee4381936
|
move inline pydantic models - fix existing models routes integration
|
2026-02-11 12:39:58 +00:00 |
|
Roland Tannous
|
b4ec0389f0
|
refactor/inference-api-routes-part-1
|
2026-02-03 16:57:57 +00:00 |
|
Roland Tannous
|
544d6944d1
|
root studio folder
|
2026-02-02 09:13:49 +00:00 |
|