Roland Tannous
|
6b839a1481
|
feat: add min_p sampling parameter to /chat/completions generation pipeline
|
2026-02-16 06:33:17 +00:00 |
|
Shine1i
|
f6397bf1ac
|
feat: add cancelation support for chat generation and streaming tasks
|
2026-02-15 18:23:27 +01:00 |
|
sshah229
|
2483b98985
|
added the inference fetching from model mappers
|
2026-02-15 02:48:53 -07:00 |
|
Roland Tannous
|
ac8128519d
|
decouple reliance of backend on frontend for is_lora
|
2026-02-14 20:13:50 +00:00 |
|
Roland Tannous
|
8403bac48d
|
feat(inference): add use_adapter field for per-request adapter toggling in compare mode
|
2026-02-14 14:52:13 +00:00 |
|
Roland Tannous
|
480418b595
|
feat(inference): accept OpenAI multimodal content parts (image_url) in /chat/completions
|
2026-02-14 09:06:25 +00:00 |
|
Roland Tannous
|
c78cb11f81
|
feat: add OpenAI-compatible POST /chat/completions endpoint with streaming and non-streaming support
|
2026-02-12 19:00:05 +00:00 |
|
Roland Tannous
|
5ae20f6099
|
move inline pydantic models - fix existing models routes integration
|
2026-02-11 12:39:58 +00:00 |
|
Roland Tannous
|
c17ba10f96
|
refactor/inference-api-routes-part-1
|
2026-02-03 16:57:57 +00:00 |
|
Roland Tannous
|
75d8dcc824
|
root studio folder
|
2026-02-02 09:13:49 +00:00 |
|