Seed each request with the model's recommended sampling (matching the Chat UI), add per-field override flags, ignore oversized overrides, warn when sampling pins cannot apply to a reused server, and apply pins to the completions endpoint. |
||
|---|---|---|
| .. | ||
| __init__.py | ||
| _password_prompt.py | ||
| chat.py | ||
| export.py | ||
| inference.py | ||
| start.py | ||
| studio.py | ||
| train.py | ||