Change all repetition_penalty defaults from 1.1 (or 1.05/1.2 in presets) to 1.0 across the entire backend and frontend. Most models handle repetition well on their own and a non-1.0 penalty can degrade output quality, especially for code, structured output, and creative tasks. Files changed: - Backend: inference.py, llama_cpp.py, orchestrator.py, worker.py, models/inference.py (Field defaults) - Frontend: chat-settings-sheet.tsx (Creative/Precise presets), runtime-provider.tsx (auto-title generation) |
||
|---|---|---|
| .. | ||
| assets | ||
| auth | ||
| core | ||
| loggers | ||
| models | ||
| plugins | ||
| requirements | ||
| routes | ||
| state | ||
| tests | ||
| utils | ||
| __init__.py | ||
| colab.py | ||
| main.py | ||
| run.py | ||