Change all repetition_penalty defaults from 1.1 (or 1.05/1.2 in presets) to 1.0 across the entire backend and frontend. Most models handle repetition well on their own and a non-1.0 penalty can degrade output quality, especially for code, structured output, and creative tasks. Files changed: - Backend: inference.py, llama_cpp.py, orchestrator.py, worker.py, models/inference.py (Field defaults) - Frontend: chat-settings-sheet.tsx (Creative/Precise presets), runtime-provider.tsx (auto-title generation) |
||
|---|---|---|
| .. | ||
| .gitkeep | ||
| __init__.py | ||
| auth.py | ||
| data_recipe.py | ||
| datasets.py | ||
| export.py | ||
| inference.py | ||
| models.py | ||
| responses.py | ||
| training.py | ||
| users.py | ||