|
__init__.py
|
vLLM Windows CUDA support [tested] (#2158)
|
2025-05-12 05:33:42 -07:00 |
|
_utils.py
|
Fix bugs
|
2025-06-20 06:09:03 -07:00 |
|
cohere.py
|
Fix renaming on other model than Llama (#2762)
|
2025-06-18 13:38:36 -07:00 |
|
dpo.py
|
Update dpo.py
|
2025-02-13 14:59:42 -08:00 |
|
gemma.py
|
Fix renaming on other model than Llama (#2762)
|
2025-06-18 13:38:36 -07:00 |
|
gemma2.py
|
Fix renaming on other model than Llama (#2762)
|
2025-06-18 13:38:36 -07:00 |
|
llama.py
|
Enable vLLM to share memory space (#2712)
|
2025-06-19 04:04:14 -07:00 |
|
llama4.py
|
Qwen 3
|
2025-05-02 03:09:44 -07:00 |
|
loader.py
|
Fix Whisper, ModernBERT (#2565)
|
2025-05-17 05:11:50 -07:00 |
|
loader_utils.py
|
Update loader_utils.py
|
2025-05-12 21:06:30 -07:00 |
|
mapper.py
|
Versioning
|
2025-06-10 06:51:07 -07:00 |
|
qwen2.py
|
Fix renaming on other model than Llama (#2762)
|
2025-06-18 13:38:36 -07:00 |
|
qwen3.py
|
Fix renaming on other model than Llama (#2762)
|
2025-06-18 13:38:36 -07:00 |
|
rl.py
|
Fix TRL 1.8.2 (#2774)
|
2025-06-20 06:28:58 -07:00 |
|
vision.py
|
Fix Whisper, ModernBERT (#2565)
|
2025-05-17 05:11:50 -07:00 |