Now that the upstream patch fixes have landed (#5319 for the three patch_* helpers, unsloth-zoo#628 for the MoE coverage canary), every observed cell-level red was one of those two things. Both are fixed, so re-run the matrix in strict mode: - Removed every per-step `continue-on-error: true`. A failing test step fails the cell. The previous green-with-fail-prints lie is gone. - Runtime patch ledger: was `assert REQUIRED helpers exist by name` (an inventory walk). Now also `assert len(fail) == 0` -- any zero-arg patch that raises is a real regression. NEEDS_PRECONDITION still skips the three patches that legitimately need real CUDA / runtime args. - patch_tiled_mlp shim: bumped seq_len from 4 to 192 with hidden=64 so divmod(192, 64) = (3, 0) and the tiled path actually runs 3 shards instead of degenerating to n_shards=1 (which is bit-exact and only confirms patching installed something). Added an explicit pre-assertion that we are exercising multi-shard. - openenv graceful-skip warning: previous text said "Weight reload still functional" which over-promised. Replaced with the literal consequence: duplicate `collective_rpc("reload_weights")` is not stripped and `wake_up(tags=["kv_cache"])` is not retagged. Most users are unaffected; openenv GRPO users on this TRL build may see redundant reload_weights or partial wake_up. Includes a merge of main into this branch so the consolidated cells pip-install the post-#5319 unsloth tree. |
||
|---|---|---|
| .. | ||
| __init__.py | ||
| _utils.py | ||
| cohere.py | ||
| dpo.py | ||
| falcon_h1.py | ||
| gemma.py | ||
| gemma2.py | ||
| glm4_moe.py | ||
| granite.py | ||
| llama.py | ||
| llama4.py | ||
| loader.py | ||
| loader_utils.py | ||
| mapper.py | ||
| mistral.py | ||
| qwen2.py | ||
| qwen3.py | ||
| qwen3_moe.py | ||
| rl.py | ||
| rl_replacements.py | ||
| sentence_transformer.py | ||
| vision.py | ||