Review round follow-ups: - Drop the machine-specific HF_HOME defaults from the four bench / reproduction scripts (fp8_layer_ablation, hunyuan_int8_profile, quant_accuracy_sweep, video_speedmem_bench); they pointed at a private workspace cache and broke the scripts on any other machine. The standard HF_HOME env override still applies. - Correct the vae_quant 'auto' descriptions (image + video request fields, select_vae_quant_scheme docstring, loader comment) to match the shipped ladder: auto engages layerwise fp8 only; fp8_dynamic is an explicit opt-in and is never picked automatically. - Enforce _TE_FAMILY_SCHEME_DENY on the explicit text-encoder path too, gating the final concrete mode (so an int8 -> fp8 fallback is re-checked), matching the table's documented contract and the VAE module's behavior. Covered by a new test. Also merges origin/image-generation (single-GPU fit-budget fix) to keep the stacked head self-consistent. |
||
|---|---|---|
| .. | ||
| assets | ||
| auth | ||
| core | ||
| hub | ||
| loggers | ||
| models | ||
| plugins | ||
| requirements | ||
| routes | ||
| state | ||
| storage | ||
| tests | ||
| utils | ||
| __init__.py | ||
| _platform_compat.py | ||
| cloudflare_tunnel.py | ||
| colab.py | ||
| main.py | ||
| run.py | ||
| startup_banner.py | ||