feat: add per-model FP8 layerwise casting for VRAM reduction - #8945
Merged
lstein merged 21 commits intoMay 12, 2026
Commits
Commits on Mar 6, 2026
Commits on Mar 9, 2026
Commits on Mar 11, 2026
Commits on Mar 20, 2026
Commits on Mar 21, 2026
- committed
- committed
- committed
- committed
Commits on Mar 26, 2026
Commits on Mar 31, 2026
Commits on May 11, 2026
- committed
- committed
- authored
- committed
- committed
- committed