Expand description
Lazy tensor views let both checkpoint formats use the same packed writer. BF16 conversion holds one quantization group at a time, including when a fused expert tensor spans an entire layer. No converted checkpoint is staged.
Structsยง
Enumsยง
- Conversion ๐
- Part ๐
Functionsยง
- folded_
norm ๐ - normalized_
name ๐ - quantized_
weight ๐ - write_
norm ๐