Skip to main content

Module source

Module source 

Source
Expand description

Lazy tensor views let both checkpoint formats use the same packed writer. BF16 conversion holds one quantization group at a time, including when a fused expert tensor spans an entire layer. No converted checkpoint is staged.

Structsยง

Quantized ๐Ÿ”’
Source ๐Ÿ”’
Tensor ๐Ÿ”’

Enumsยง

Conversion ๐Ÿ”’
Part ๐Ÿ”’

Functionsยง

folded_norm ๐Ÿ”’
normalized_name ๐Ÿ”’
quantized_weight ๐Ÿ”’
write_norm ๐Ÿ”’