Model techniques map
Techniquesinference & servinginference quantization

specific method · filed under inference & serving

MSE-based scaling

Chooses quantization scales to minimize reconstruction error.

Also called MSE-based weight scaling.

sources
2
model
1
labs adopt it
0
strongest
evaluated

How sources treat it

One count per evidence span, weakest treatment to strongest.

evaluated 2

Documented in

Evidence

2 spans quoted from the sources, strongest treatment first.

MSE-based scaling minimizes reconstruction error

evaluatedpost trainingin Nemotron 3 UltraNVIDIA

we experimented with max-based, MSE-based, and Four-Over-Six scaling

evaluatedpost trainingin Nemotron 3 UltraNVIDIA

Filed alongside

Other methods under inference & serving :: inference quantization.