Model techniques map
Techniquesinference & servinginference quantization

specific method · filed under inference & serving

AWQ INT4 weight quantization (W4A16)

Uses AWQ to quantize weights to INT4 while retaining higher-precision activations, with calibration on agentic trajectories in the cited setup.

Also called INT4 AWQ weight quantization.

source
1
model
1
lab adopt it
1
strongest
used

How sources treat it

One count per evidence span, weakest treatment to strongest.

used 1

Documented in

Evidence

1 span quoted from the sources, strongest treatment first.

For INT4 weight quantization (W4A16), we applied AWQ using a calibration set of 128 long-context agentic trajectories. This approach initially introduced a non-negotiable quality drop.

usedpost trainingin Laguna XS.2Poolside

Filed alongside

Other methods under inference & serving :: inference quantization.