ambiguous · filed under post-training
Extended post-training
A broad description attributing model gains to post-training extended beyond the unspecified baseline; the evidence does not identify a particular procedure.
- source
- 1
- model
- 1
- lab adopt it
- 1
- strongest
- used
How sources treat it
One count per evidence span, weakest treatment to strongest.
used 1
Documented in
Evidence
1 span quoted from the sources, strongest treatment first.
all gains come from extended post-training alone
usedpost trainingin GLM-5.3Z.ai
Filed alongside
Other methods under post-training.
Scalable RL at Agent ScaleReinforcement learning for low-pass-rate tasksScalable reinforcement learning post-trainingSFT followed by RL and on-policy distillationJoint SFT and RLPost-training on diverse domainsPost-training optimizationThree-stage post-training with multi-teacher on-policy distillationThree-stage Thinker post-training strategy