specific method · filed under model architecture
Configurable visual token budget
A configurable number of visual tokens used to represent images at variable resolutions.
- sources
- 3
- model
- 1
- lab adopt it
- 1
- strongest
- optional
How sources treat it
One count per evidence span, weakest treatment to strongest.
Documented in
Evidence
3 spans quoted from the sources, strongest treatment first.
Gemma 4 supports variable image resolution through a configurable visual token budget, which controls how many tokens are used to represent an image.
Gemma 4 supports variable image resolution through a configurable visual token budget, which controls how many tokens are used to represent an image.
Gemma 4 supports variable image resolution through a configurable visual token budget, which controls how many tokens are used to represent an image.
Filed alongside
Other methods under model architecture :: multimodal architecture.