Model techniques map
Techniquesmodel architecturemultimodal architecture

specific method · filed under model architecture

Configurable visual token budget

A configurable number of visual tokens used to represent images at variable resolutions.

sources
3
model
1
lab adopt it
1
strongest
optional

How sources treat it

One count per evidence span, weakest treatment to strongest.

optional 3

Documented in

Evidence

3 spans quoted from the sources, strongest treatment first.

Gemma 4 supports variable image resolution through a configurable visual token budget, which controls how many tokens are used to represent an image.

optionalinference servingin Gemma 4Google DeepMind

Gemma 4 supports variable image resolution through a configurable visual token budget, which controls how many tokens are used to represent an image.

optionalinference servingin Gemma 4Google DeepMind

Gemma 4 supports variable image resolution through a configurable visual token budget, which controls how many tokens are used to represent an image.

optionalinference servingin Gemma 4Google DeepMind

Filed alongside

Other methods under model architecture :: multimodal architecture.