image routing

Cut token costs by routing text-heavy images away from pixel processing

Routing text-heavy images through pixel processing is a waste of VRAM and context space.

4 min readMachine Learning
Cut token costs by routing text-heavy images away from pixel processing
Cut token consumption by 88% with deterministic image routing (P50: 59ms) [P]
From Machine Learning

Recently I have been working on image attachment integration for my VEX agent runtime, and I came up with an idea for a significant optimization step.

Typically images attached to LLM context as raw bytes (base64 encoded string or publicly accessible URL), and the VLM processes image data by scaling it down to batches. But why attach an image as raw pixels if it primarily consists of text ?

Read the original at Machine Learning