Image decision model glance-qwen3-vl-4b comes to Hugging Face and Replicate
A user reports roughly 220 milliseconds of compute while running the model locally on two RTX 3080 Ti GPUs to filter Pinterest pins, asking three questions per image.
TLDR
A post announcing glance-qwen3-vl-4b on Hugging Face and Replicate describes a “Hot dog or not?” demo at 0.4 seconds. A user reports running it locally on two RTX 3080 Ti GPUs to filter Pinterest pins at scale, checking text overlays, usability and relevance—three questions per image—with roughly 220 milliseconds of compute.
