Posts tagged “vision-language-models”
Cohere’s North Micro Vision Gives Documents Their Pixels Back
Cohere’s 2.4B North Micro Vision preserves native page detail for local OCR, but builders must supply the runtime and reasoning layer.
LFM2.5-VL-3B Gives the Edge Eyes. Keep Its Hands Tied.
Liquid AI’s LFM2.5-VL-3B brings screen grounding, OCR, and tool calling to local devices—but builders should separate sight from action.