

Ultra-efficient 1.3B vision-language model for mobile
MiniCPM-V 4.6 is an open MLLM for image and video understanding on phones and consumer hardware, with mixed 4x/16x visual token compression, iOS/Android/HarmonyOS demos, and support for vLLM, SGLang, llama.cpp, and Ollama.
No comments yet. Be the first!
Real conversations about MiniCPM-V 4.6 on X
Post on X