-
minicpm-o2.6
A GPT-4o Level MLLM for Vision, Speech and Multimodal Live Streaming on Your Phone
vision 8b28.1K Pulls 13 Tags Updated 1 year ago
-
minicpm-v4.5
A GPT-4o Level MLLM for Single Image, Multi Image and High-FPS Video Understanding on Your Phone
vision 8b22.9K Pulls 11 Tags Updated 3 months ago
-
minicpm-v4.6
A Pocket-Sized MLLM for Ultra-Efficient Image and Video Understanding on Your Phone
vision 1b13.6K Pulls 13 Tags Updated 3 months ago
-
minicpm5
highly efficient large language models (LLMs) designed explicitly for end-side devices
10K Pulls 4 Tags Updated 3 months ago
-
minicpm-o4.5
A Gemini 2.5 Flash Level MLLM for Vision, Speech, and Full-Duplex Mulitmodal Live Streaming on Your Phone
vision 8b8,873 Pulls 12 Tags Updated 7 months ago
-
minicpm5-2b
SOTA on-device LLMs, small yet powerful
thinking 2b3,328 Pulls 13 Tags Updated 6 days ago
-
minicpm-v2.6
A GPT-4V Level MLLM for Single Image, Multi Image and Video on Your Phone
vision 8b3,081 Pulls 12 Tags Updated 1 year ago
-
minicpm-v4
A GPT-4V Level MLLM for Single Image, Multi Image and Video on Your Phone
vision 4b2,443 Pulls 12 Tags Updated 1 year ago
-
minicpm4.1
highly efficient large language models (LLMs) designed explicitly for end-side devices
1,545 Pulls 1 Tag Updated 1 year ago
-
minicpm-v2.5
A GPT-4V Level Multimodal LLM on Your Phone
vision 8b507 Pulls 13 Tags Updated 1 year ago