148 1 year ago

is a novel end-to-end trained large multimodal model that combines a vision encoder and Vicuna for general-purpose visual and language understanding.

vision
ollama run AgentricAi/AgentricAI_LLaVa

Models

View all →

Readme

No readme