A series of multimodal LLMs (MLLMs) designed for vision-language understanding.
vision
8b
42.9K Pulls Updated 6 days ago
f02dd72bb242 · 59B
{
"stop": [
"<|im_start|>",
"<|im_end|>"
]
}