← all repositories

zai-org/GLM-V

GLM-V is an open-source series of vision-language models (GLM-4.6V, GLM-4.5V, GLM-4.1V) for multimodal reasoning tasks including image understanding, video comprehension, and complex problem solving.

GLM-V
Not currently ranked — collecting fresh signals.
star history

The repository provides pre-trained vision-language models that process and reason over images and video alongside text. It includes model weights, training code, and inference scripts for the GLM-4.6V, GLM-4.5V, and GLM-4.1V model families. The models are trained using scalable reinforcement learning to enhance complex reasoning and multimodal understanding capabilities.

Frequently asked

What is zai-org/GLM-V?
GLM-V is an open-source series of vision-language models (GLM-4.6V, GLM-4.5V, GLM-4.1V) for multimodal reasoning tasks including image understanding, video comprehension, and complex problem solving.
Is GLM-V open source?
Yes — zai-org/GLM-V is open source, released under the Apache-2.0 license.
What language is GLM-V written in?
zai-org/GLM-V is primarily written in Python.
How popular is GLM-V?
zai-org/GLM-V has 2.4k stars on GitHub.
Where can I find GLM-V?
zai-org/GLM-V is on GitHub at https://github.com/zai-org/GLM-V.

heatdrop uses Google Analytics to see which pages get read — nothing else. Your call. How we handle data.