GLM 4.6 Vision
Z.AI · text, image, video → text
Zhipu's latest visual reasoning model achieves state-of-the-art visual understanding accuracy at the same scale upon release. It natively supports tool invocation, can automatically complete tasks, supports ultra-long 128K context length, and allows flexible toggling of reasoning.
Input$0.137 /M
Output$0.411 /M
Cache read$0.0274 /M

GLM 4.6 Vision