GLM 4.6 Vision
Z.AI · text, image, video, document → text
Zhipu's latest visual reasoning model achieves state-of-the-art visual understanding accuracy at the same scale upon release. It natively supports tool invocation, can automatically complete tasks, supports ultra-long 128K context length, and allows flexible toggling of reasoning.
Input$0.137 /M
Output$0.411 /M
Cache read$0.0274 /M
GLM 4.6 Vision