Zhipu AI
Profile Overview
Zhipu AI (智谱AI) is a Chinese artificial intelligence company founded in 2019 as a spin-off from Tsinghua University's esteemed Knowledge Engineering Group (KEG) at the Department of Computer Science. Headquartered in Beijing, Zhipu AI was co-founded by Professor Tang Jie and Dr. Zhang Peng, building on years of academic research in knowledge graph reasoning and large-scale language modeling. Zhipu AI is best known for developing the GLM (General Language Model) series and the ChatGLM family of conversational LLMs, which are among the most widely adopted Chinese-language AI models globally. ChatGLM-6B, released in open-source, became a landmark model in China for its ability to run on consumer-grade hardware while delivering strong Chinese and English language understanding. The company's flagship commercial model, GLM-4, consistently ranks at the top of Chinese AI benchmarks alongside GPT-4-class models, offering strong performance in long-context reasoning, multimodal understanding, and agentic task execution. Zhipu AI also launched CodeGeeX, an open-source code generation model, and provides enterprise solutions across finance, healthcare, and education. Zhipu AI has raised substantial funding from prominent investors including Alibaba, Tencent, and Legend Capital. In February 2024, the company raised $438 million in a Series D round, bringing its total funding to over $1.5 billion at a valuation exceeding $2.4 billion. Zhipu AI is regarded as one of China's "AI Tiger" companies and a leading candidate in the race to develop Chinese-language foundation models.
Last Financing Round
Flagship Offerings
Vetted AI Models
GLM Flash Latest
activeGLM Flash Latest is an advanced artificial intelligence model engineered by Zhipu AI. Tailored for complex multi-step reasoning, programming synthesis, and extended document comprehension, GLM Flash Latest represents a key architectural iteration in the Zhipu AI model family. First released in 2026-08-27, it serves enterprise developers, research teams, and autonomous system architects requiring strict instruction compliance. Featuring an input capacity of 1,310,720 tokens (approximately 1,748 words), GLM Flash Latest processes multi-file code repositories, lengthy technical reports, and complex prompts in a single inference call.
GLM 5.3 Flash
activeGLM 5.3 Flash is an advanced artificial intelligence model engineered by Zhipu AI. Tailored for complex multi-step reasoning, programming synthesis, and extended document comprehension, GLM 5.3 Flash represents a key architectural iteration in the Zhipu AI model family. First released in 2026-08-26, it serves enterprise developers, research teams, and autonomous system architects requiring strict instruction compliance. Featuring an input capacity of 1,310,720 tokens (approximately 1,748 words), GLM 5.3 Flash processes multi-file code repositories, lengthy technical reports, and complex prompts in a single inference call.
GLM Latest
activeGLM Latest is an advanced artificial intelligence model engineered by Zhipu AI. Tailored for complex multi-step reasoning, programming synthesis, and extended document comprehension, GLM Latest represents a key architectural iteration in the Zhipu AI model family. First released in 2026-08-19, it serves enterprise developers, research teams, and autonomous system architects requiring strict instruction compliance. Featuring an input capacity of 1,310,720 tokens (approximately 1,748 words), GLM Latest processes multi-file code repositories, lengthy technical reports, and complex prompts in a single inference call.
GLM 5.3
activeGLM 5.3 is an advanced artificial intelligence model engineered by Zhipu AI. Tailored for complex multi-step reasoning, programming synthesis, and extended document comprehension, GLM 5.3 represents a key architectural iteration in the Zhipu AI model family. First released in 2026-08-18, it serves enterprise developers, research teams, and autonomous system architects requiring strict instruction compliance. Featuring an input capacity of 1,310,720 tokens (approximately 1,748 words), GLM 5.3 processes multi-file code repositories, lengthy technical reports, and complex prompts in a single inference call.
GLM 5.2
activeGLM 5.2 is an advanced artificial intelligence model engineered by Zhipu AI. Tailored for complex multi-step reasoning, programming synthesis, and extended document comprehension, GLM 5.2 represents a key architectural iteration in the Zhipu AI model family. First released in 2026-06-16, it serves enterprise developers, research teams, and autonomous system architects requiring strict instruction compliance. Featuring an input capacity of 1,048,576 tokens (approximately 1,398 words), GLM 5.2 processes multi-file code repositories, lengthy technical reports, and complex prompts in a single inference call.
GLM 5.1
activeGLM 5.1 is an advanced artificial intelligence model engineered by Zhipu AI. Tailored for complex multi-step reasoning, programming synthesis, and extended document comprehension, GLM 5.1 represents a key architectural iteration in the Zhipu AI model family. First released in 2026-04-07, it serves enterprise developers, research teams, and autonomous system architects requiring strict instruction compliance. Featuring an input capacity of 204,800 tokens (approximately 273 words), GLM 5.1 processes multi-file code repositories, lengthy technical reports, and complex prompts in a single inference call.
GLM 5V Turbo
activeGLM 5V Turbo is an advanced artificial intelligence model engineered by Zhipu AI. Tailored for complex multi-step reasoning, programming synthesis, and extended document comprehension, GLM 5V Turbo represents a key architectural iteration in the Zhipu AI model family. First released in 2026-04-01, it serves enterprise developers, research teams, and autonomous system architects requiring strict instruction compliance. Featuring an input capacity of 202,752 tokens (approximately 270 words), GLM 5V Turbo processes multi-file code repositories, lengthy technical reports, and complex prompts in a single inference call.
GLM 5 Turbo
activeGLM 5 Turbo is an advanced artificial intelligence model engineered by Zhipu AI. Tailored for complex multi-step reasoning, programming synthesis, and extended document comprehension, GLM 5 Turbo represents a key architectural iteration in the Zhipu AI model family. First released in 2026-03-15, it serves enterprise developers, research teams, and autonomous system architects requiring strict instruction compliance. Featuring an input capacity of 202,752 tokens (approximately 270 words), GLM 5 Turbo processes multi-file code repositories, lengthy technical reports, and complex prompts in a single inference call.
GLM 5
activeGLM 5 is an advanced artificial intelligence model engineered by Zhipu AI. Tailored for complex multi-step reasoning, programming synthesis, and extended document comprehension, GLM 5 represents a key architectural iteration in the Zhipu AI model family. First released in 2026-02-11, it serves enterprise developers, research teams, and autonomous system architects requiring strict instruction compliance. Featuring an input capacity of 204,800 tokens (approximately 273 words), GLM 5 processes multi-file code repositories, lengthy technical reports, and complex prompts in a single inference call.
GLM 4.7 Flash
activeGLM 4.7 Flash is an advanced artificial intelligence model engineered by Zhipu AI. Tailored for complex multi-step reasoning, programming synthesis, and extended document comprehension, GLM 4.7 Flash represents a key architectural iteration in the Zhipu AI model family. First released in 2026-01-19, it serves enterprise developers, research teams, and autonomous system architects requiring strict instruction compliance. Featuring an input capacity of 200,000 tokens (approximately 267 words), GLM 4.7 Flash processes multi-file code repositories, lengthy technical reports, and complex prompts in a single inference call.
GLM 4.7
activeGLM 4.7 is an advanced artificial intelligence model engineered by Zhipu AI. Tailored for complex multi-step reasoning, programming synthesis, and extended document comprehension, GLM 4.7 represents a key architectural iteration in the Zhipu AI model family. First released in 2025-12-22, it serves enterprise developers, research teams, and autonomous system architects requiring strict instruction compliance. Featuring an input capacity of 204,800 tokens (approximately 273 words), GLM 4.7 processes multi-file code repositories, lengthy technical reports, and complex prompts in a single inference call.
GLM 4.6V
activeGLM 4.6V is an advanced artificial intelligence model engineered by Zhipu AI. Tailored for complex multi-step reasoning, programming synthesis, and extended document comprehension, GLM 4.6V represents a key architectural iteration in the Zhipu AI model family. First released in 2025-12-08, it serves enterprise developers, research teams, and autonomous system architects requiring strict instruction compliance. Featuring an input capacity of 131,072 tokens (approximately 175 words), GLM 4.6V processes multi-file code repositories, lengthy technical reports, and complex prompts in a single inference call.
GLM 4.6
activeGLM 4.6 is an advanced artificial intelligence model engineered by Zhipu AI. Tailored for complex multi-step reasoning, programming synthesis, and extended document comprehension, GLM 4.6 represents a key architectural iteration in the Zhipu AI model family. First released in 2025-09-30, it serves enterprise developers, research teams, and autonomous system architects requiring strict instruction compliance. Featuring an input capacity of 204,800 tokens (approximately 273 words), GLM 4.6 processes multi-file code repositories, lengthy technical reports, and complex prompts in a single inference call.
GLM 4.5V
activeGLM 4.5V is an advanced artificial intelligence model engineered by Zhipu AI. Tailored for complex multi-step reasoning, programming synthesis, and extended document comprehension, GLM 4.5V represents a key architectural iteration in the Zhipu AI model family. First released in 2025-08-11, it serves enterprise developers, research teams, and autonomous system architects requiring strict instruction compliance. Featuring an input capacity of 65,536 tokens (approximately 87 words), GLM 4.5V processes multi-file code repositories, lengthy technical reports, and complex prompts in a single inference call.
GLM 4.5 Air
activeGLM 4.5 Air is an advanced artificial intelligence model engineered by Zhipu AI. Tailored for complex multi-step reasoning, programming synthesis, and extended document comprehension, GLM 4.5 Air represents a key architectural iteration in the Zhipu AI model family. First released in 2025-07-25, it serves enterprise developers, research teams, and autonomous system architects requiring strict instruction compliance. Featuring an input capacity of 131,072 tokens (approximately 175 words), GLM 4.5 Air processes multi-file code repositories, lengthy technical reports, and complex prompts in a single inference call.
GLM 4.5
activeGLM 4.5 is an advanced artificial intelligence model engineered by Zhipu AI. Tailored for complex multi-step reasoning, programming synthesis, and extended document comprehension, GLM 4.5 represents a key architectural iteration in the Zhipu AI model family. First released in 2025-07-25, it serves enterprise developers, research teams, and autonomous system architects requiring strict instruction compliance. Featuring an input capacity of 131,072 tokens (approximately 175 words), GLM 4.5 processes multi-file code repositories, lengthy technical reports, and complex prompts in a single inference call.