Average Ratings 0 Ratings
Average Ratings 0 Ratings
Description
GLM-5.1 represents the latest advancement in Z.ai’s GLM series, crafted as a cutting-edge, agent-focused AI model tailored for coding, reasoning, and managing long-term workflows. This iteration builds upon the framework of GLM-5, which employs a Mixture-of-Experts (MoE) architecture to achieve high performance without incurring excessive inference expenses, aligning with a larger initiative towards open-weight models that are accessible to developers. A significant emphasis of GLM-5.1 is on fostering agentic behavior, allowing it to plan, execute, and refine multi-step tasks instead of merely reacting to isolated prompts. Its capabilities are specifically engineered to manage intricate workflows, such as debugging code, exploring repositories, and performing sequential operations while maintaining context over time. In comparison to its predecessors, GLM-5.1 enhances reliability during lengthy interactions, ensuring coherence throughout extended sessions and minimizing failures in multi-step reasoning processes. Overall, this model signifies a leap forward in AI development, particularly in its ability to support complex task management seamlessly.
Description
K2 Horizon comprises a network of six open models, including the 375B-A23B, 36B-A4B, 32B, 7B, 3.7B, and 0.9B, all engineered to excel in various domains such as reasoning, mathematics, coding, agentic tasks, and overall capabilities. These models share a unified architecture, vocabulary, training techniques, interfaces, evaluation frameworks, and deployment tools, facilitating seamless transitions between sizes and dynamic workload management. The fleet's flagship, the 375B-A23B model, is particularly adept at handling intricate reasoning, software development, research projects, and long-term agentic functions, while the 32B and 36B-A4B models focus on delivering robust local deployment solutions. Notably, the 36B-A4B model features an innovative Mixture-of-Value Attention mechanism, which merges sparse attention with Mixture-of-Experts layers, allowing it to engage approximately 4 billion parameters per token and compete closely with the performance of the denser 32B model. This architecture not only enhances flexibility but also maximizes resource efficiency across a wide range of applications.
API Access
Has API
Yes
API Access
Has API
No
Integrations
C
Yes
Cheaper Inference
Yes
Cherry Studio
Yes
Dessix
Yes
GLM Coding Plan
Yes
GLM-5-Turbo
Yes
HTML
Yes
Java
Yes
JavaScript
Yes
Kotlin
Yes
Integrations
C
No
Cheaper Inference
No
Cherry Studio
No
Dessix
No
GLM Coding Plan
No
GLM-5-Turbo
No
HTML
No
Java
No
JavaScript
No
Kotlin
No
Pricing Details
Free
Open source
Free Trial
No
Free Version
Yes
Pricing Details
No price information available.
Free Trial
No
Free Version
No
Deployment
Web-Based
Yes
On-Premises
Yes
iPhone App
No
iPad App
No
Android App
No
Windows
Yes
Mac
Yes
Linux
Yes
Chromebook
No
Deployment
Web-Based
Yes
On-Premises
No
iPhone App
No
iPad App
No
Android App
No
Windows
No
Mac
No
Linux
No
Chromebook
No
Customer Support
Business Hours
No
Live Rep (24/7)
No
Online Support
Yes
Customer Support
Business Hours
No
Live Rep (24/7)
No
Online Support
Yes
Types of Training
Training Docs
Yes
Webinars
No
Live Training (Online)
No
In Person
No
Types of Training
Training Docs
Yes
Webinars
No
Live Training (Online)
No
In Person
No
Vendor Details
Company Name
Z.ai
Founded
2023
Country
China
Website
z.ai/
Vendor Details
Company Name
Institute of Foundation Models
Founded
2025
Country
United States
Website
ifm.ai/blog/k2/