Average Ratings 0 Ratings
Average Ratings 0 Ratings
Description
K2 Horizon comprises a network of six open models, including the 375B-A23B, 36B-A4B, 32B, 7B, 3.7B, and 0.9B, all engineered to excel in various domains such as reasoning, mathematics, coding, agentic tasks, and overall capabilities. These models share a unified architecture, vocabulary, training techniques, interfaces, evaluation frameworks, and deployment tools, facilitating seamless transitions between sizes and dynamic workload management. The fleet's flagship, the 375B-A23B model, is particularly adept at handling intricate reasoning, software development, research projects, and long-term agentic functions, while the 32B and 36B-A4B models focus on delivering robust local deployment solutions. Notably, the 36B-A4B model features an innovative Mixture-of-Value Attention mechanism, which merges sparse attention with Mixture-of-Experts layers, allowing it to engage approximately 4 billion parameters per token and compete closely with the performance of the denser 32B model. This architecture not only enhances flexibility but also maximizes resource efficiency across a wide range of applications.
Description
Kimi K2 represents a cutting-edge series of open-source large language models utilizing a mixture-of-experts (MoE) architecture, with a staggering 1 trillion parameters in total and 32 billion activated parameters tailored for optimized task execution. Utilizing the Muon optimizer, it has been trained on a substantial dataset of over 15.5 trillion tokens, with its performance enhanced by MuonClip’s attention-logit clamping mechanism, resulting in remarkable capabilities in areas such as advanced knowledge comprehension, logical reasoning, mathematics, programming, and various agentic operations. Moonshot AI offers two distinct versions: Kimi-K2-Base, designed for research-level fine-tuning, and Kimi-K2-Instruct, which is pre-trained for immediate applications in chat and tool interactions, facilitating both customized development and seamless integration of agentic features. Comparative benchmarks indicate that Kimi K2 surpasses other leading open-source models and competes effectively with top proprietary systems, particularly excelling in coding and intricate task analysis. Furthermore, it boasts a generous context length of 128 K tokens, compatibility with tool-calling APIs, and support for industry-standard inference engines, making it a versatile option for various applications. The innovative design and features of Kimi K2 position it as a significant advancement in the field of artificial intelligence language processing.
API Access
Has API
No
API Access
Has API
Yes
Integrations
AiAssistWorks
No
Brokk
No
EaseMate AI
No
Kimi
No
NVIDIA TensorRT
No
Nebius Token Factory
No
Okara
No
OpenClaw
No
OpenCode
No
PenguinBot
No
Integrations
AiAssistWorks
Yes
Brokk
Yes
EaseMate AI
Yes
Kimi
Yes
NVIDIA TensorRT
Yes
Nebius Token Factory
Yes
Okara
Yes
OpenClaw
Yes
OpenCode
Yes
PenguinBot
Yes
Pricing Details
No price information available.
Free Trial
No
Free Version
No
Pricing Details
Free
Open source
Free Trial
No
Free Version
Yes
Deployment
Web-Based
Yes
On-Premises
No
iPhone App
No
iPad App
No
Android App
No
Windows
No
Mac
No
Linux
No
Chromebook
No
Deployment
Web-Based
Yes
On-Premises
Yes
iPhone App
No
iPad App
No
Android App
No
Windows
Yes
Mac
Yes
Linux
Yes
Chromebook
No
Customer Support
Business Hours
No
Live Rep (24/7)
No
Online Support
Yes
Customer Support
Business Hours
No
Live Rep (24/7)
No
Online Support
Yes
Types of Training
Training Docs
Yes
Webinars
No
Live Training (Online)
No
In Person
No
Types of Training
Training Docs
Yes
Webinars
No
Live Training (Online)
No
In Person
No
Vendor Details
Company Name
Institute of Foundation Models
Founded
2025
Country
United States
Website
ifm.ai/blog/k2/
Vendor Details
Company Name
Moonshot AI
Founded
2023
Country
China
Website
moonshotai.github.io/Kimi-K2/