Average Ratings 0 Ratings
Average Ratings 0 Ratings
Description
Hy4 preview represents a cutting-edge open source Mixture-of-Experts flagship model tailored for a variety of real-world productivity tasks, including software engineering, office activities, game development, and scientific exploration. This model boasts a staggering total of 770 billion parameters, with 49 billion activated per token, and features an impressive 1 million-token context window, allowing it to efficiently manage large codebases, vast document collections, and complex multi-step processes. The architecture consists of 78 layers that integrate Gated DeepSeek Sparse Attention alongside IndexCache for reusing sparse indices across layers, while also employing identity Hyper-Connections to enhance the flow of information between layers. Additionally, a dedicated Multi-Token Prediction layer facilitates speculative decoding, further enhancing its capabilities. Hy4 preview is crafted to comprehend, plan, debug, and validate intricate engineering projects, while also achieving notable improvements in the quality of front-end visuals and interaction design, thereby making it an invaluable asset for professionals across various domains.
Description
Ling 2.6 represents an independently developed and open-source series of large language models created by Ant Group, utilizing a Mixture of Experts (MoE) architecture to enhance inference efficiency, long context modeling, training methodologies, and collaborative reasoning for AI agents. By employing this MoE architecture, Ling effectively directs each token to engage only the most pertinent expert subnetworks, significantly reducing the computational load while preserving the extensive capabilities of the model. This series makes strides in long-sequence modeling, exemplified by Ling-2.6-1T, which accommodates a native context window of up to 1 million tokens and offers a 256K context window through its official API; additionally, Ling-2.6-flash features a native 256K context window, enabling it to handle around 200,000 characters in lengthy inputs. These models are meticulously crafted to ensure dependable retrieval of long-range information without any discernible loss of quality, regardless of whether the data is located at the start, middle, or end of the context. This innovative approach to long-context processing sets a new benchmark for efficiency and reliability in language model performance.
API Access
Has API
Yes
API Access
Has API
Yes
Integrations
Claude Code
No
Hermes Agent
No
Kilo Code
No
OpenAI
No
OpenClaw
No
OpenRouter
No
Tencent Hy
Yes
Integrations
Claude Code
Yes
Hermes Agent
Yes
Kilo Code
Yes
OpenAI
Yes
OpenClaw
Yes
OpenRouter
Yes
Tencent Hy
No
Pricing Details
No price information available.
Free Trial
No
Free Version
No
Pricing Details
$0.0028 per 1M tokens
Free Trial
No
Free Version
No
Deployment
Web-Based
Yes
On-Premises
No
iPhone App
No
iPad App
No
Android App
No
Windows
No
Mac
No
Linux
No
Chromebook
No
Deployment
Web-Based
Yes
On-Premises
No
iPhone App
No
iPad App
No
Android App
No
Windows
No
Mac
No
Linux
No
Chromebook
No
Customer Support
Business Hours
No
Live Rep (24/7)
No
Online Support
Yes
Customer Support
Business Hours
No
Live Rep (24/7)
No
Online Support
Yes
Types of Training
Training Docs
Yes
Webinars
No
Live Training (Online)
No
In Person
No
Types of Training
Training Docs
Yes
Webinars
No
Live Training (Online)
No
In Person
No
Vendor Details
Company Name
Tencent
Founded
1998
Country
China
Website
hy.tencent.ai/research/hy4-preview
Vendor Details
Company Name
Ant Group
Founded
2014
Country
China
Website
developer.ant-ling.com/en/docs/models/ling/