Average Ratings 1 Rating
Average Ratings 0 Ratings
Description
MiMo-V2.6-Pro is an open-source, natively omnimodal AI model from Xiaomi MiMo designed for coding, agentic workflows, visual creation, research, and computer-based tasks. It is the highest-capability model in the MiMo-V2.6 series and was trained using a large-scale reinforcement learning process focused on verifiable, complex tasks. The model can handle software engineering, automation, tool use, visual reasoning, and long-running agent workflows across multiple environments. Its multimodal capabilities extend beyond traditional coding into 3D scene generation, Blender modeling, embodied simulation, frontend development, presentation design, video creation, and music composition. MiMo-V2.6-Pro can take text, images, video, and other visual inputs and coordinate multi-agent workflows to produce and iteratively refine complex outputs. Research applications demonstrated by Xiaomi include materials discovery, literature analysis, computational simulation, hypothesis generation, and formalizing mathematical proofs in Lean. Xiaomi has released the model alongside its technical report, reinforcement learning environments, and RL code so researchers can study and reproduce the training approach. MiMo-V2.6-Pro is also available in an UltraSpeed configuration that provides substantially faster output for latency-sensitive workflows. Users can access the model through MiMo Desktop, AI Studio, MiMo Code, the MiMo API Platform, OpenRouter, and Hugging Face.
Description
Pixtral Large is an expansive multimodal model featuring 124 billion parameters, crafted by Mistral AI and enhancing their previous Mistral Large 2 framework. This model combines a 123-billion-parameter multimodal decoder with a 1-billion-parameter vision encoder, allowing it to excel in the interpretation of various content types, including documents, charts, and natural images, all while retaining superior text comprehension abilities. With the capability to manage a context window of 128,000 tokens, Pixtral Large can efficiently analyze at least 30 high-resolution images at once. It has achieved remarkable results on benchmarks like MathVista, DocVQA, and VQAv2, outpacing competitors such as GPT-4o and Gemini-1.5 Pro. Available for research and educational purposes under the Mistral Research License, it also has a Mistral Commercial License for business applications. This versatility makes Pixtral Large a valuable tool for both academic research and commercial innovations.
API Access
Has API
Yes
API Access
Has API
Yes
Screenshots View All
No images available
Integrations
APIPark
No
Amazon Bedrock
No
AnythingLLM
No
Cline
Yes
Expanse
No
Fleak
No
Hyperbolic
No
Kilo Code
Yes
Langflow
No
Lunary
No
Integrations
APIPark
Yes
Amazon Bedrock
Yes
AnythingLLM
Yes
Cline
No
Expanse
Yes
Fleak
Yes
Hyperbolic
Yes
Kilo Code
No
Langflow
Yes
Lunary
Yes
Pricing Details
Free
$0.435 per 1 million tokens input
$0.87 per 1 million tokens output
$0.87 per 1 million tokens output
Free Trial
No
Free Version
Yes
Pricing Details
Free
Open source
Free Trial
No
Free Version
Yes
Deployment
Web-Based
Yes
On-Premises
No
iPhone App
No
iPad App
No
Android App
No
Windows
No
Mac
No
Linux
No
Chromebook
No
Deployment
Web-Based
Yes
On-Premises
Yes
iPhone App
No
iPad App
No
Android App
No
Windows
Yes
Mac
Yes
Linux
Yes
Chromebook
No
Customer Support
Business Hours
No
Live Rep (24/7)
No
Online Support
Yes
Customer Support
Business Hours
No
Live Rep (24/7)
No
Online Support
No
Types of Training
Training Docs
Yes
Webinars
No
Live Training (Online)
No
In Person
No
Types of Training
Training Docs
Yes
Webinars
No
Live Training (Online)
No
In Person
No
Vendor Details
Company Name
Xiaomi Technology
Founded
2010
Country
China
Website
mimo.xiaomi.com
Vendor Details
Company Name
Mistral AI
Founded
2023
Country
France
Website
mistral.ai/news/pixtral-large/