What Integrates with Ollama?
Find out what Ollama integrations exist in 2026. Learn what software and services currently integrate with Ollama, and sort them by reviews, cost, features, and more. Below is a list of products that Ollama currently integrates with:
-
1
Devika
Devika
FreeDevika is an innovative open-source AI software engineer that interprets high-level commands, dissects them into actionable steps, gathers pertinent information, and writes code to achieve specified goals. By leveraging advanced language models, reasoning techniques, and browsing functionalities, Devika effectively aids in software development, handling intricate coding challenges with little human oversight. The platform is compatible with various programming languages and boasts essential features such as sophisticated AI planning, contextual keyword identification, and real-time agent monitoring. With the intention of becoming a formidable competitor to proprietary AI solutions, Devika presents a bold, open-source alternative for developers seeking versatile support in their projects. Ultimately, it seeks to empower programmers by streamlining the coding process and enhancing productivity. -
2
E2B
E2B
FreeE2B is an open-source runtime that provides a secure environment for executing AI-generated code within isolated cloud sandboxes. This platform allows developers to enhance their AI applications and agents with code interpretation features, enabling the safe execution of dynamic code snippets in a regulated setting. Supporting a variety of programming languages like Python and JavaScript, E2B offers software development kits (SDKs) for easy integration into existing projects. It employs Firecracker microVMs to guarantee strong security and isolation during code execution. Developers have the flexibility to implement E2B on their own infrastructure or take advantage of the available cloud service. The platform is crafted to be agnostic to large language models, ensuring compatibility with numerous options, including OpenAI, Llama, Anthropic, and Mistral. Among its key features are quick sandbox initialization, customizable execution environments, and the capability to manage long-running sessions lasting up to 24 hours. With E2B, developers can confidently run AI-generated code while maintaining high standards of security and efficiency. -
3
LiteLLM
LiteLLM
FreeLiteLLM serves as a comprehensive platform that simplifies engagement with more than 100 Large Language Models (LLMs) via a single, cohesive interface. It includes both a Proxy Server (LLM Gateway) and a Python SDK, which allow developers to effectively incorporate a variety of LLMs into their applications without hassle. The Proxy Server provides a centralized approach to management, enabling load balancing, monitoring costs across different projects, and ensuring that input/output formats align with OpenAI standards. Supporting a wide range of providers, this system enhances operational oversight by creating distinct call IDs for each request, which is essential for accurate tracking and logging within various systems. Additionally, developers can utilize pre-configured callbacks to log information with different tools, further enhancing functionality. For enterprise clients, LiteLLM presents a suite of sophisticated features, including Single Sign-On (SSO), comprehensive user management, and dedicated support channels such as Discord and Slack, ensuring that businesses have the resources they need to thrive. This holistic approach not only improves efficiency but also fosters a collaborative environment where innovation can flourish. -
4
Gemma 3
Google
FreeGemma 3, launched by Google, represents a cutting-edge AI model constructed upon the Gemini 2.0 framework, aimed at delivering superior efficiency and adaptability. This innovative model can operate seamlessly on a single GPU or TPU, which opens up opportunities for a diverse group of developers and researchers. Focusing on enhancing natural language comprehension, generation, and other AI-related functions, Gemma 3 is designed to elevate the capabilities of AI systems. With its scalable and robust features, Gemma 3 aspires to propel the evolution of AI applications in numerous sectors and scenarios, potentially transforming the landscape of technology as we know it. -
5
MacWhisper
MacWhisper
€59 one-time paymentMacWhisper is a Mac transcription and dictation app that helps users transcribe audio, video, meetings, podcasts, lectures, interviews, subtitles, voice memos, and private files. The app supports drag-and-drop transcription for common media formats and can record meetings from Zoom, Teams, Webex, Skype, Chime, Discord, and other online meeting tools. MacWhisper can also capture and transcribe audio from any app on a Mac, making it useful for videos, calls, recordings, and media workflows. The platform is built with privacy in mind, offering local AI models and offline processing for sensitive content. Users can generate accurate transcripts, recognize speakers, remove filler words, translate text, search transcripts, edit content, and export files in formats such as subtitles, text, Markdown, PDF, HTML, and DOCX. Batch transcription helps professionals process multiple files at once. MacWhisper Pro adds AI services, custom prompts, cloud and local model options, app-specific dictation prompts, automatic meeting detection, watched folders, workflow uploads, and CLI control. The app can connect to AI providers such as OpenAI, Anthropic, xAI, Google Gemini, DeepSeek, Azure, OpenRouter, Ollama, LM Studio, Deepgram, ElevenLabs, and others. By combining transcription, meeting recording, dictation, privacy-focused local processing, AI summaries, exports, integrations, and workflow automation, MacWhisper helps users turn spoken content into useful text. -
6
EXAONE Deep
LG
FreeEXAONE Deep represents a collection of advanced language models that are enhanced for reasoning, created by LG AI Research, and come in sizes of 2.4 billion, 7.8 billion, and 32 billion parameters. These models excel in a variety of reasoning challenges, particularly in areas such as mathematics and coding assessments. Significantly, the EXAONE Deep 2.4B model outshines other models of its size, while the 7.8B variant outperforms both open-weight models of similar dimensions and the proprietary reasoning model known as OpenAI o1-mini. Furthermore, the EXAONE Deep 32B model competes effectively with top-tier open-weight models in the field. The accompanying repository offers extensive documentation that includes performance assessments, quick-start guides for leveraging EXAONE Deep models with the Transformers library, detailed explanations of quantized EXAONE Deep weights formatted in AWQ and GGUF, as well as guidance on how to run these models locally through platforms like llama.cpp and Ollama. Additionally, this resource serves to enhance user understanding and accessibility to the capabilities of EXAONE Deep models. -
7
NeoBase
NeoBase
FreeNeoBase serves as an intelligent assistant for databases, allowing users to perform queries, conduct analyses, and oversee database management through natural language interaction. It is compatible with various databases, enabling users to connect and communicate with them via a chat interface, which enhances the efficiency of transaction management and performance tuning. Being self-hosted and open-source, NeoBase grants users full control over their data while ensuring privacy. Its design embodies a sleek Neo Brutalism aesthetic, facilitating intuitive and effective database visualization. With NeoBase, users can convert natural language into optimized queries, thereby streamlining the execution of intricate database tasks. Additionally, it takes care of database schema management while providing users the autonomy to adjust it as needed. Users can execute queries, revert changes when necessary, and easily visualize extensive datasets. Moreover, NeoBase offers AI-driven recommendations to enhance database performance, making database management a more manageable and efficient process overall. -
8
Mongo Pilot
Mongo Pilot
$49MongoPilot transforms MongoDB management by offering a user-friendly GUI combined with powerful AI capabilities. Its visual query builder lets you easily construct queries through a drag-and-drop interface, while the local AI assistant helps you generate precise MongoDB queries through plain English commands. The platform’s local-first design means your data stays private and secure on your machine, and with MongoPilot’s smart automation, you can streamline your workflow, manage aggregation pipelines, and enjoy efficient, no-syntax-required querying. -
9
CodeNext
CodeNext
$15 per monthCodeNext.ai is an innovative AI-driven coding assistant tailored for Xcode developers, featuring advanced context-aware code completion alongside interactive chat capabilities. It is compatible with numerous top-tier AI models, such as OpenAI, Azure OpenAI, Google AI, Mistral, Anthropic, Deepseek, Ollama, and others, allowing developers the convenience to select and switch models according to their preferences. The tool offers smart, instant code suggestions as you type, significantly boosting productivity and coding effectiveness. Additionally, its chat functionality empowers developers to communicate in natural language for tasks like writing code, debugging, refactoring, and executing various coding operations within or outside the codebase. CodeNext.ai also incorporates custom chat plugins, facilitating the execution of terminal commands and shortcuts right within the chat interface, thereby optimizing the overall development process. Ultimately, this sophisticated assistant not only simplifies coding tasks but also enhances collaboration and streamlines the workflow for developers. -
10
Devstral
Mistral AI
$0.1 per million input tokensDevstral is a collaborative effort between Mistral AI and All Hands AI, resulting in an open-source large language model specifically tailored for software engineering. This model demonstrates remarkable proficiency in navigating intricate codebases, managing edits across numerous files, and addressing practical problems, achieving a notable score of 46.8% on the SWE-Bench Verified benchmark, which is superior to all other open-source models. Based on Mistral-Small-3.1, Devstral boasts an extensive context window supporting up to 128,000 tokens. It is designed for optimal performance on high-performance hardware setups, such as Macs equipped with 32GB of RAM or Nvidia RTX 4090 GPUs, and supports various inference frameworks including vLLM, Transformers, and Ollama. Released under the Apache 2.0 license, Devstral is freely accessible on platforms like Hugging Face, Ollama, Kaggle, Unsloth, and LM Studio, allowing developers to integrate its capabilities into their projects seamlessly. This model not only enhances productivity for software engineers but also serves as a valuable resource for anyone working with code. -
11
Thunder Compute
Thunder Compute
$0.35 per hourThunder Compute delivers cheap cloud GPUs for companies, researchers, and developers running demanding AI and machine learning workloads. The platform gives users fast access to H100, A100, and RTX A6000 GPUs for LLM training, inference, fine-tuning, image generation, ComfyUI workflows, PyTorch jobs, CUDA applications, deep learning pipelines, model serving, and other GPU-intensive compute tasks. Thunder Compute is designed for teams that want affordable GPU cloud infrastructure with a strong developer experience, clear pricing, and minimal operational friction. Instead of dealing with the cost and complexity of legacy cloud vendors, users can deploy on-demand GPU instances with persistent storage, rapid provisioning, straightforward management, and scalable compute capacity. Thunder Compute is a strong fit for startups building AI products, engineering teams that need cloud GPUs for inference, and organizations looking for GPU hosting that is both economical and reliable. If you are searching for cheap H100s, A100 cloud instances, affordable GPUs for AI, or a RunPod alternative with transparent pricing and a simple interface, Thunder Compute provides a modern option for high-performance cloud GPU rental and AI infrastructure. Thunder Compute supports teams building and deploying modern AI applications that need dependable access to cheap cloud GPUs for both experimentation and production. From prototype training runs to large-scale inference and batch processing, the platform is designed to reduce infrastructure friction and accelerate iteration. For users comparing GPU cloud providers, Thunder Compute stands out with affordable pricing, fast access to top-tier GPUs, and a developer-friendly experience built around real AI workflows. -
12
NativeMind
NativeMind
FreeNativeMind serves as a completely open-source AI assistant that operates directly within your browser through Ollama integration, maintaining total privacy by refraining from sending any data to external servers. All processes, including model inference and prompt handling, take place locally, which eliminates concerns about syncing, logging, or data leaks. Users can effortlessly transition between various powerful open models like DeepSeek, Qwen, Llama, Gemma, and Mistral, requiring no extra configurations, while taking advantage of native browser capabilities to enhance their workflows. Additionally, NativeMind provides efficient webpage summarization; it maintains ongoing, context-aware conversations across multiple tabs; offers local web searches that can answer questions straight from the page; and delivers immersive translations that keep the original format intact. Designed with an emphasis on both efficiency and security, this extension is fully auditable and supported by the community, ensuring enterprise-level performance suitable for real-world applications without the risk of vendor lock-in or obscure telemetry. Moreover, the user-friendly interface and seamless integration make it an appealing choice for those seeking a reliable AI assistant that prioritizes their privacy. -
13
Dyad
Dyad
FreeDyad is an open-source AI app builder that allows users to transform their ideas into fully functional applications directly on their local machines without the need for any coding, simply by conversing with an AI. This powerful tool supports the creation of unlimited applications with features such as real-time previews, instant undo capabilities, and seamless workflows that enhance productivity. With robust integration of Supabase, users can develop both the user interface and backend logic in a unified environment, while its flexible, model-agnostic framework allows connections to any AI system—be it cloud-based models like Gemini 2.5 Pro and GPT-4, or local solutions such as Ollama—ensuring no vendor lock-in. Moreover, all source code is kept securely on your device and works effortlessly with your favorite IDE, providing a personalized development environment. A natural-language API facilitates sophisticated data queries and updates, enabling automation of tasks while staying within the chat interface. By operating entirely on local machines, Dyad maximizes user privacy, reduces latency, and offers a smooth developer experience, free from the uncertainties associated with cloud services. Thus, users can enjoy a powerful combination of functionality and privacy in their app development journey. -
14
DeepSeek V3.1
DeepSeek
FreeDeepSeek V3.1 stands as a revolutionary open-weight large language model, boasting an impressive 685-billion parameters and an expansive 128,000-token context window, which allows it to analyze extensive documents akin to 400-page books in a single invocation. This model offers integrated functionalities for chatting, reasoning, and code creation, all within a cohesive hybrid architecture that harmonizes these diverse capabilities. Furthermore, V3.1 accommodates multiple tensor formats, granting developers the versatility to enhance performance across various hardware setups. Preliminary benchmark evaluations reveal strong results, including a remarkable 71.6% on the Aider coding benchmark, positioning it competitively with or even superior to systems such as Claude Opus 4, while achieving this at a significantly reduced cost. Released under an open-source license on Hugging Face with little publicity, DeepSeek V3.1 is set to revolutionize access to advanced AI technologies, potentially disrupting the landscape dominated by conventional proprietary models. Its innovative features and cost-effectiveness may attract a wide range of developers eager to leverage cutting-edge AI in their projects. -
15
Broxi AI
Broxi AI
$25 per monthBroxi AI is an innovative no-code platform that empowers users to transform a basic text description into a fully operational AI agent in just minutes, utilizing intuitive visual drag-and-drop functionalities that eliminate the need for any technical expertise. Its unique Broxi Autopilot feature allows users to input natural language commands, like “create an agent to handle FAQs from our PDF handbook,” and seamlessly specify various input types such as PDFs, chat interfaces, or websites, along with diverse output options like emails, messages, or API interactions. With a single click, Broxi efficiently builds, tests within an interactive sandbox, and enables immediate deployment of your AI agent through various channels, including API, web widgets, Slack integration, or embedded applications. Additionally, it boasts compatibility with numerous tools and systems, provides real-time monitoring and centralized management capabilities, and upholds enterprise-level security standards, ensuring that even non-technical teams can easily automate tasks related to customer support, internal processes, sales interactions, content creation, and data extraction without the necessity of coding. This makes Broxi a powerful ally for organizations aiming to enhance their efficiency and service delivery through AI. -
16
Novelcrafter
Novelcrafter
$4.64 per monthNovelcrafter is an innovative writing platform powered by AI, designed to assist authors at every stage of their storytelling journey, encompassing everything from brainstorming ideas and developing characters to drafting, reviewing, and finalizing their written works. The platform features a specialized “Codex” wiki that allows writers to organize essential elements such as characters, settings, lore, and world-building details, promoting consistency and ease of access. Additionally, it provides various structured planning modes that include acts, chapters, and scenes, enabling authors to transition effortlessly between the planning phase and the writing interface. Authors have the flexibility to utilize customizable AI tools, allowing them to link their own API keys (such as OpenAI, Claude, or local LLMs) and create specific prompts, or they can choose to write manually without any AI assistance. Furthermore, Novelcrafter boasts a distraction-free writing mode, keeps a revision history, supports the import and export of documents in formats like Word, Markdown, and HTML, and offers mobile compatibility for writers who need to jot down ideas while on the move. This platform seeks to empower writers by providing a comprehensive suite of tools tailored to enhance creativity and streamline the writing process. -
17
BrowserOS
BrowserOS
FreeBrowserOS is an open-source web browser that is agent-enabled and built on a fork of Chromium, integrating AI agents seamlessly into the online experience to facilitate task automation, navigation, and interaction with web applications using natural language commands. Users can log into websites as they normally would, and by issuing simple instructions such as “extract the quarterly results from this webpage and update a spreadsheet,” BrowserOS creates and executes a local, repeatable agent that takes care of clicks, form submissions, and other navigational tasks on their behalf. It comes equipped with a split-view feature that provides access to prominent large language models like ChatGPT, Claude, or Gemini, while also allowing for local model execution through platforms such as Ollama, ensuring it works harmoniously with existing Chrome extensions, bookmarks, and passwords. The browser enhances productivity by offering semantic search capabilities for browsing history and bookmarks, highlighting tools, and the option to set up MCP (Model-Context-Protocol) servers specifically for applications like Gmail, Calendar, Docs, and Notion, transforming it into a comprehensive productivity tool. Additionally, its user-friendly interface encourages a smooth transition for those accustomed to traditional browsing, as it simplifies complex tasks with the power of AI-driven automation. -
18
BotDojo
BotDojo
$89 per monthBotDojo serves as a robust AI enablement platform tailored for enterprises, allowing companies to create, implement, oversee, and expand intelligent agents across various communication channels like chat, voice, email, and web, all through an intuitive low-code visual workflow designer that seamlessly integrates with existing enterprise data systems. It boasts a library of over 100 pre-built templates aimed at streamlining typical applications, including support automation, knowledge retrieval, sales analytics, and internal operations, while also facilitating branching logic, memory capabilities, and the orchestration of tools such as code, RPA, and web browsing. In addition, BotDojo establishes connections with essential business tools like CRMs, ticketing platforms, and databases to enhance its functionality. The platform further fosters continuous improvement and learning for agents through human feedback loops, enabling employees to mentor agents by providing feedback, embedding corrections into agent memory and responses, and assessing performance using comprehensive observability metrics, including deflection rates, first-contact resolution, and cost per interaction. Ultimately, BotDojo not only optimizes operational efficiency but also ensures that intelligent agents evolve and adapt to meet organizational needs effectively. -
19
Ekinox
Ekinox
$30 per monthEkinox serves as a visual AI automation platform that allows users to create, implement, and oversee AI-driven workflows without the need for coding; its user-friendly drag-and-drop interface facilitates the design of intelligent agents that can link to over 100 pre-existing integrations, triggering actions across numerous productivity, data, and communication applications. The platform is designed for real-time processing and encourages collaboration by offering team workspaces, version control, and immediate deployment capabilities. In addition, it boasts enterprise-level security that adheres to SOC 2 standards, features bank-level encryption, supports custom API connectors, and includes sophisticated access controls. Users benefit from the ability to monitor their workflows through comprehensive analytics dashboards, enabling them to assess costs and performance across various models and integrations while utilizing predictive auto-scaling and log retention for enhanced functionality. With setup times cut down to mere minutes, Ekinox optimizes processes ranging from straightforward task automation to more complex workflows, making it an invaluable tool. This efficiency not only improves productivity but also enhances the overall user experience. -
20
schnell.digital AI Kit
schnell.digital GmbH
160 EUR/month schnell.digital AI Kit is a no-code AI automation and workflow platform that lets teams describe business processes in natural language and run them as autonomous agents. Instead of stitching together prompts, scripts, and SaaS tools, users build workflows in a visual story editor, connect them to company knowledge via built-in RAG, and let AI Kit execute them across existing systems. The platform is model-agnostic and BYOK: connect OpenAI, Anthropic, or Mistral via your own API keys, or run fully local with open-source models for sensitive workloads. RAG indexing supports common document formats and integrates with Microsoft 365, Google Workspace, and custom APIs. Workflows can chain LLM calls, retrieval, tool use, conditional logic, and human-in-the-loop approvals. Deployment is flexible: managed EU cloud (hosted in Germany) or full on-premise installation behind your firewall. On-premise tiers ship with unlimited storage, audit logging, and role-based access control. GDPR compliance is built in, with a DPA included by default. A metrics module tracks runs, latency, token costs, and outcomes per workflow, making ROI measurable and ops auditable. Tiered licensing scales from single-team Cloud Starter to multi-workspace Inhouse Enterprise, with implementation support from schnell.digital or certified partners — typical pilot rollout in 4–6 weeks. Built for mid-market companies that want measurable AI automation without vendor lock-in or a dedicated AI team. -
21
Scraib
Scraib
$3.99 per monthScraib.app is a macOS writing assistant powered by AI that resides in the menu bar, allowing users to select text from any application and improve it by pressing Control + R, which enhances grammar, clarity, and style. Users have the flexibility to set custom rules to align with their preferred tone, and unlike other writing software that requires switching between applications, Scraib seamlessly integrates with various platforms, including Slack, Outlook, Pages, Word, Chrome, and Figma. It prioritizes user privacy by offering options to work with different AI providers like ChatGPT, Claude, and others, while also allowing for local operation with supported models, ensuring that sensitive data remains secure. Designed for efficiency, it minimizes workflow interruptions, enabling users to refine their text without leaving their current application, making it an ideal tool for enhancing written communication on the fly. Additionally, Scraib's intuitive shortcut-based system enhances productivity, allowing for quick adjustments and refinements directly where the text exists. -
22
DeepSeek-V3.2
DeepSeek
FreeDeepSeek-V3.2 is a highly optimized large language model engineered to balance top-tier reasoning performance with significant computational efficiency. It builds on DeepSeek's innovations by introducing DeepSeek Sparse Attention (DSA), a custom attention algorithm that reduces complexity and excels in long-context environments. The model is trained using a sophisticated reinforcement learning approach that scales post-training compute, enabling it to perform on par with GPT-5 and match the reasoning skill of Gemini-3.0-Pro. Its Speciale variant overachieves in demanding reasoning benchmarks and does not include tool-calling capabilities, making it ideal for deep problem-solving tasks. DeepSeek-V3.2 is also trained using an agentic synthesis pipeline that creates high-quality, multi-step interactive data to improve decision-making, compliance, and tool-integration skills. It introduces a new chat template design featuring explicit thinking sections, improved tool-calling syntax, and a dedicated developer role used strictly for search-agent workflows. Users can encode messages using provided Python utilities that convert OpenAI-style chat messages into the expected DeepSeek format. Fully open-source under the MIT license, DeepSeek-V3.2 is a flexible, cutting-edge model for researchers, developers, and enterprise AI teams. -
23
MiniMax-M2.1
MiniMax
FreeMiniMax-M2.1 is a state-of-the-art open-source AI model built specifically for agent-based development and real-world automation. It focuses on delivering strong performance in coding, tool calling, and long-term task execution. Unlike closed models, MiniMax-M2.1 is fully transparent and can be deployed locally or integrated through APIs. The model excels in multilingual software engineering tasks and complex workflow automation. It demonstrates strong generalization across different agent frameworks and development environments. MiniMax-M2.1 supports advanced use cases such as autonomous coding, application building, and office task automation. Benchmarks show significant improvements over previous MiniMax versions. The model balances high reasoning ability with stability and control. Developers can fine-tune or extend it for specialized agent workflows. MiniMax-M2.1 empowers teams to build reliable AI agents without vendor lock-in. -
24
MiniMax M2.5
MiniMax
FreeMiniMax M2.5 is a next-generation foundation model built to power complex, economically valuable tasks with speed and cost efficiency. Trained using large-scale reinforcement learning across hundreds of thousands of real-world task environments, it excels in coding, tool use, search, and professional office workflows. In programming benchmarks such as SWE-Bench Verified and Multi-SWE-Bench, M2.5 reaches state-of-the-art levels while demonstrating improved multilingual coding performance. The model exhibits architect-level reasoning, planning system structure and feature decomposition before writing code. With throughput speeds of up to 100 tokens per second, it completes complex evaluations significantly faster than earlier versions. Reinforcement learning optimizations enable more precise search rounds and fewer reasoning steps, improving overall efficiency. M2.5 is available in two variants—standard and Lightning—offering identical capabilities with different speed configurations. Pricing is designed to be dramatically lower than competing frontier models, reducing cost barriers for large-scale agent deployment. Integrated into MiniMax Agent, the model supports advanced office skills including Word formatting, Excel financial modeling, and PowerPoint editing. By combining high performance, efficiency, and affordability, MiniMax M2.5 aims to make agent-powered productivity accessible at scale. -
25
Interpreter
Interpreter
FreeInterpreter is an innovative desktop AI solution that enables users to collaborate with smart assistants for tasks such as document editing, PDF form completion, and spreadsheet management all within a cohesive AI-driven platform. It accommodates both interactive and non-interactive PDF forms, allowing users to efficiently populate and process documents in real-time, eliminating the need for manual data input. Featuring a comprehensive AI-powered spreadsheet interface, it supports advanced functions like pivot tables, charts, and complex data manipulation, thereby serving as a contemporary substitute for conventional Excel methods. Additionally, Interpreter includes an integrated Word editor equipped with features like tracked changes, diverse formatting options, and image embedding, which facilitates document creation and modification with AI support directly within the software. Users have the flexibility to log in using OpenAI, utilize their own API keys, or operate the system offline through Ollama for local model execution, ensuring adaptability in deploying AI functionalities. This combination of features positions Interpreter as a versatile tool that enhances productivity while simplifying various administrative tasks. -
26
Qwen3.5
Alibaba
FreeQwen3.5 represents a major advancement in open-weight multimodal AI models, engineered to function as a native vision-language agent system. Its flagship model, Qwen3.5-397B-A17B, leverages a hybrid architecture that fuses Gated DeltaNet linear attention with a high-sparsity mixture-of-experts framework, allowing only 17 billion parameters to activate during inference for improved speed and cost efficiency. Despite its sparse activation, the full 397-billion-parameter model achieves competitive performance across reasoning, coding, multilingual benchmarks, and complex agent evaluations. The hosted Qwen3.5-Plus version supports a one-million-token context window and includes built-in tool use for search, code interpretation, and adaptive reasoning. The model significantly expands multilingual coverage to 201 languages and dialects while improving encoding efficiency with a larger vocabulary. Native multimodal training enables strong performance in image understanding, video processing, document analysis, and spatial reasoning tasks. Its infrastructure includes FP8 precision pipelines and heterogeneous parallelism to boost throughput and reduce memory consumption. Reinforcement learning at scale enhances multi-step planning and general agent behavior across text and multimodal environments. Overall, Qwen3.5 positions itself as a high-efficiency foundation for autonomous digital agents capable of reasoning, searching, coding, and interacting with complex environments. -
27
Agent Zero
Agent Zero
$2.65 per monthAgent Zero is an innovative open source framework for AI agents that enables the development of autonomous assistants capable of executing intricate tasks through direct interaction with computer systems. This platform offers a unique setting where AI agents can access real system functions, empowering them to run commands, write and execute code, navigate the internet, analyze data, and oversee workflows as part of comprehensive automation solutions. Unlike a standard chat interface, Agent Zero operates within its isolated virtual environment, enabling it to engage with the operating system, install necessary tools, run scripts, and manage tasks across various components seamlessly. The framework prioritizes transparency and developer control, allowing users to monitor, adjust, and personalize agent behavior, tool accessibility, and information processing methods. With a modular architecture, Agent Zero facilitates the dynamic creation and utilization of tools, all while maintaining a consistent memory for enhanced performance. This makes it an ideal choice for developers aiming to build highly customizable and efficient AI-driven workflows. -
28
GLM-5-Turbo
Z.ai
FreeGLM-5-Turbo represents a rapid iteration of Z.ai’s GLM-5 model, engineered to offer both efficient and stable performance specifically tailored for agent-driven scenarios, all while preserving robust reasoning and programming abilities. This model is fine-tuned to handle high-throughput demands, especially in complex long-chain agent tasks that necessitate a series of sequential steps, tools, and decisions executed reliably and with minimal latency. With its support for sophisticated agentic workflows, GLM-5-Turbo enhances multi-step planning, tool utilization, and task execution, delivering superior responsiveness compared to larger flagship models in the lineup. Drawing from the foundational strengths of the GLM-5 family, it maintains strong capabilities in reasoning, coding, and processing extensive contexts, but prioritizes the optimization of essential aspects like speed, efficiency, and stability within production settings. Furthermore, it is crafted to seamlessly integrate with agent frameworks such as OpenClaw, allowing it to proficiently coordinate actions, manage inputs, and carry out tasks effectively. This ensures that users benefit from a responsive and reliable tool that can adapt to various operational demands and complexities. -
29
DockClaw
DockClaw
$19.99 per monthDockClaw serves as a managed hosting solution for OpenClaw, facilitating the rapid deployment and operation of autonomous AI agents in mere seconds, all without the complexities of server management, Docker, or DevOps configurations. This platform empowers users to effortlessly launch AI-driven agents capable of integrating with various messaging services like Telegram and other communication avenues, enabling them to function continuously for automating workflows, interacting with users, and performing various tasks. With one-click deployment options available on dedicated virtual machines or isolated containers, DockClaw guarantees 24/7 uptime, persistent storage, and health monitoring, ensuring that agents stay consistently operational and reliable. Users benefit from the flexibility of selecting from a range of AI models, such as Claude, GPT, Gemini, Llama, and other systems compatible with OpenAI, with the ability to switch models easily without any vendor lock-in. Furthermore, DockClaw incorporates native configuration tools that allow for the fine-tuning of agent behavior, memory management, and system prompts, while also ensuring secure API key management through encrypted environments and a zero-knowledge architecture. This comprehensive approach not only enhances user experience but also fosters a versatile environment for AI development and deployment. -
30
MiniMax M2.7
MiniMax
FreeMiniMax M2.7 is a powerful AI model built to drive real-world productivity across coding, search, and office-based workflows. It is trained using reinforcement learning across a wide range of real-world environments, enabling it to execute complex, multi-step tasks with precision and efficiency. The model demonstrates strong problem-solving capabilities by breaking down challenges into structured steps before generating solutions across multiple programming languages. It delivers high-speed performance with rapid token output, ensuring faster completion of demanding tasks. With optimized reasoning, it reduces token usage and execution time, making it more efficient than previous models. M2.7 also achieves state-of-the-art results in software engineering benchmarks, significantly improving response times for technical issues. Its advanced agentic capabilities allow it to work seamlessly with tools and support complex workflows with high skill accuracy. The model is designed to handle professional tasks, including multi-turn interactions and high-quality document editing. It also provides strong support for office productivity, enabling efficient handling of structured data and business tasks. With competitive pricing, it delivers high performance while remaining cost-effective. Overall, it combines speed, intelligence, and versatility to meet the needs of modern professionals and teams. -
31
Octrafic
Octrafic
FreeOctrafic is a command-line tool that leverages AI and is available as open source, aimed at simplifying the process of automated API testing and exploration by allowing users to communicate with APIs in natural language rather than having to write complex scripts or set up intricate testing frameworks. By simply directing the tool to any HTTP API or OpenAPI specification, users can articulate their testing requirements in straightforward English, prompting the integrated AI agent to create test scenarios, perform actual HTTP requests, verify responses, and generate organized results. This tool streamlines the entire testing process, encompassing endpoint discovery, request formulation, schema checks, and error identification, which enables developers to prioritize testing logic without getting bogged down by the underlying implementation specifics. Additionally, it accommodates real-time execution against live APIs, ensuring the accuracy of status codes and behaviors without the need for mock setups, and it can also produce aesthetically formatted PDF reports for effective communication with teams or stakeholders. With its user-friendly approach, Octrafic represents a significant advancement in making API testing more accessible and efficient. -
32
Qwen3.6-35B-A3B
Alibaba
FreeQwen3.5-35B-A3B is a member of the Qwen3.5 "Medium" model series, meticulously crafted as an effective multimodal foundation model that strikes a balance between robust reasoning capabilities and practical application needs. Utilizing a Mixture-of-Experts (MoE) architecture, it boasts a total of 35 billion parameters, yet activates only around 3 billion for each token, enabling it to achieve performance levels similar to much larger models while significantly cutting down on computational expenses. The model employs a hybrid attention mechanism that merges linear attention with traditional attention layers, which enhances its ability to handle extensive context and boosts scalability for intricate tasks. As an inherently vision-language model, it processes both textual and visual data, catering to a variety of applications, including multimodal reasoning, programming, and automated workflows. Furthermore, it is engineered to operate as a versatile "AI agent," proficient in planning, utilizing tools, and systematically solving problems, extending its functionality beyond mere conversational interactions. This capability positions it as a valuable asset across diverse domains, where advanced AI-driven solutions are increasingly required. -
33
HiClaw
AgentScope
FreeHiClaw is a multi-agent operating system that is open source and operates on the Matrix framework, allowing various AI agents to work together within Matrix rooms, where their activities are fully accessible to humans in real-time. The system features a Manager Agent that oversees multiple Worker Agents, efficiently breaking down complex tasks and facilitating simultaneous execution, which enhances the management of these intricate operations. Designed with a focus on enterprise-level security and collaborative capabilities, HiClaw utilizes the open Matrix instant messaging protocol, ensuring that all communications between agents are transparent, easily auditable, and fit for distributed systems and federated environments. Humans have the ability to join any Matrix room whenever they wish, which allows them to monitor agent discussions, intervene as necessary, or adjust agent actions in real-time, thereby safeguarding oversight and control. This structured two-tier system, consisting of Manager and Worker Agents, delineates clear responsibilities for each agent, simplifying the process of integrating custom Worker Agents tailored for various applications, while also promoting adaptability within the architecture. Consequently, the design of HiClaw not only enhances operational efficiency but also paves the way for innovative uses of AI collaboration across diverse scenarios. -
34
PyGPT
PyGPT
FreePyGPT is a versatile open-source AI assistant designed for personal use on desktop systems such as Linux, Windows, and Mac, and it is developed using Python. It operates in a manner akin to ChatGPT but functions locally on your computer, providing features like chat, image and video generation, vision capabilities, voice control, and more. Supporting a variety of models, PyGPT includes options like OpenAI's GPT-5, GPT-4, o1, o3, o4, Google Gemini, Anthropic Claude, xAI Grok, Perplexity Sonar, DeepSeek, Mistral AI, alongside models from Ollama and LlamaIndex. Users can choose from 12 operational modes, including chatting with files, real-time audio interactions, research, completion tasks, and various imaging capabilities. With integrated LlamaIndex support, users can engage with their personal files and data seamlessly. Additionally, PyGPT features built-in vector database capabilities, automated embedding of files and data, and maintains full conversation context alongside both short- and long-term memory. The assistant is equipped with internet access through platforms like Google, Microsoft Bing, and DuckDuckGo, enhancing its functionality, which also includes speech synthesis and recognition, making it a comprehensive tool for productivity. Overall, PyGPT stands out as an innovative solution for those seeking a powerful local AI assistant. -
35
OllaCoder
OllaCoder
FreeOllaCoder serves as a private AI coding assistant tailored for VS Code, catering specifically to developers who prefer not to upload their source code to external servers. Operating locally, it utilizes your personal Ollama models and integrates features such as agent mode, inline edits, codebase chat, intelligent autocomplete, MCP servers, and a local-first runtime all within a single editor interface. The core philosophy behind OllaCoder emphasizes the notion that software development is a personal endeavor, asserting that your code should remain under your control while providing an AI assistant that is robust, transparent, and unobtrusive. It primarily communicates with your local Ollama instance, ensuring that prompts, completions, and modifications remain on your device; cloud services are optional, with API keys securely stored in the OS keychain. OllaCoder's agent mode is capable of planning tasks, modifying files, executing terminal commands, and confirming the accuracy of its work, allowing users to approve, reject, or revert any action taken. Additionally, the inline edits feature enables users to select a function, specify the desired change, and examine a real diff change by change, enhancing the coding experience. Overall, OllaCoder represents a significant step forward in maintaining code privacy while providing powerful AI-assisted development tools. -
36
Dock
Dock
$19 per monthDock serves as a collaborative AI workspace designed for you, your team, and the various agents you deploy. It enables both humans and AI agents to share a unified cloud environment, allowing everyone to access and modify the same information in real-time, rather than navigating through disjointed chats, files, and isolated outputs. The platform is structured around tables with defined columns, rich-text documents, and recognizes agents as primary entities, each equipped with their own API keys, permissions, and audit trails, eliminating the need for delegated human tokens. Teams can leverage Dock for a multitude of tasks, including planning, researching, decision-making, and executing projects, all within a shared interface that accommodates both human and AI contributions. Use cases for Dock span various domains, including engineering, go-to-market strategies, research, operations, individual projects, and agency tasks. Engineering teams can utilize Dock to facilitate sprint planning, create specification documents, and respond to incidents efficiently; marketing teams can streamline content calendars, manage sales pipelines, and enhance customer success initiatives; research teams can effectively document interviews, identify themes, and analyze competitive intelligence; and operations teams can oversee runbooks, manage recruitment processes, ensure compliance, and coordinate onboarding efforts. In essence, Dock fosters a seamless collaboration environment that enhances productivity and innovation across all team functions. -
37
Laguna XS.2
Poolside
FreeLaguna XS.2 represents Poolside’s innovative open-weight coding model, distinguished as the lightest and quickest member of the Laguna series. This model features a total of 33 billion parameters in a Mixture of Experts setup, with 3 billion parameters activated, and has been meticulously trained in-house using 30 trillion tokens. As the latest generation model accessible to the public, it embodies a second-generation architecture and marks Poolside’s inaugural open-weight offering, drawing from insights gained during the training of Laguna M.1 with synthetic data and reinforcement learning techniques. Specifically designed to enhance agentic coding workflows, Laguna XS.2 excels in coding, acting, and rapidly iterating, particularly within Poolside’s coding agent environment. This model is particularly advantageous for developers and teams seeking a lightweight, efficient coding solution rather than a more cumbersome frontier system. Released under the permissive Apache 2.0 license, it empowers the community to assess, fine-tune, quantize, and build upon its weights, fostering a collaborative development atmosphere. In essence, Laguna XS.2 not only provides a robust platform for agentic coding but also encourages innovation and experimentation among its users. -
38
Laguna M.1
Poolside
FreeLaguna M.1 stands out as Poolside's most proficient model for agentic coding, meticulously developed in-house specifically for enhancing software development workflows. This model features a total of 225 billion parameters, utilizing a Mixture of Experts architecture with 23 billion activated parameters, and has been trained entirely within the organization on a dataset consisting of 30 trillion tokens, leveraging the power of 6,144 interconnected NVIDIA H200 GPUs. Poolside undertook the task of training Laguna M.1 from the ground up, employing its proprietary data, dedicated training codebase, and an asynchronous on-policy reinforcement learning approach within its agent framework, all tailored for agentic coding applications. The design of the model ensures optimal performance within Poolside's coding agent, enabling it to effectively reason through software tasks, interact with various tools, edit code, execute tests, and facilitate extended autonomous development sessions. Specifically crafted for developers and teams tackling intricate coding challenges, Laguna M.1 offers enhanced capabilities in reasoning, architectural comprehension, terminal operations, and multi-step execution, surpassing what lighter models can achieve. Ultimately, its robust feature set positions it as an essential asset for those engaged in demanding software projects. -
39
Workers by Delos
Delos
$30 per monthAI Workers are independent agents crafted specifically for your organization; these advanced AI entities function like genuine colleagues rather than mere chatbots awaiting instructions. Each comes equipped with its own professional identity, including an email, phone number, and presence on platforms like Slack and Teams, demonstrating initiative and the capacity to operate around the clock without needing prompts. Rather than dictating every detail of their tasks, you simply define the objectives, and they autonomously develop the necessary workflows, accommodating a variety of tasks such as generating daily reports, conducting weekly follow-ups, managing CRM updates, handling client interactions, performing research, overseeing content creation, executing finance responsibilities, coordinating HR activities, supporting design projects, and assisting with development operations. This system includes a diverse range of specialized AI Workers tailored for functions in marketing, development, design, HR, finance, and more, with each worker crafted for a specific role and practical applications. Furthermore, AI Workers can integrate seamlessly with over 3,000 tools, such as Slack, Microsoft Teams, Gmail, Notion, HubSpot, Salesforce, and various other business applications, ensuring they fit effortlessly into your existing workflows. Their versatility and adaptability make them an invaluable asset for enhancing productivity and streamlining processes in any business environment. -
40
OpenWorker
OpenWorker
FreeOpenWorker serves as an open-source, locally-focused AI assistant designed to complete various daily tasks from initiation to conclusion rather than merely providing answers. Users can request specific results like a renewal brief, incident report, follow-up message, calendar update, sprint summary, or finalized document, and OpenWorker seamlessly operates across multiple platforms where the relevant data is stored. It offers integration with a range of services including Slack, Gmail, Outlook, Google Calendar, Notion, HubSpot, GitHub, Attio, Google Drive, Jira, Linear, Asana, Dropbox, Box, and an array of other applications through both one-click and manual connections. The platform accommodates cloud, open-weight, and fully local models, supporting providers such as OpenAI, Anthropic, Google, xAI, Mistral, DeepSeek, Kimi, Qwen, and Ollama, allowing users the flexibility to switch models based on task requirements. OpenWorker excels at researching, gathering necessary context, executing multi-step tasks, and generating refined outputs in various formats like chat, Slack, Markdown, PDF, images, or files, all while ensuring to check in prior to making significant decisions. This comprehensive suite of functionalities empowers users to streamline their workflows and enhances overall productivity. -
41
FluentDB
FluentDB
$54 one-time paymentFluentDB is an advanced database client designed specifically for Apple Silicon Macs, emphasizing an AI-centric approach by integrating a schema-aware assistant, a contemporary SQL editor, a user-friendly keyboard navigator, efficient data tables, and essential safety measures within a singular, streamlined interface. It offers seamless connectivity to various databases, including PostgreSQL, MySQL, SQLite, and Microsoft SQL Server, allowing users to pose inquiries in plain English and receive accurately generated SQL statements tailored to their actual tables and columns. The AI prioritizes schema interpretation over row data, refrains from executing any statements autonomously, and seeks user consent prior to executing any write queries, ensuring a high level of control. Features like read-only connections, local query history, server-side cancellation for unresponsive queries, and transparent SQL execution empower users to maintain oversight of their operations. The editor enhances productivity with schema-aware autocomplete, formatting options, saved queries, instant previews, and the ability to effortlessly toggle between manual SQL and AI-driven modes without any loss of context. Users can access a comprehensive range of database components—such as tables, views, functions, sequences, types, and extensions—by using Command + P, while the intuitive native grid allows for smooth scrolling, editing, sorting, copying, and exporting of data, enhancing the overall user experience. This innovative tool is designed to elevate database management by merging advanced AI capabilities with user-friendly functionality. -
42
Antalogy
Antalogy
0Antalogy is a Markdown editor designed to function like a conventional word processor, featuring a Word-style Ribbon for a familiar experience. You can type freely while the interface remains unobtrusive, ensuring that clean Markdown is saved in the background. - Completely Local Documents: Your files are stored exclusively on your device, devoid of cloud syncing, telemetry, or any risk of document format dependence. - Import from Word .docx: With just one click, you can transform .docx files into .md formats while retaining tables, lists, and converting embedded images to PNG. - Comprehensive Mermaid Diagrams and AI Capabilities: It fully accommodates Mermaid syntax for creating visual data representations and charts directly from text. The built-in AI Assistant can analyze your document's text and automatically generate the corresponding Mermaid code to create accurate visual flowcharts and architecture diagrams instantly. - AI Assistant for Custom LLM Integration: This feature allows you to link to any OpenAI-compatible API for your own local quantized LLMs, whether they are hosted through LMStudio, Ollama, or other private cloud and on-premises inference servers, enhancing your document's functionality even further. This flexibility opens up a world of possibilities for users who want to leverage advanced AI tools in their writing process. -
43
LensHH-LT
Synapse Optics
$385LensHH-LT is a powerful optical design software tailored for engineers focused on refractive and reflective imaging systems, providing essential functionality at a price that doesn't require an enterprise-level investment. The software offers comprehensive analyses, including ray tracing, Seidel aberrations, ray and OPD fans, as well as FFT and geometric MTF, PSF, wavefront maps, spot diagrams, field curvature, distortion, and relative illumination. Users can optimize designs through both local and global methods—utilizing multi-start approaches, basin-hopping searches, glass substitutions from complete catalogs, stock-lens matching, and saddle-point construction for discovering topologies. It supports seamless integration by reading and writing files from Zemax, Code V, OSLO, Optalix, and Optiland, ensuring that existing designs can be imported without loss. In addition to its user-friendly GUI, the software features an API, a command-line interface, and an open MCP server, enabling AI agents like Claude to automate design processes, perform optimizations, and retrieve analysis results. Priced at a one-time fee of $385, it does not require a subscription and includes a full-featured 45-day trial without the need for an account. For users with air-gapped systems, offline activation is also available, offering flexibility for various working environments. Overall, LensHH-LT stands out for its robust functionality and accessibility to engineers seeking a cost-effective solution. -
44
AtomCode
AtomGit
FreeAtomCode is an innovative open-source AI coding assistant that operates directly within the terminal, enabling it to autonomously read and edit files, run commands, search the web, conduct tests, and verify its own work until each task is accomplished. Serving as a multi-model alternative to platforms like Claude Code and Cursor Agent, it is compatible with a range of models including Claude, OpenAI, DeepSeek, GLM, Qwen, Ollama, SiliconFlow, and any other API that aligns with OpenAI's standards. The agent's advanced code graph functionalities facilitate symbol indexing, reference lookup, caller and callee tracing, dependency analysis, and blast-radius analysis, allowing it to navigate extensive codebases with a depth of understanding that transcends simple text searches. Additionally, developers have the capability to attach screenshots and images, with vision preprocessing available to derive valuable context when the primary model lacks direct image support. AtomCode also features seamless integration with AtomGit for managing OAuth logins, repositories, issue tracking, and pull requests, while further enhancing its utility with support for MCP, reusable Skills, plugins, custom slash commands, hooks, and workflows. This comprehensive set of features makes AtomCode a robust tool for developers seeking efficiency and versatility in their coding tasks. -
45
traxy
traxy.ai
$149/month Traxy serves as a signal-to-leads platform designed for go-to-market strategies driven by LinkedIn; it monitors activities surrounding the accounts and individuals you prioritize, transforms those interactions into meaningful “signals,” and then translates those signals into prioritized, actionable leads for your team to pursue—eliminating the need for you to constantly engage with LinkedIn throughout the day. Additionally, this system streamlines the lead generation process, allowing your team to focus on high-impact opportunities. -
46
Amelia's Agent
Amelia's Agent
$4.99 per monthAmelia’s Agent is an innovative AI tool that transforms simple English descriptions into a variety of digital products, including websites, applications, games, tools, automations, and code. Users initiate the process with a straightforward sentence, respond to a set of clarifying questions, and then review and approve a proposed plan before any execution occurs. The AI then constructs the project on a secure private cloud server, with a transparent display of all files and commands as they are generated. Every project is archived in a genuine Git repository from the very first build, ensuring that each version is meticulously documented and that the user retains full ownership of the code. Users have the capability to clone, download, upload to GitHub, or execute the project on their local machines. Additionally, Amelia is capable of enhancing existing websites and repositories, analyzing them prior to implementing any accepted modifications. Supporting over 300 AI models, it intelligently directs each task to the most appropriate model while giving advanced users the flexibility to select models directly, utilize their personal API keys, or execute open-source models via Ollama or LM Studio. This versatility makes Amelia's Agent a powerful ally in the realm of AI-driven development. -
47
Llama 2
Meta
FreeIntroducing the next iteration of our open-source large language model, this version features model weights along with initial code for the pretrained and fine-tuned Llama language models, which span from 7 billion to 70 billion parameters. The Llama 2 pretrained models have been developed using an impressive 2 trillion tokens and offer double the context length compared to their predecessor, Llama 1. Furthermore, the fine-tuned models have been enhanced through the analysis of over 1 million human annotations. Llama 2 demonstrates superior performance against various other open-source language models across multiple external benchmarks, excelling in areas such as reasoning, coding capabilities, proficiency, and knowledge assessments. For its training, Llama 2 utilized publicly accessible online data sources, while the fine-tuned variant, Llama-2-chat, incorporates publicly available instruction datasets along with the aforementioned extensive human annotations. Our initiative enjoys strong support from a diverse array of global stakeholders who are enthusiastic about our open approach to AI, including companies that have provided valuable early feedback and are eager to collaborate using Llama 2. The excitement surrounding Llama 2 signifies a pivotal shift in how AI can be developed and utilized collectively. -
48
Code Llama
Meta
FreeCode Llama is an advanced language model designed to generate code through text prompts, distinguishing itself as a leading tool among publicly accessible models for coding tasks. This innovative model not only streamlines workflows for existing developers but also aids beginners in overcoming challenges associated with learning to code. Its versatility positions Code Llama as both a valuable productivity enhancer and an educational resource, assisting programmers in creating more robust and well-documented software solutions. Additionally, users can generate both code and natural language explanations by providing either type of prompt, making it an adaptable tool for various programming needs. Available for free for both research and commercial applications, Code Llama is built upon Llama 2 architecture and comes in three distinct versions: the foundational Code Llama model, Code Llama - Python which is tailored specifically for Python programming, and Code Llama - Instruct, optimized for comprehending and executing natural language directives effectively. -
49
GPT Pilot
Pythagora
FreeGPT Pilot is an innovative open-source AI application designed to function as a comprehensive AI developer, generating fully operational applications with very little human intervention. In contrast to basic code autocompletion utilities, GPT Pilot is capable of creating entire features, troubleshooting code, discussing problems, and even soliciting code reviews. This tool seeks to expand the horizons of AI-driven software development by managing as much as 95% of coding responsibilities, reserving the remaining 5% for human developers. Additionally, it is designed to work seamlessly with platforms such as VS Code, allowing for real-time collaboration between developers and AI. By facilitating this partnership, GPT Pilot empowers developers to focus on more complex tasks while the AI handles routine coding challenges. -
50
RouteLLM
LMSYS
Created by LM-SYS, RouteLLM is a publicly available toolkit that enables users to direct tasks among various large language models to enhance resource management and efficiency. It features strategy-driven routing, which assists developers in optimizing speed, precision, and expenses by dynamically choosing the most suitable model for each specific input. This innovative approach not only streamlines workflows but also enhances the overall performance of language model applications.