List of Ollama Integrations
This is a list of platforms and tools that integrate with Ollama. This list is updated as of September 2026.
-
1
Gemma 3n
Google DeepMind
Empower your apps with efficient, intelligent, on-device capabilities!Meet Gemma 3n, our state-of-the-art open multimodal model engineered for exceptional performance and efficiency on devices. Emphasizing responsive and low-footprint local inference, Gemma 3n sets the stage for a new era of intelligent applications that can be deployed while on the go. It possesses the ability to interpret and react to a combination of images and text, with upcoming plans to add video and audio capabilities shortly. This allows developers to build smart, interactive functionalities that uphold user privacy and operate smoothly without relying on an internet connection. The model features a mobile-centric design that significantly reduces memory consumption. Jointly developed by Google's mobile hardware teams and industry specialists, it maintains a 4B active memory footprint while providing the option to create submodels for enhanced quality and reduced latency. Furthermore, Gemma 3n is our first open model constructed on this groundbreaking shared architecture, allowing developers to begin experimenting with this sophisticated technology today in its initial preview. As the landscape of technology continues to evolve, we foresee an array of innovative applications emerging from this powerful framework, further expanding its potential in various domains. The future looks promising as more features and enhancements are anticipated to enrich the user experience. -
2
Elestio
Elestio
Effortlessly deploy and manage software with complete security.Elestio is an all-encompassing DevOps platform that empowers users to swiftly deploy more than 350 open-source software applications on dedicated virtual machines in under three minutes. By handling essential tasks such as installation, configuration, encryption, backups, and updates for both software and operating systems, it allows users to focus on optimizing the software's functionality. The platform's adaptability in deployment options is noteworthy, as it accommodates a range of cloud providers, including DigitalOcean, AWS, VULTR, Hetzner, Linode, and Scaleway, as well as on-premise configurations, thus enhancing flexibility and reducing the chances of vendor lock-in. Powered entirely by dedicated hardware, Elestio ensures users enjoy full resource access along with superior kernel-level security measures. It prioritizes data safety through end-to-end TLS encryption for all connections between the user’s device, the dashboard, and the services. Moreover, Elestio incorporates a fully managed CI/CD system that easily integrates with GitHub, GitLab, and Docker registries, while also being compatible with any Linux technology stack. This broad functionality positions Elestio as an exceptional option for developers who seek a trustworthy and secure environment for their deployment needs. In addition, the platform's user-friendly interface makes it accessible to both novice and experienced developers alike. -
3
GLM-5V-Turbo
Z.ai
Transforming visions into code with seamless multimodal intelligence.The GLM-5V-Turbo stands as a cutting-edge multimodal coding foundation model, expertly designed for scenarios necessitating visual inputs, proficient in interpreting various formats including images, videos, texts, and files to produce text-based results. This model is particularly optimized for agent workflows, enabling it to grasp environments effectively, devise suitable actions, and execute tasks, while also maintaining compatibility with agent frameworks such as Claude Code and OpenClaw. Notably, it excels in managing long-context interactions, offering an impressive context capacity of 200K tokens alongside an output limit of up to 128K tokens, making it exceptionally suited for complex, long-duration projects. Moreover, it presents an array of thinking modes tailored for different situations, demonstrates strong visual understanding of both images and videos, and streams outputs in real-time to improve user interaction. It also incorporates advanced function-calling capabilities that allow seamless integration of external tools, with its context caching feature significantly enhancing performance during extended dialogues. In real-world applications, the model is capable of skillfully converting design mockups into operational frontend projects, highlighting its adaptability and depth in practical coding environments. Furthermore, this adaptability empowers users to approach a diverse array of intricate tasks with assurance and effectiveness, greatly enhancing their productivity. -
4
Qwen3.6
Alibaba
Unlock powerful AI solutions for coding and reasoning.Qwen3.6 is a next-generation large language model developed by Alibaba, designed to deliver advanced reasoning, coding, and multimodal capabilities. It builds on the Qwen3.5 series with a strong emphasis on stability, efficiency, and real-world usability. The model supports multimodal inputs, enabling it to process text, images, and video for more complex analysis and decision-making. One of its key strengths is agentic AI, allowing it to perform multi-step tasks and operate more autonomously in workflows. Qwen3.6 is particularly optimized for coding, capable of handling complex engineering tasks at a repository level rather than just individual functions. It uses a mixture-of-experts architecture, with billions of parameters but only a subset activated during each inference, improving efficiency. The model is available in both open-weight and proprietary versions, giving developers flexibility in deployment and customization. It can be integrated into enterprise systems, APIs, and cloud environments for production use. Qwen3.6 also offers strong multimodal reasoning, enabling it to analyze documents, visuals, and structured data together. It is designed to support a wide range of applications, from software development to data analysis and automation. The model includes enhancements in performance, scalability, and usability compared to earlier versions. It reflects a broader shift toward agent-based AI systems that can execute tasks rather than just provide responses. Overall, Qwen3.6 represents a powerful and versatile AI model for modern enterprise and developer use cases. -
5
Osaurus
Osaurus
Empower your Mac with intelligent, versatile AI agents.Osaurus stands out as a cutting-edge AI platform tailored for macOS, allowing users to run open models directly on their Macs while also offering the flexibility to utilize cloud models for improved performance and seamless memory sharing. Crafted with Swift specifically for Apple Silicon, it ensures full offline capability with local models through interfaces such as Ollama, MLX, or LM Studio, safeguarding conversations, code, files, settings, agents, skills, and provider keys on the device unless the user chooses to engage with cloud services. Users benefit from a handy system-wide chat overlay that provides immediate access to the AI across any application, and they can create unique agents designed for specific tasks like coding, research, and file management, with each agent preserving its own prompt, history, and memory. Osaurus adeptly synthesizes prior discussions into relevant insights, loads necessary skills based on specific tasks, and equips agents with focused access to important directories, file search functions, Git, and other vital tools. Furthermore, agents function within a secure sandbox environment, capable of assigning tasks to subagents, following scheduled actions, responding to folder changes, utilizing voice commands, and generating images, reflecting the platform's extensive versatility and feature-rich capabilities. This holistic approach not only boosts productivity but also gives users the power to tailor their workflows to meet their precise needs, ultimately enriching their overall experience. -
6
Qwen3.8-Flash-Next
Alibaba
Revolutionizing AI with efficient, powerful multimodal capabilities.Qwen3.8-Flash-Next is a pioneering open-weight multimodal Mixture-of-Experts architecture that offers an initial look at the design meant for its successor, Qwen4. This model has been expertly crafted to enhance various aspects such as attention mechanisms, residual pathways, embeddings, and optimization strategies, thereby increasing its overall functionality, enhancing computational efficiency, expanding its model capacity, and ensuring stability during training. Its unique hybrid structure combines Gated DeltaNet, which effectively condenses historical information, with Qwen Sparse Attention, facilitating the selection of meaningful context on a micro-block scale to reduce both attention and indexing expenses for lengthy sequences. The Gated Residual feature enhances the residual pathway by incorporating four streams, which helps in dynamically regulating the information flow across different layers. Moreover, the N-gram Embedding cleverly merges large-scale local-pattern memory with minimal computational overhead for each token, with the capability to transfer to host memory for added efficiency. The entire model is built around a main network comprising 125 billion parameters, supplemented by an additional 51 billion parameters specifically for N-gram embeddings, activating only 6 billion parameters for each token processed. This advanced framework underscores the continuous evolution in machine learning architectures, laying the groundwork for exciting future innovations, and it exemplifies the increasing sophistication and potential of multimodal models in various applications. -
7
Step 5 Preview
StepFun
Empower your productivity with advanced multimodal task mastery.Step 5 Preview epitomizes the apex of StepFun’s offerings for agentic tasks, specifically designed for practical applications in the realms of software engineering and professional knowledge, with particular excellence in financial settings. This model is adept at processing text, images, and videos, featuring an impressive 1M-token context window that suits tasks requiring extensive data, tool utilization, and continuous advancement toward specific objectives. It can effectively analyze extensive documents, amalgamate information from diverse sources, and leverage conversation threads for efficient cross-document queries and research organization. In the field of programming and software development, it showcases proficiency in a range of programming languages, assisting with debugging, code adjustments, verification tasks, and test generation. Moreover, its sophisticated multi-step agent capabilities allow applications to leverage tools for information retrieval, document analysis, detailed research, and the development of analytical reports. The model's multimodal understanding enables it to integrate images, videos, and text, facilitating tasks like chart analysis and responding to queries based on screenshots. This extensive suite of capabilities not only enhances productivity but also solidifies Step 5 Preview as an essential tool for professionals across a multitude of industries, ensuring they remain at the forefront of their respective fields. -
8
Second State
Second State
Lightweight, powerful solutions for seamless AI integration everywhere.Our solution, which is lightweight, swift, portable, and powered by Rust, is specifically engineered for compatibility with OpenAI technologies. To enhance microservices designed for web applications, we partner with cloud providers that focus on edge cloud and CDN compute. Our offerings address a diverse range of use cases, including AI inference, database interactions, CRM systems, ecommerce, workflow management, and server-side rendering. We also incorporate streaming frameworks and databases to support embedded serverless functions aimed at data filtering and analytics. These serverless functions may act as user-defined functions (UDFs) in databases or be involved in data ingestion and query result streams. With an emphasis on optimizing GPU utilization, our platform provides a "write once, deploy anywhere" experience. In just five minutes, users can begin leveraging the Llama 2 series of models directly on their devices. A notable strategy for developing AI agents that can access external knowledge bases is retrieval-augmented generation (RAG), which we support seamlessly. Additionally, you can effortlessly set up an HTTP microservice for image classification that effectively runs YOLO and Mediapipe models at peak GPU performance, reflecting our dedication to delivering robust and efficient computing solutions. This functionality not only enhances performance but also paves the way for groundbreaking applications in sectors such as security, healthcare, and automatic content moderation, thereby expanding the potential impact of our technology across various industries. -
9
Gemma
Google
Revolutionary lightweight models empowering developers through innovative AI.Gemma encompasses a series of innovative, lightweight open models inspired by the foundational research and technology that drive the Gemini models. Developed by Google DeepMind in collaboration with various teams at Google, the term "gemma" derives from Latin, meaning "precious stone." Alongside the release of our model weights, we are also providing resources designed to foster developer creativity, promote collaboration, and uphold ethical standards in the use of Gemma models. Sharing essential technical and infrastructural components with Gemini, our leading AI model available today, the 2B and 7B versions of Gemma demonstrate exceptional performance in their weight classes relative to other open models. Notably, these models are capable of running seamlessly on a developer's laptop or desktop, showcasing their adaptability. Moreover, Gemma has proven to not only surpass much larger models on key performance benchmarks but also adhere to our rigorous standards for producing safe and responsible outputs, thereby serving as an invaluable tool for developers seeking to leverage advanced AI capabilities. As such, Gemma represents a significant advancement in accessible AI technology. -
10
EvalsOne
EvalsOne
Unlock AI potential with streamlined evaluations and expert insights.Explore an intuitive yet comprehensive evaluation platform aimed at the continuous improvement of your AI-driven products. By streamlining the LLMOps workflow, you can build trust and gain a competitive edge in the market. EvalsOne acts as an all-in-one toolkit to enhance your application evaluation methodology. Think of it as a multifunctional Swiss Army knife for AI, equipped to tackle any evaluation obstacle you may face. It is perfect for crafting LLM prompts, refining retrieval-augmented generation strategies, and evaluating AI agents effectively. You have the option to choose between rule-based methods or LLM-centric approaches to automate your evaluations. In addition, EvalsOne facilitates the effortless incorporation of human assessments, leveraging expert feedback for improved accuracy. This platform is useful at every stage of LLMOps, from initial concept development to final production rollout. With its user-friendly design, EvalsOne supports a wide range of professionals in the AI field, including developers, researchers, and industry experts. Initiating evaluation runs and organizing them by various levels is a straightforward process. The platform also allows for rapid iterations and comprehensive analyses through forked runs, ensuring that your evaluation process is both efficient and effective. As the landscape of AI development continues to evolve, EvalsOne is tailored to meet these changing demands, making it an indispensable resource for any team aiming for excellence in their AI initiatives. Whether you are looking to push the boundaries of your technology or simply streamline your workflow, EvalsOne stands ready to assist you. -
11
Gemma 2
Google
Unleashing powerful, adaptable AI models for every need.The Gemma family is composed of advanced and lightweight models that are built upon the same groundbreaking research and technology as the Gemini line. These state-of-the-art models come with powerful security features that foster responsible and trustworthy AI usage, a result of meticulously selected data sets and comprehensive refinements. Remarkably, the Gemma models perform exceptionally well in their varied sizes—2B, 7B, 9B, and 27B—frequently surpassing the capabilities of some larger open models. With the launch of Keras 3.0, users benefit from seamless integration with JAX, TensorFlow, and PyTorch, allowing for adaptable framework choices tailored to specific tasks. Optimized for peak performance and exceptional efficiency, Gemma 2 in particular is designed for swift inference on a wide range of hardware platforms. Moreover, the Gemma family encompasses a variety of models tailored to meet different use cases, ensuring effective adaptation to user needs. These lightweight language models are equipped with a decoder and have undergone training on a broad spectrum of textual data, programming code, and mathematical concepts, which significantly boosts their versatility and utility across numerous applications. This diverse approach not only enhances their performance but also positions them as a valuable resource for developers and researchers alike. -
12
Continue
Continue
Empower your coding with personalized, seamless AI assistance!The premier open-source AI assistant allows you to design personalized autocomplete functions and chat interfaces by seamlessly linking various models to diverse contexts. By eliminating obstacles that disrupt your productivity during software development, you can maintain a smooth workflow. With a user-friendly plug-and-play system that integrates effortlessly into your entire tech stack, you can expedite your development process. Customize your code assistant to adapt as it gains new features and enhancements. Effortlessly autocomplete entire blocks of code or individual lines across multiple programming languages as you continue typing. You can inquire about files, functions, the full codebase, and more by associating relevant code or context. Additionally, highlight specific segments of code and utilize a keyboard shortcut to transform that code into easily understandable natural language, enhancing collaboration and comprehension. This innovative tool empowers developers to work more efficiently and creatively than ever before. -
13
Azure Marketplace
Microsoft
Unlock cloud potential with diverse solutions for businesses.The Azure Marketplace operates as a vast digital platform, offering users access to a multitude of certified software applications, services, and solutions from Microsoft along with numerous third-party vendors. This marketplace enables businesses to efficiently find, obtain, and deploy software directly within the Azure cloud ecosystem. It showcases a wide range of offerings, including virtual machine images, frameworks for AI and machine learning, developer tools, security solutions, and niche applications designed for specific sectors. With a variety of pricing options such as pay-as-you-go, free trials, and subscription-based plans, the Azure Marketplace streamlines the purchasing process while allowing for consolidated billing through a unified Azure invoice. Additionally, it guarantees seamless integration with Azure services, which empowers organizations to strengthen their cloud infrastructure, improve operational efficiency, and accelerate their journeys toward digital transformation. In essence, the Azure Marketplace is crucial for enterprises aiming to stay ahead in a rapidly changing technological environment while fostering innovation and adaptability. This platform is not just a marketplace; it is a gateway to unlocking the potential of cloud capabilities for businesses worldwide. -
14
Langflow
Langflow
Empower your AI projects with seamless low-code innovation.Langflow is a low-code platform designed for AI application development that empowers users to harness agentic capabilities alongside retrieval-augmented generation. Its user-friendly visual interface allows developers to construct complex AI workflows effortlessly through drag-and-drop components, facilitating a more efficient experimentation and prototyping process. Since it is based on Python and does not rely on any particular model, API, or database, Langflow offers seamless integration with a broad spectrum of tools and technology stacks. This flexibility enables the creation of sophisticated applications such as intelligent chatbots, document processing systems, and multi-agent frameworks. The platform provides dynamic input variables, fine-tuning capabilities, and the option to create custom components tailored to individual project requirements. Additionally, Langflow integrates smoothly with a variety of services, including Cohere, Bing, Anthropic, HuggingFace, OpenAI, and Pinecone, among others. Developers can choose to utilize pre-built components or develop their own code, enhancing the platform's adaptability for AI application development. Furthermore, Langflow includes a complimentary cloud service, allowing users to swiftly deploy and test their projects, which promotes innovation and rapid iteration in AI solution creation. Overall, Langflow emerges as an all-encompassing solution for anyone eager to effectively utilize AI technology in their projects. This comprehensive approach ensures that users can maximize their productivity while exploring the vast potential of AI applications. -
15
Witsy
Witsy
Unlock creativity and productivity with seamless AI collaboration.Witsy is a versatile desktop application that provides users with access to a wide variety of generative AI models from top AI providers, serving as a one-stop solution for all generative AI needs. Designed as a BYOK (Bring Your Own Keys) application, Witsy requires users to supply their own API keys for the LLM providers they wish to engage with. Additionally, users can choose to utilize Ollama to run models locally at no charge and seamlessly incorporate them into Witsy. A key aspect of Witsy's design is its commitment to user privacy; it does not collect or process any personal information, ensuring that all data remains securely on the user's device. Furthermore, the application avoids using cookies or any tracking mechanisms, which adds an extra layer of protection for user privacy. All features within Witsy are accessible via convenient keyboard shortcuts, allowing users to easily start chats, use the scratchpad, execute commands, and perform other tasks efficiently. Users also have the ability to customize these shortcuts according to their personal preferences, enhancing their interaction with the application. Another beneficial feature is the capability to engage with the AI model in the scratchpad, which streamlines the document creation process. This unique approach fosters a collaborative environment, enabling users to work with AI as though they are partnering with a colleague, ultimately making Witsy an essential asset for boosting productivity and creativity. As a result, Witsy stands out in the market as a user-friendly tool that adapts to individual workflows and enhances overall efficiency. -
16
Open WebUI
Open WebUI
Empower your AI journey with versatile, offline functionality.Open WebUI is a powerful, adaptable, and user-friendly AI platform that can be self-hosted and operates fully offline. It accommodates various LLM runners, including Ollama, and adheres to OpenAI-compliant APIs while featuring an integrated inference engine that enhances Retrieval Augmented Generation (RAG), making it a compelling option for AI deployment. Key features encompass an easy installation via Docker or Kubernetes, seamless integration with OpenAI-compatible APIs, comprehensive user group management and permissions for enhanced security, and a mobile-responsive design that supports both Markdown and LaTeX. Additionally, Open WebUI offers a Progressive Web App (PWA) version for mobile devices, enabling offline access and a user experience comparable to that of native apps. The platform also includes a Model Builder, allowing users to create customized models based on foundational Ollama models directly within the interface. With a thriving community exceeding 156,000 members, Open WebUI stands out as a versatile and secure solution for managing and deploying AI models, making it a superb choice for both individuals and businesses that require offline functionality. Its ongoing updates and enhancements ensure that it remains relevant and beneficial in the rapidly changing AI technology landscape, continually attracting new users and fostering innovation. -
17
TensorWave
TensorWave
Unleash unmatched AI performance with scalable, efficient cloud technology.TensorWave is a dedicated cloud platform tailored for artificial intelligence and high-performance computing, exclusively leveraging AMD Instinct Series GPUs to guarantee peak performance. It boasts a robust infrastructure that is both high-bandwidth and memory-optimized, allowing it to effortlessly scale to meet the demands of even the most challenging training or inference workloads. Users can quickly access AMD’s premier GPUs within seconds, including cutting-edge models like the MI300X and MI325X, which are celebrated for their impressive memory capacity and bandwidth, featuring up to 256GB of HBM3E and speeds reaching 6.0TB/s. The architecture of TensorWave is enhanced with UEC-ready capabilities, advancing the future of Ethernet technology for AI and HPC networking, while its direct liquid cooling systems contribute to a significantly lower total cost of ownership, yielding energy savings of up to 51% in data centers. The platform also integrates high-speed network storage, delivering transformative enhancements in performance, security, and scalability essential for AI workflows. In addition, TensorWave ensures smooth compatibility with a diverse array of tools and platforms, accommodating multiple models and libraries to enrich the user experience. This platform not only excels in performance and efficiency but also adapts to the rapidly changing landscape of AI technology, solidifying its role as a leader in the industry. Overall, TensorWave is committed to empowering users with cutting-edge solutions that drive innovation and productivity in AI initiatives. -
18
Sim Studio
Sim Studio
Empower your workflow design with seamless multi-agent application development.Sim Studio is a powerful platform that harnesses artificial intelligence to enable the design, testing, and launch of workflows driven by agents, boasting a user-friendly visual editor akin to Figma that eliminates the requirement for repetitive coding while easing the infrastructure challenges. Developers can quickly embark on the journey of creating multi-agent applications, gaining full command over system prompts, defining tool parameters, adjusting sampling configurations, and organizing output formats, all while seamlessly switching between various LLM providers like OpenAI, Anthropic, Claude, Llama, and Gemini without the hassle of rewriting their code. The platform enhances local development capabilities through its integration with Ollama, which ensures user privacy and reduces costs during the initial prototyping phase, and it later accommodates scalable deployment in the cloud as projects evolve. With Sim Studio, users can efficiently link their agents to current tools and data repositories, facilitating automatic importation of knowledge bases and providing access to an extensive library of over 40 pre-built integrations. This effortless integration feature greatly boosts productivity, streamlining the workflow creation process even further, allowing for rapid iteration and refinement of applications. As a result, developers can focus on innovation rather than getting bogged down by technical complexities. -
19
Droidrun
Droidrun
Effortless mobile automation with natural-language control and scalability.Droidrun functions as an advanced mobile agent platform that enables users to command real Android devices using natural language, thereby automating a wide range of mobile app activities such as logging in, making reservations, purchasing products, and retrieving data, including access to content often restricted by app logins or platform constraints. The platform’s cloud-based architecture supports the quick deployment of agents that come pre-equipped with necessary applications, allowing users to execute tasks on multiple devices at once and to create complex, multi-step workflows driven by conversational commands; furthermore, users can replay recorded workflows at faster speeds for efficiency. Credential management features facilitate easy storage and retrieval of login information for future tasks, and the system is designed for smooth integration with various existing technologies, such as LLMs, N8N, or personalized scripts, thus bolstering larger automation efforts. Developers benefit from access to SDK examples, which include Python integrations with platforms like Gemini and Ollama, simplifying the incorporation of Droidrun into their pre-existing development environments. This all-encompassing strategy not only optimizes mobile automation processes but also encourages innovation, enabling developers to create customized solutions that precisely meet their unique requirements. As a result, organizations can enhance productivity and achieve greater operational efficiency through the tailored use of Droidrun’s capabilities. -
20
gpt-oss-20b
OpenAI
Empower your AI workflows with advanced, explainable reasoning.gpt-oss-20b is a robust text-only reasoning model featuring 20 billion parameters, released under the Apache 2.0 license and shaped by OpenAI’s gpt-oss usage guidelines, aimed at simplifying the integration into customized AI workflows via the Responses API without reliance on proprietary systems. It has been meticulously designed to perform exceptionally in following instructions, offering capabilities like adjustable reasoning effort, detailed chain-of-thought outputs, and the option to leverage native tools such as web search and Python execution, which leads to well-structured and coherent responses. Developers must take responsibility for implementing their own deployment safeguards, including input filtering, output monitoring, and compliance with usage policies, to ensure alignment with protective measures typically associated with hosted solutions and to minimize the risk of malicious or unintended actions. Furthermore, its open-weight architecture is particularly advantageous for on-premises or edge deployments, highlighting the significance of control, customization, and transparency to cater to specific user requirements. This flexibility empowers organizations to adapt the model to their distinct needs while upholding a high standard of operational integrity and performance. As a result, gpt-oss-20b not only enhances user experience but also promotes responsible AI usage across various applications. -
21
gpt-oss-120b
OpenAI
Powerful reasoning model for advanced text-based applications.gpt-oss-120b is a reasoning model focused solely on text, boasting 120 billion parameters, and is released under the Apache 2.0 license while adhering to OpenAI’s usage policies; it has been developed with contributions from the open-source community and is compatible with the Responses API. This model excels at executing instructions and utilizes various tools, including web searches and Python code execution, which allows for a customizable level of reasoning effort and results in detailed chain-of-thought outputs that can seamlessly fit into different workflows. Although it is constructed to comply with OpenAI's safety policies, its open-weight nature poses a risk, as adept users might modify it to bypass these protections, thereby prompting developers and organizations to implement additional safety measures akin to those of managed models. Assessments reveal that gpt-oss-120b falls short of high performance in specialized fields such as biology, chemistry, or cybersecurity, even after attempts at adversarial fine-tuning. Moreover, its introduction does not represent a substantial advancement in biological capabilities, indicating a cautious stance regarding its use. Consequently, it is advisable for users to stay alert to the potential risks associated with its open-weight attributes, and to consider the implications of its deployment in sensitive environments. As awareness of these factors grows, the community's approach to managing such technologies will evolve and adapt. -
22
ChatKit
OpenAI
Empower your apps with seamless, intelligent chat integration.ChatKit is a multifunctional toolkit tailored for developers aiming to effortlessly integrate and manage chat agents across a variety of applications and websites. It provides a diverse array of features, including the capacity to interact with external documents, text-to-speech capabilities, customizable prompt templates, and convenient shortcut triggers for quick access. Users can either employ their personal OpenAI API key, which entails costs according to OpenAI’s token pricing, or opt for ChatKit's credit system, which requires a license for use. This platform supports multiple model backends, such as OpenAI, Azure OpenAI, Google Gemini, and Ollama, alongside various routing frameworks like OpenRouter. Moreover, ChatKit includes functionalities like cloud synchronization, tools for team collaboration, web accessibility, launcher widgets, and organized conversation flows, which collectively enhance its usability. Ultimately, ChatKit simplifies the deployment of advanced chat agents, enabling developers to concentrate on enhancing functionality rather than building an entire chat infrastructure from scratch. With its wide-ranging capabilities, it not only empowers teams to create more engaging user interactions but also facilitates a more streamlined development process. By leveraging these features, developers can significantly improve the overall efficiency and effectiveness of their chat applications. -
23
TaskMaster AI
TaskMaster AI
Streamline AI workflows with structured tasks and dependencies.Taskmaster serves as an innovative project management tool driven by artificial intelligence, designed to streamline the coordination and supervision of AI agents as they tackle complex workflows by breaking down broad objectives into specific, manageable tasks with clear dependencies. Functioning as a flexible "project manager" for AI-driven initiatives, it empowers users to define requirements, automatically generate detailed task lists, and oversee execution while maintaining contextual awareness across extensive, multi-step processes. Additionally, this platform can create Product Requirement Documents (PRDs) that can be transformed into actionable tasks and subtasks, enabling agents to work in a logical and cohesive manner while remaining conscious of prior activities. Moreover, it integrates smoothly with a variety of AI providers and models, facilitating customized setups for primary, research, and backup agents, which significantly boosts both efficiency and reliability. The intuitive interface of Taskmaster further enhances the workflow management experience, making it user-friendly for teams that operate with a range of AI technologies. As a result, organizations can achieve improved collaboration and productivity when utilizing this tool. -
24
Singulr
Singulr
Empowering organizations to secure and optimize AI seamlessly.Singulr serves as a holistic platform tailored for enterprise AI governance and security, offering a unified control structure that supports organizations in discovering, securing, and optimizing their extensive AI deployments. By addressing the growing disparity between the swift adoption of AI technologies and the limitations of governance, it provides unmatched insight into all AI systems employed within the organization, encompassing bespoke applications, integrated AI solutions, publicly available tools, and shadow AI, which frequently bypass security detection. The platform meticulously identifies and inventories AI resources across the enterprise, establishing a dynamic record of agents, models, and services, while assessing their respective risks through comprehensive evaluations of data management, model lineage, vulnerabilities, and compliance standards. Furthermore, Singulr Pulse, the platform’s intelligence layer, analyzes millions of AI systems, assigns risk classifications, and streamlines automated onboarding processes, dramatically reducing approval timelines from weeks to just hours, all while maintaining stringent security protocols. This forward-thinking methodology not only improves the efficiency of AI adoption but also enables organizations to uphold a robust governance structure as they navigate the intricate landscape of AI integration. In doing so, it positions organizations to better respond to the challenges and opportunities presented by the evolving AI landscape. -
25
Cherry Studio
Cherry Studio
Unify your AI experience with seamless, powerful productivity.Cherry Studio is a versatile AI assistant and multi-platform desktop application that amalgamates various AI models into a unified workspace suitable for Windows, macOS, and Linux systems. By establishing connections with top-tier model providers, it allows users to effortlessly shift between different AI services, eliminating the need to juggle multiple applications, browser tabs, or fragmented workflows. Designed to serve as a powerful local AI productivity hub, the tool supports a wide array of tasks such as chatting, writing, translation, research, coding help, document analysis, image interpretation, and multimodal AI workflows, all accessible through a single interface. Users can personalize the model providers, manage assistants, organize conversations, and choose different models tailored to their specific needs, making Cherry Studio particularly beneficial for both casual users and those involved in complex experimentation. Moreover, its assistant system enables users to create, subscribe to, and manage role-based assistants with customized prompts for diverse situations, including product management, community engagement, technical support, and strategic planning, which not only enhances user efficiency but also enriches the overall experience. This adaptability empowers both individuals and teams to effectively leverage AI, allowing them to align their tools with their distinct workflows and objectives, ultimately maximizing productivity and innovation in their endeavors. -
26
Qwen3.7-Plus
Alibaba
Empower your insights with seamless vision-language integration.Qwen3.7-Plus represents a cutting-edge multimodal agent model that effectively merges vision and language into a flexible foundation for intelligent agents. Building on the agentic capabilities of Qwen3.7, it expands its functionality to encompass visual understanding, reasoning, grounded interactions, and the utilization of diverse multimodal tools, enabling agents to interpret, analyze, and navigate through text, images, documents, screens, and complex real-world environments. This model is specifically designed for dynamic tasks that extend beyond simple question answering, facilitating a range of activities such as visual searches, document comprehension, evaluations of charts and tables, screen analysis, GUI interactions, image-based reasoning, and workflows that integrate perception, planning, and action. Qwen3.7-Plus strengthens the connection between linguistic reasoning and visual signals, equipping users to ask questions about images, interpret intricate multimodal data, extract structured information, and generate replies that blend contextual and visual components, thereby enhancing the potential for interactive AI applications. With these advancements, users are empowered to engage in more complex and refined interactions with the system, transforming it into a highly effective tool for a multitude of practical uses across various fields. The model’s ability to adapt to different scenarios further solidifies its relevance in today’s rapidly evolving technological landscape. -
27
Laguna XS 2.1
Poolside
Empowering coding agents for seamless, long-horizon workflows.The Laguna XS 2.1 represents a sophisticated advancement in coding models, functioning as an open weight agentic system that excels in executing long-duration tasks on local machines. It boasts a robust 33-billion-parameter Mixture-of-Experts architecture, activating 3 billion parameters per token, while preserving the efficient design of its predecessor, Laguna XS.2, and significantly enhancing its capabilities in multilingual software engineering and terminal-related tasks. This model is meticulously crafted to support coding agents in reviewing code repositories, navigating complex changes, leveraging diverse tools, executing commands, and ensuring seamless progress throughout extensive projects. With an impressive context window of 256K, it empowers agents to adeptly handle large codebases, maintain extensive histories, and navigate intricate multi-step workflows. The Laguna XS 2.1 also enjoys compatibility with various platforms like vLLM, SGLang, NVIDIA TensorRT-LLM, Hugging Face Transformers, and Ollama, with aspirations for future native support from llama.cpp. Offered in multiple checkpoint formats such as BF16, FP8, INT4, and NVFP4, it allows developers to choose between high fidelity and configurations designed for environments with restricted VRAM or processing capacity. This versatility not only enhances its usability across different development frameworks but also positions it as a prime choice for diverse programming needs and settings. Furthermore, its ability to adapt to varying project demands makes it a valuable asset for developers seeking efficiency and performance in their workflows. -
28
scribe
scribe
Transform your development knowledge into a curated, accessible wiki.Scribe functions as an automated knowledge base on a self-hosted platform, leveraging your tools for content creation. It meticulously analyzes Git history, Claude Code and Codex sessions, as well as self-sent URLs and drop files, converting this data into a well-structured, cross-project wiki formatted in Markdown and preserved within Git repositories. By alleviating the necessity for developers to maintain an extra cognitive framework or to reconstruct context at the start of each agent session, Scribe adeptly captures essential decisions, solutions, assessments, and the reasoning behind them, allowing agents to access this accumulated knowledge before taking any actions. Its operational pipeline is orchestrated through cron jobs, which enables it to recognize projects and filter out low-value data using FTS5 before interacting with an LLM, thereby extracting validated facts through structured bounded workflows executed in two stages. The resulting output is organized into entity-centric pages that include YAML frontmatter, wikilinks, backlinks, retrieval context, and a variety of relationship types, such as supersedes, contradicts, derived_from, specializes, and extends, which ensures thorough documentation and clarity across various projects. This organized methodology not only fosters enhanced collaboration and knowledge sharing among teams but also streamlines the entire development process, ultimately leading to more efficient project management. As a result, teams can work more cohesively, drawing on a shared pool of insights and decisions that improve overall productivity. -
29
Qwen3.8-2.4T-A95B
Alibaba
Unleashing unparalleled capabilities for complex, multi-step tasks.Qwen3.8-2.4T-A95B emerges as the largest open model in the Qwen3.8 series, presenting advanced Qwen-Max-class capabilities in a format that is accessible to the public. Built on the robust foundation of Qwen3.5, this model offers marked improvements in performance across various domains, including coding, professional applications, research, and complex, extended agentic tasks, underscoring its ability to reliably execute intricate, multi-step workflows to completion. With its innovative mixture-of-experts architecture, it features a remarkable total of 2.4 trillion parameters, of which 95 billion are activated, utilizing 512 experts and allowing for simultaneous engagement of 10 routed experts alongside one shared expert. The model supports a native context length of 262,144 tokens, extendable to about 1.01 million tokens, thereby enabling considerable adaptability for diverse applications. Additionally, enhancements in agent execution, such as superior autonomous planning and improved responsiveness to environmental cues, enhance its overall efficiency. Its extensive compatibility with popular agent frameworks and development tools further aids in smooth integration into current systems, making it an appealing option for both developers and researchers. This versatility is particularly beneficial for those seeking to leverage advanced AI capabilities in their projects. -
30
NVIDIA Personal AI Router (PAIR)
NVIDIA
"Seamlessly unite your systems for efficient local AI."The NVIDIA Personal AI Router (PAIR) acts as a bridge for Windows, Linux, and macOS systems, creating a personal AI inference cluster and effectively managing AI application and agent workloads via a single local endpoint. This groundbreaking device allows RTX, DGX Spark, and Mac systems already linked to the same network to operate in unison as a local AI cluster, without requiring specialized cables, racks, or intricate setup processes. PAIR adeptly detects compatible machines and distributes inference requests among the available nodes, enabling intensive AI workflows to harness unused computing power across different operating systems. It integrates smoothly with popular local inference backends, such as Ollama and LM Studio, ensuring applications have access to a consistent endpoint while intelligently directing requests to the appropriate local computational resources. Tailored specifically for private local inference, PAIR guarantees that prompts, files, and agent contexts remain securely within the user’s local network, thereby negating the need to transfer data to cloud-based inference services. Additionally, this method not only bolsters data privacy but also maximizes resource utilization across the various systems engaged in AI processes, leading to improved efficiency and performance in workload management. Ultimately, PAIR represents a significant advancement in personal AI infrastructure, allowing users to leverage their existing hardware for enhanced AI capabilities. -
31
Qwen3.8-Omni-Flash
Alibaba
Empower your productivity with advanced multimodal capabilities today!Qwen3.8-Omni-Flash is a groundbreaking omnimodal model designed to significantly boost the efficiency of agents operating in productivity-focused settings, transitioning from basic understanding of multimodal inputs to actively performing tasks, utilizing diverse tools, and engaging in creative projects. Built upon the sophisticated Qwen3.8-Flash-Next architecture, it adeptly handles text, images, audio, and video inputs with an extraordinary context window of up to 1 million tokens, while maintaining strong performance in text-centric applications. This model transcends traditional coding and knowledge-based tasks, enriching workflows related to audio and video through capabilities such as video editing, crafting music videos, providing film commentary, summarizing audiovisual content, and facilitating real-time discussions. It particularly excels at enhancing the interpretation of long-form audio and video through organized descriptions, enabling agents to gather compelling evidence, grasp meeting content, and conduct thorough research focused on video materials. Users are empowered to specify parameters including subject matter, time frame, level of detail, and output format for video assessments, allowing for comprehensive overviews and customized analyses. This adaptability positions it as an indispensable resource for both professionals and creatives eager to optimize their productivity across a variety of multimedia platforms, ensuring that every project reaches its full potential. Furthermore, the model's seamless integration into diverse workflows opens up new possibilities for collaboration and innovation in content creation. -
32
WordRaptor
Curtis Duggan Software
Unlock lifetime SEO mastery with effortless content creation!Presenting WordRaptor, the all-in-one SEO tool that you invest in once for lifetime access. It handles everything from keyword generation and content planning to bulk drafting, image management, and both manual and automated publishing, ensuring a complete solution for your diverse needs. With WordRaptor, you gain back control and privacy right on your Mac, eliminating the ongoing expenses associated with subscription-based AI writing services. Effortlessly produce content from titles, keywords, or descriptions while offering specific directions for tailored results. It automatically creates vital meta titles, descriptions, and open graph metadata to boost your digital footprint. You can efficiently manage extensive content volumes through our streamlined article queue system and choose your preferred AI model by integrating your own API key for enhanced adaptability. Your sensitive data remains securely stored on your Mac, enabling you to publish articles on various platforms such as Wix, WordPress, Ghost, Shopify, and Webflow without hassle. Additionally, with an interface designed for ease of use, even individuals unfamiliar with SEO can confidently navigate the software, making it an ideal choice for everyone. -
33
Qwen 4
Alibaba
Unleashing the future of AI with unparalleled intelligence.Qwen 4 is Alibaba’s forthcoming next-generation foundation model and the planned successor to the company’s Qwen3.x model family. Alibaba announced Qwen 4 at the 2026 Apsara Conference on September 22 and confirmed that the model is currently in training. The company has not yet disclosed Qwen 4’s architecture, parameter count, context length, training-compute requirements, benchmark scores, pricing, licensing terms, or release schedule. Qwen 4 is being developed as Alibaba expands its broader AI stack across foundation models, multimodal systems, AI infrastructure, and agent-oriented cloud services. A major research direction surrounding Alibaba’s next generation of models is recursive self-improvement based on real-world tasks and empirical feedback. The company has already experimented with this approach using Qwen3.8-Max, allowing the model to participate in automated pipeline design, data validation, experimentation, error diagnosis, and post-training optimization. Alibaba reported that Qwen3.8-Max completed 33 iterative cycles during one such experiment and increased its Artificial Analysis score from 40 to 45. In a separate chip-design experiment, a Qwen model performed more than 10,000 EDA tool calls during over 60 hours of automated improvement work, illustrating Alibaba’s interest in long-horizon agentic tasks. These demonstrations describe the research program surrounding future Qwen development rather than confirmed features of Qwen 4 itself. Alibaba has also announced a longer-term roadmap in which Qwen 4.5 and Qwen 5 models are projected to reach between 5 trillion and 10 trillion parameters. Qwen 4 therefore remains a pre-release model, with detailed capabilities and access information expected to become clearer when Alibaba publishes its formal launch materials. -
34
Llama
Meta
Empowering researchers with inclusive, efficient AI language models.Llama, a leading-edge foundational large language model developed by Meta AI, is designed to assist researchers in expanding the frontiers of artificial intelligence research. By offering streamlined yet powerful models like Llama, even those with limited resources can access advanced tools, thereby enhancing inclusivity in this fast-paced and ever-evolving field. The development of more compact foundational models, such as Llama, proves beneficial in the realm of large language models since they require considerably less computational power and resources, which allows for the exploration of novel approaches, validation of existing studies, and examination of potential new applications. These models harness vast amounts of unlabeled data, rendering them particularly effective for fine-tuning across diverse tasks. We are introducing Llama in various sizes, including 7B, 13B, 33B, and 65B parameters, each supported by a comprehensive model card that details our development methodology while maintaining our dedication to Responsible AI practices. By providing these resources, we seek to empower a wider array of researchers to actively participate in and drive forward the developments in the field of AI. Ultimately, our goal is to foster an environment where innovation thrives and collaboration flourishes. -
35
Surf.new
Steel.dev
Explore AI agents effortlessly, enhancing productivity and creativity.Surf.new is an innovative, free, and open-source platform created for the exploration of AI agents capable of navigating the internet. These agents replicate human-like browsing and interactions with websites, making tasks like automation and online research more efficient. This platform serves a dual purpose: it is perfect for developers looking to evaluate web agents for future use, as well as for everyday users aiming to simplify repetitive tasks such as tracking flight prices, collecting product information, or booking reservations. Surf.new provides an accessible environment where users can test and assess the efficacy of these web agents effortlessly. Noteworthy Features: Seamless AI Agent Framework Switching: Users can easily switch between numerous frameworks with a single click, including options for browser use, an experimental Claude Computer-use-based agent, and smooth integration with LangChain, promoting a variety of experimentation approaches. Extensive AI Model Compatibility: The platform supports a wide array of well-known models, including Claude 3.7, DeepSeek R1, OpenAI models, and Gemini 2.0 Flash, allowing users to choose the most fitting model for their specific requirements. Moreover, the intuitive interface of Surf.new fosters creativity and exploration, making it a prime choice for those eager to delve into the potential of AI-driven web agents while enhancing their own productivity. By encouraging users to engage with various tools, Surf.new not only simplifies tasks but also inspires innovative solutions.