-
1
GPT-5.6 Sol
OpenAI
Unleash advanced reasoning and accelerate your complex workflows.
GPT-5.6 Sol is a next-generation OpenAI model previewed as the flagship option in the GPT-5.6 family. The series includes Sol for the strongest capability, Terra for balanced everyday work, and Luna for faster, lower-cost use cases. GPT-5.6 Sol is built for demanding work across coding, agentic automation, biology, cybersecurity, research, and enterprise knowledge workflows. The model introduces a new max reasoning effort that allows it to spend more time reasoning through difficult problems. It also adds ultra mode, which coordinates subagents to help accelerate complex tasks that benefit from parallel or multi-agent execution. In coding workflows, GPT-5.6 Sol is designed for command-line tasks that require planning, iteration, testing, tool coordination, and long-horizon software engineering judgment. In biology workflows, it is positioned for genomics and quantitative-biology analysis where efficient reasoning over complex scientific tasks matters. In cybersecurity, GPT-5.6 Sol supports legitimate defensive work such as vulnerability discovery, patch development, debugging, security education, code review, and authorized testing. OpenAI describes GPT-5.6 Sol as more capable at helping users find and fix vulnerabilities than reliably carrying out end-to-end attacks under tested conditions. The model’s release is paired with a layered safeguard system that includes model-level refusals, real-time misuse classifiers, paused generation for higher-risk cases, account-level review, automated red-teaming, third-party testing, differentiated access, and enterprise safety controls. GPT-5.6 Sol helps developers, researchers, enterprises, and cyber defenders use frontier AI for advanced technical work while supporting safer deployment, stronger oversight, and phased access.
-
2
GPT-6 Astra
OpenAI
Revolutionizing professional workflows with advanced AI capabilities.
GPT-6 Astra is OpenAI’s advanced frontier model for computer use, coding, browsing, scientific research, cybersecurity, professional knowledge work, and long-running agentic tasks. It is designed to combine high-level reasoning with the ability to directly operate software and tools rather than only generating text responses. Astra can navigate websites, complete forms, update business systems, organize calendars, conduct research, analyze data, generate plots, test applications, and troubleshoot problems that appear on screen. For professional users, the model can create documents, spreadsheets, presentations, analyses, websites, and other artifacts while following existing templates, formatting requirements, and organizational styles. Its software engineering capabilities include codebase analysis, implementation, debugging, verification, browser testing, system configuration, and other terminal-based development workflows. In Codex, Astra can preserve notes across context windows and search earlier requirements, test results, messages, and tool outputs during lengthy development sessions. The model also combines scientific reasoning with computer use so researchers can work with specialized applications, inspect data, explore results, and assist with computational research processes. OpenAI reports substantial advances in Astra’s cybersecurity capabilities, while the production model applies safeguards to restrict higher-risk activities such as advanced exploit creation. Alignment improvements focus on interpreting user intent, respecting authorization boundaries, avoiding attempts to circumvent system restrictions, and communicating more accurately about what the model can and cannot do. GPT-6 Astra supports enterprise-oriented deployment features including eligible Zero Data Retention API configurations and is available through ChatGPT, the OpenAI API, Amazon Web Services, and Amazon Bedrock.
-
3
GPT-6 Sol
OpenAI
Unlock professional potential with streamlined, intelligent collaboration tools.
GPT-6 Sol is an advanced OpenAI model positioned between the cost-efficient GPT-6 Luna and the higher-capability GPT-6 Astra for demanding professional and agentic workloads. The model is designed for coding, knowledge work, business automation, computer use, research, and other tasks that require sustained reasoning across multiple steps. It inherits advances from the GPT-6 generation while emphasizing a balance of intelligence, speed, and operating cost for applications that need to run at scale. GPT-6 Sol supports multiple reasoning-effort levels so applications can spend more computation on difficult tasks and reduce effort for straightforward requests. In software development, it can handle complex real-codebase tasks, generate merge-ready changes, debug software, work through terminal workflows, and operate as part of coding agents. Its professional-work capabilities support multi-application processes spanning functions such as finance, operations, sales, marketing, customer support, and human resources. Computer-use abilities allow agents powered by GPT-6 Sol to interact with graphical interfaces and complete long-horizon workflows involving everyday and professional software. OpenAI has also improved the model’s factual reliability, communication style, and alignment compared with GPT-5.6 Sol, including lower rates of misleading claims in challenging coding evaluations. GPT-6 prompt caching provides higher cache-hit rates, supports changing reasoning effort or available tools without invalidating earlier cached context, and offers substantial discounts for cached input tokens. Developers can monitor caching behavior, configure prompt-cache breakpoints, and incorporate Sol into persistent agents that repeatedly reuse large amounts of context. GPT-6 Sol is accessible through ChatGPT Work, Codex, and the OpenAI API under the gpt-6-sol model identifier.
-
4
GPT-6 Luna
OpenAI
Maximize efficiency with advanced, cost-effective AI solutions.
GPT-6 Luna is OpenAI’s efficiency-focused GPT-6 model for developers and users who need capable reasoning, coding, computer use, and agentic workflows at very low inference cost. It is positioned below GPT-6 Sol and GPT-6 Astra in the model family while bringing many of the GPT-6 generation’s improvements to applications that prioritize scale and affordability. The model supports configurable reasoning effort so developers can allocate additional computation to complex tasks while keeping simpler interactions fast and economical. GPT-6 Luna can power business automation across applications used for sales, marketing, finance, operations, customer support, and human resources. Its coding capabilities support work on real software repositories, including multi-step engineering tasks that require analysis, modification, testing, and iteration. Luna can also operate in computer-use environments, allowing agents to navigate graphical interfaces and complete extended workflows across software applications. OpenAI reports that GPT-6 Luna substantially improves factual reliability compared with GPT-5.6 Luna and can approach the capabilities of more expensive models on some tasks when used at higher reasoning levels. The model also benefits from GPT-6’s improved collaboration style, with clearer technical communication, less unnecessary jargon, and fewer low-value details. Enhanced prompt caching allows applications to reuse previously processed context at a discount while preserving cache reuse when reasoning effort or available tools change. These efficiency improvements make Luna suitable for high-volume agents, coding assistants, automated workflows, customer-facing applications, and other systems where per-request cost is important. GPT-6 Luna is available through the OpenAI API as gpt-6-luna, as well as through ChatGPT Work, Codex, and supported ChatGPT desktop experiences.
-
5
GPT-5.6 Terra
OpenAI
Empowering your workflows with balanced intelligence, speed, affordability.
GPT-5.6 Terra is a balanced model in OpenAI’s GPT-5.6 series, designed to provide strong performance for everyday work while keeping costs lower than the flagship Sol tier. The GPT-5.6 family includes Sol for the highest capability, Terra for balanced work, and Luna for fast and affordable use cases. Terra is positioned as a practical option for developers, businesses, and enterprise teams that need capable reasoning, coding, automation, research support, and defensive security assistance without always using the most expensive model. According to the pasted preview text, Terra offers competitive performance to GPT-5.5 while being 2x cheaper. It appears in GPT-5.6 benchmark previews for Terminal-Bench 2.1, GeneBench v1, ExploitBench, and ExploitGym, showing that the model is intended for technical and long-horizon tasks as well as general work. Terra can support coding workflows that require planning, iteration, command-line reasoning, and tool coordination. It can also support legitimate cybersecurity workflows such as code review, vulnerability research, patch development, debugging, security education, and defensive testing. The model is developed with layered safeguards matched to its capabilities, including trained refusals, real-time checks, misuse classifiers, monitoring, enforcement, and account-level review. OpenAI also describes automated red-teaming and third-party human expert red-teaming as part of the broader GPT-5.6 safety process. Terra is priced below Sol in the pasted API pricing structure, with lower input and output costs per 1 million tokens. GPT-5.6 Terra helps organizations use a capable GPT-5.6 model for production workflows where performance, cost efficiency, and safety controls all matter.
-
6
GPT-5.6 Luna
OpenAI
Fast, affordable AI intelligence for practical user needs.
GPT-5.6 Luna is the lowest-cost model in OpenAI’s GPT-5.6 family, built for fast and affordable AI assistance across everyday and technical workflows. The GPT-5.6 lineup includes Sol as the flagship model, Terra as the balanced model for everyday work, and Luna as the efficient model for users who need strong capability at lower cost. Luna is intended for developers, businesses, and teams that need scalable AI for coding help, workflow automation, research support, analysis, customer-facing applications, and high-volume API usage. In the pasted preview text, Luna is presented as part of the same GPT-5.6 release process and benchmark set as Sol and Terra. It appears in evaluations for command-line coding workflows, long-horizon biology tasks, ExploitBench, and ExploitGym, indicating that it is designed to handle more than simple chat use cases. The model is priced at a lower per-token rate than Sol and Terra, making it more suitable for applications where cost efficiency is a major priority. GPT-5.6 Luna also supports the new GPT-5.6 prompt caching approach, including explicit cache breakpoints, a 30-minute minimum cache life, cache writes billed above the uncached input rate, and discounted cached-input reads. Like the rest of the GPT-5.6 family, Luna is developed with layered safeguards matched to model capability. These safeguards include trained refusals for prohibited cyber assistance, real-time misuse classifiers, paused generation for higher-risk cases, account-level review, monitoring, enforcement, automated red-teaming, and third-party human expert red-teaming. Luna is expected to support legitimate defensive and technical workflows such as code review, debugging, patch development, security education, and defensive testing while making prohibited misuse more difficult and detectable. GPT-5.6 Luna helps organizations deploy GPT-5.6-class AI where speed, affordability, scalability, and safe production use are the most important requirements.
-
7
GPT-5.5
OpenAI
Transform your ideas into execution with unmatched efficiency.
GPT-5.5 represents a new class of AI built to transform how work is done across digital environments. It combines advanced reasoning, tool usage, and task execution capabilities to manage complex, multi-step workflows with minimal human intervention. The model performs strongly in software engineering, data analysis, business operations, and scientific research, where it can plan tasks, gather information, test solutions, and refine outputs iteratively. It supports generating documents, building applications, analyzing large datasets, and navigating software systems as part of a unified workflow. A key capability is its integration with workspace agents—customizable AI agents that can be created once and deployed across teams to automate entire processes. These agents can run continuously, interact with tools like CRM systems, messaging platforms, and document editors, and keep workflows moving without constant supervision. Organizations can define permissions, approval checkpoints, and monitoring to maintain full control over automation. GPT-5.5 also improves collaboration by standardizing workflows and scaling best practices across teams. With enterprise-grade security and governance, it is designed for safe deployment in complex environments. Its ability to persist through ambiguity and long-running tasks makes it highly effective for execution-heavy work. By reducing manual intervention and increasing speed, GPT-5.5 enables teams to focus on higher-value activities and operate at a significantly higher level of productivity.
-
8
ChatGPT
OpenAI
Unlock your potential with efficient, AI-powered assistance today!
ChatGPT is an advanced AI-powered assistant designed to help users accomplish tasks, generate ideas, and improve productivity across a wide range of use cases. It enables users to perform activities such as writing, editing, coding, research, and brainstorming with ease. The platform supports both text and voice interactions, allowing users to communicate in the way that suits them best. ChatGPT can summarize meetings, analyze data, and provide actionable insights to support better decision-making. It also assists with creative tasks, including content creation, marketing strategies, and personal planning. One of its most powerful capabilities is workspace agents, which allow users to build automated systems that handle entire workflows. These agents can operate across different tools, gather information, and take actions such as updating documents, sending communications, or managing tasks without constant supervision. They can be scheduled to run recurring processes, ensuring work continues even when teams are not actively involved. Workspace agents can be shared across teams, helping organizations standardize workflows and scale best practices efficiently. Built-in governance features, such as permissions, approval checkpoints, and monitoring, ensure secure and controlled automation. ChatGPT integrates seamlessly into existing workflows, reducing the need for multiple tools and manual coordination. It supports collaboration by allowing teams to refine, edit, and manage work in real time. The platform adapts to various industries and use cases, from personal productivity to enterprise operations. By combining intelligent assistance with automation, ChatGPT enables users to focus on higher-impact work. Ultimately, it acts as a comprehensive solution for both everyday tasks and complex organizational workflows.
-
9
GPT-3
OpenAI
Unleashing powerful language models for diverse, effective communication.
Our models are crafted to understand and generate natural language effectively. We offer four main models, each designed with different complexities and speeds to meet a variety of needs. Among these options, Davinci emerges as the most robust, while Ada is known for its remarkable speed. The principal GPT-3 models are mainly focused on the text completion endpoint, yet we also provide specific models that are fine-tuned for other endpoints. Not only is Davinci the most advanced in its lineup, but it also performs tasks with minimal direction compared to its counterparts. For tasks that require a nuanced understanding of content, like customized summarization and creative writing, Davinci reliably produces outstanding results. Nevertheless, its superior capabilities come at the cost of requiring more computational power, which leads to higher expenses per API call and slower response times when compared to other models. Consequently, the choice of model should align with the particular demands of the task in question, ensuring optimal performance for the user's needs. Ultimately, understanding the strengths and limitations of each model is essential for achieving the best results.
-
10
GPT-4
OpenAI
Revolutionizing language understanding with unparalleled AI capabilities.
The fourth iteration of the Generative Pre-trained Transformer, known as GPT-4, is an advanced language model expected to be launched by OpenAI. As the next generation following GPT-3, it is part of the series of models designed for natural language processing and has been built on an extensive dataset of 45TB of text, allowing it to produce and understand language in a way that closely resembles human interaction. Unlike traditional natural language processing models, GPT-4 does not require additional training on specific datasets for particular tasks. It generates responses and creates context solely based on its internal mechanisms. This remarkable capacity enables GPT-4 to perform a wide range of functions, including translation, summarization, answering questions, sentiment analysis, and more, all without the need for specialized training for each task. The model’s ability to handle such a variety of applications underscores its significant potential to influence advancements in artificial intelligence and natural language processing fields. Furthermore, as it continues to evolve, GPT-4 may pave the way for even more sophisticated applications in the future.
-
11
GPT-4 Turbo
OpenAI
Revolutionary AI model redefining text and image interaction.
The GPT-4 model signifies a remarkable leap in artificial intelligence, functioning as a large multimodal system adept at processing both text and image inputs, while generating text outputs that enable it to address intricate problems with an accuracy that surpasses previous iterations due to its vast general knowledge and superior reasoning abilities. Available through the OpenAI API for subscribers, GPT-4 is tailored for chat-based interactions, akin to gpt-3.5-turbo, and excels in traditional completion tasks via the Chat Completions API. This cutting-edge version of GPT-4 features advancements such as enhanced instruction compliance, a JSON mode, reliable output consistency, and the capability to execute functions in parallel, rendering it an invaluable resource for developers. It is crucial to understand, however, that this preview version is not entirely equipped for high-volume production environments, having a constraint of 4,096 output tokens. Users are invited to delve into its functionalities while remaining aware of its existing restrictions, which may affect their overall experience. The ongoing updates and potential future enhancements promise to further elevate its performance and usability.
-
12
GPT-4o
OpenAI
Revolutionizing interactions with swift, multi-modal communication capabilities.
GPT-4o, with the "o" symbolizing "omni," marks a notable leap forward in human-computer interaction by supporting a variety of input types, including text, audio, images, and video, and generating outputs in these same formats. It boasts the ability to swiftly process audio inputs, achieving response times as quick as 232 milliseconds, with an average of 320 milliseconds, closely mirroring the natural flow of human conversations. In terms of overall performance, it retains the effectiveness of GPT-4 Turbo for English text and programming tasks, while significantly improving its proficiency in processing text in other languages, all while functioning at a much quicker rate and at a cost that is 50% less through the API. Moreover, GPT-4o demonstrates exceptional skills in understanding both visual and auditory data, outpacing the abilities of earlier models and establishing itself as a formidable asset for multi-modal interactions. This groundbreaking model not only enhances communication efficiency but also expands the potential for diverse applications across various industries. As technology continues to evolve, the implications of such advancements could reshape the future of user interaction in multifaceted ways.
-
13
GPT-4.5
OpenAI
Revolutionizing AI with enhanced learning, reasoning, and collaboration.
GPT-4.5 marks a substantial leap forward in artificial intelligence, building upon its predecessors by enhancing its unsupervised learning methods, honing its reasoning capabilities, and improving its collaborative functionalities. Designed to better interpret human intentions, this model enables more fluid and instinctive interactions, leading to increased precision and fewer instances of misinformation across a wide range of topics. Its advanced capabilities not only foster the generation of creative and intellectually stimulating content but also aid in tackling complex problems while offering assistance in various domains such as writing, design, and even aerospace endeavors. In addition, the model's improved human engagement opens doors for practical applications, making it more user-friendly and reliable for both businesses and developers. As it continues to innovate, GPT-4.5 establishes a new benchmark for the role of AI in numerous sectors and applications, demonstrating its potential to transform how we approach technology in everyday life. The ongoing developments in this field suggest a promising future where AI can seamlessly integrate into our daily routines and professional tasks.
-
14
GPT-6.1 Sol
OpenAI
Unlock professional potential with streamlined, intelligent collaboration tools.
GPT-6.1 Sol is OpenAI's upgraded Sol model for developers and professionals who need advanced reasoning and agentic capabilities without the higher cost of GPT-6 Astra. It is designed for coding, professional knowledge work, computer use, scientific research, factual question answering, and multi-step business workflows. OpenAI describes GPT-6.1 Sol as approaching GPT-6 Astra's intelligence across several important workloads while charging one-fifth of Astra's standard input and output token prices. On DeepSWE v1.1, which evaluates long-horizon software engineering in real codebases, GPT-6.1 Sol matches GPT-6 Astra at approximately one-fifth of the cost and exceeds GPT-6 Sol's best score by 6.4 percentage points. Its professional-work capabilities include understanding complex PDFs containing tables, charts, diagrams, and fine-print details across fields such as finance, healthcare, and legal work. On AutomationBench, GPT-6.1 Sol improves on GPT-6 Sol by 4.8 percentage points at the same reasoning setting and scores 2.2 points above Opus 5.5 at medium reasoning effort. Computer-use performance also advances significantly, with GPT-6.1 Sol outperforming GPT-6 Sol by seven percentage points on the OSWorld 2.0 offline set at maximum reasoning effort and coming within 2.1 points of GPT-6 Astra. For scientific research, the model can work with code and terminal tools on workflows involving data analysis, simulations, model fitting, and theorem proving, more than doubling GPT-6 Sol's Terminal-Bench Science 0.1 score at maximum effort. OpenAI also reports improved factual accuracy, including a reduction in the factual-error rate from 11.4% with GPT-6 Sol to 7.7% with GPT-6.1 Sol at low reasoning effort on its deliberately difficult factuality evaluation.
-
15
GPT-4o mini
OpenAI
Streamlined, efficient AI for text and visual mastery.
A streamlined model that excels in both text comprehension and multimodal reasoning abilities.
The GPT-4o mini has been crafted to efficiently manage a vast range of tasks, characterized by its affordability and quick response times, which make it particularly suitable for scenarios requiring the simultaneous execution of multiple model calls, such as activating various APIs at once, analyzing large sets of information like complete codebases or lengthy conversation histories, and delivering prompt, real-time text interactions for customer support chatbots. At present, the API for GPT-4o mini supports both textual and visual inputs, with future enhancements planned to incorporate support for text, images, videos, and audio. This model features an impressive context window of 128K tokens and can produce outputs of up to 16K tokens per request, all while maintaining a knowledge base that is updated to October 2023. Furthermore, the advanced tokenizer utilized in GPT-4o enhances its efficiency in handling non-English text, thus expanding its applicability across a wider range of uses. Consequently, the GPT-4o mini is recognized as an adaptable resource for developers and enterprises, making it a valuable asset in various technological endeavors. Its flexibility and efficiency position it as a leader in the evolving landscape of AI-driven solutions.
-
16
OpenAI o1-pro
OpenAI
Unleash advanced problem-solving with unparalleled speed and accuracy.
The o1-pro from OpenAI is a more sophisticated version of the original o1 model, designed to tackle complex and demanding challenges with greater reliability. This enhanced model exhibits significant improvements over the prior o1 preview, achieving an impressive 34% reduction in critical errors and a 50% boost in processing speed. It excels in areas such as mathematics, physics, and programming, providing detailed and accurate solutions. Additionally, the o1-pro can handle multimodal inputs, including both text and images, and demonstrates exceptional skills in complex reasoning tasks that require deep analytical thinking. It is accessible through a ChatGPT Pro subscription, granting users not just unlimited access, but also enhanced functionalities for those in need of advanced AI assistance. With these capabilities, users are empowered to efficiently and effectively tackle a broader array of challenges, making the o1-pro an invaluable tool for problem-solving. Overall, the advancements in this model signify a leap forward in AI technology, offering new possibilities for various applications.
-
17
OpenAI o1
OpenAI
Revolutionizing problem-solving with advanced reasoning and cognitive engagement.
OpenAI has unveiled the o1 series, which heralds a new era of AI models tailored to improve reasoning abilities. This series includes models such as o1-preview and o1-mini, which implement a cutting-edge reinforcement learning strategy that prompts them to invest additional time "thinking" through various challenges prior to providing answers. This approach allows the o1 models to excel in complex problem-solving environments, especially in disciplines like coding, mathematics, and science, where they have demonstrated superiority over previous iterations like GPT-4o in certain benchmarks. The purpose of the o1 series is to tackle issues that require deeper cognitive engagement, marking a significant step forward in developing AI systems that can reason more like humans do. Currently, the series is still in the process of refinement and evaluation, showcasing OpenAI's dedication to the ongoing enhancement of these technologies. As the o1 models evolve, they underscore the promising trajectory of AI, illustrating its capacity to adapt and fulfill increasingly sophisticated requirements in the future. This ongoing innovation signifies a commitment not only to technological advancement but also to addressing real-world challenges with more effective AI solutions.
-
18
OpenAI o1-mini
OpenAI
Affordable AI powerhouse for STEM problems and coding!
The o1-mini, developed by OpenAI, represents a cost-effective innovation in AI, focusing on enhanced reasoning skills particularly in STEM fields like math and programming. As part of the o1 series, this model is designed to address complex problems by spending more time on analysis and thoughtful solution development. Despite being smaller and priced at 80% less than the o1-preview model, the o1-mini proves to be quite powerful in handling coding tasks and mathematical reasoning. This effectiveness makes it a desirable option for both developers and businesses looking for dependable AI solutions. Additionally, its economical price point ensures that a broader audience can access and leverage advanced AI technology without sacrificing quality. Overall, the o1-mini stands out as a remarkable tool for those needing efficient support in technical areas.
-
19
OpenAI has developed a sophisticated research tool that leverages artificial intelligence to autonomously perform complex, multi-faceted research tasks across various domains, such as science, programming, and mathematics. By interpreting user inputs—which may include questions, documents, images, PDFs, or spreadsheets—the tool formulates a comprehensive research plan, gathers relevant data, and delivers detailed responses within minutes. Furthermore, it provides summaries of the research workflow along with citations, allowing users to verify the origins of the information presented. While this tool significantly boosts research productivity, it is not without its flaws, as it can occasionally produce inaccuracies or struggle to differentiate between reliable sources and misinformation. Currently, it is available to users of ChatGPT Pro, representing a major leap forward in AI-driven knowledge discovery, and ongoing improvements aim to enhance both the accuracy and speed of responses. This continuous evolution highlights a dedication to perfecting the tool's functionalities and ensuring that users access the most trustworthy information possible, paving the way for more informed decision-making in research practices.
-
20
GPT-5.1 Pro
OpenAI
Unleash advanced reasoning for complex problem-solving excellence.
GPT-5.1 Pro represents the top tier of OpenAI’s GPT-5 generation, delivering the most advanced reasoning, depth, and analytical intelligence available in ChatGPT. It is optimized for high-stakes, high-complexity scenarios where rigorous logic and verifiable accuracy are essential. Professionals use GPT-5.1 Pro for scientific research, large-scale codebases, legal reasoning, quantitative finance, data analysis, and multi-step decision workflows that exceed the capabilities of general models. With a significantly expanded context window, GPT-5.1 Pro can ingest and analyze long documents, datasets, transcripts, and multi-file projects in a single session. The model’s reasoning engine is tuned for deeper internal deliberation, enabling structured explanations, defensible conclusions, and clearer thought processes. GPT-5.1 Pro also features enhanced adherence to instructions, producing responses that are more predictable, consistent, and aligned with user goals. Compared to Instant and Thinking modes, it is built for reliability rather than speed, prioritizing quality of reasoning over quick output. While it supports most ChatGPT tools, it is intentionally restricted from Canvas and image generation to preserve dedicated compute for reasoning-heavy tasks. GPT-5.1 Pro is exclusive to ChatGPT Pro and Business subscribers, offering unlimited access within standard safety guardrails. It is the model tier best suited for users who depend on ChatGPT as a trusted research partner and analytical assistant.
-
21
GPT-5 mini
OpenAI
Streamlined AI for fast, precise, and cost-effective tasks.
GPT-5 mini is a faster, more affordable variant of OpenAI’s advanced GPT-5 language model, specifically tailored for well-defined and precise tasks that benefit from high reasoning ability. It accepts both text and image inputs (image input only), and generates high-quality text outputs, supported by a large 400,000-token context window and a maximum of 128,000 tokens in output, enabling complex multi-step reasoning and detailed responses. The model excels in providing rapid response times, making it ideal for use cases where speed and efficiency are critical, such as chatbots, customer service, or real-time analytics. GPT-5 mini’s pricing structure significantly reduces costs, with input tokens priced at $0.25 per million and output tokens at $2 per million, offering a more economical option compared to the flagship GPT-5. While it supports advanced features like streaming, function calling, structured output generation, and fine-tuning, it does not currently support audio input or image generation capabilities. GPT-5 mini integrates seamlessly with multiple API endpoints including chat completions, responses, embeddings, and batch processing, providing versatility for a wide array of applications. Rate limits are tier-based, scaling from 500 requests per minute up to 30,000 per minute for higher tiers, accommodating small to large scale deployments. The model also supports snapshots to lock in performance and behavior, ensuring consistency across applications. GPT-5 mini is ideal for developers and businesses seeking a cost-effective solution with high reasoning power and fast throughput. It balances cutting-edge AI capabilities with efficiency, making it a practical choice for applications demanding speed, precision, and scalability.
-
22
GPT-5 nano
OpenAI
Lightning-fast, budget-friendly AI for text and images!
GPT-5 nano is OpenAI’s fastest and most cost-efficient version of the GPT-5 model, engineered to handle high-speed text and image input processing for tasks such as summarization, classification, and content generation. It features an extensive 400,000-token context window and can output up to 128,000 tokens, allowing for complex, multi-step language understanding despite its focus on speed. With ultra-low pricing—$0.05 per million input tokens and $0.40 per million output tokens—GPT-5 nano makes advanced AI accessible to budget-conscious users and developers working at scale. The model supports a variety of advanced API features, including streaming output, function calling for interactive applications, structured outputs for precise control, and fine-tuning for customization. While it lacks support for audio input and web search, GPT-5 nano supports image input, code interpretation, and file search, broadening its utility. Developers benefit from tiered rate limits that scale from 500 to 30,000 requests per minute and up to 180 million tokens per minute, supporting everything from small projects to enterprise workloads. The model also offers snapshots to lock performance and behavior, ensuring consistent results over time. GPT-5 nano strikes a practical balance between speed, cost, and capability, making it ideal for fast, efficient AI implementations where rapid turnaround and budget are critical. It fits well for applications requiring real-time summarization, classification, chatbots, or lightweight natural language processing tasks. Overall, GPT-5 nano expands the accessibility of OpenAI’s powerful AI technology to a broader user base.
-
23
GPT-5.1-Codex
OpenAI
Elevate coding efficiency with intelligent, adaptive software solutions.
GPT-5.1-Codex represents a sophisticated evolution of the GPT-5.1 framework, tailored specifically for coding and software development tasks that necessitate a degree of independence. This model shines in interactive programming scenarios as well as in the sustained execution of complex engineering endeavors, encompassing activities such as building applications from scratch, improving functionalities, debugging, performing comprehensive code refactoring, and conducting code reviews. It adeptly harnesses a variety of tools while merging seamlessly into development environments, modulating its reasoning skills according to the complexity of the tasks at hand; it swiftly resolves straightforward issues while allocating additional resources to more complex challenges. Users have noted that GPT-5.1-Codex consistently produces cleaner and higher-quality code compared to its general-purpose alternatives, demonstrating a better alignment with developer needs and a significant decrease in errors. Moreover, access to the model is provided via the Responses API rather than the typical chat API, and it includes distinct configurations such as a “mini” version for those on a budget and a “max” variant that offers the highest level of performance. This specialized iteration is designed not only to improve productivity but also to significantly enhance efficiency in software development processes, ultimately leading to a smoother workflow for engineers. Its adaptability and targeted features make it a valuable asset in the fast-evolving landscape of software engineering.
-
24
GPT-5.5 Pro
OpenAI
Transform your workflow with a an intelligent, efficient AI model
GPT-5.5 Pro represents a new class of AI designed to transform how work gets done across digital environments. It combines advanced reasoning, tool usage, and task execution capabilities to handle complex, multi-step workflows with minimal human intervention. The model excels in areas such as software engineering, data analysis, business operations, and scientific research, where it can plan tasks, gather information, test solutions, and refine outputs continuously. It supports creating applications, generating reports, building spreadsheets, and navigating software systems as part of a complete workflow. A key capability is its integration with workspace agents—custom AI agents that can be built once and deployed across teams to automate entire processes. These agents can run tasks on schedules, interact with tools like CRM systems, messaging platforms, and document editors, and keep workflows moving without constant supervision. Organizations can define permissions, approval checkpoints, and monitoring to maintain control over automated processes. GPT-5.5 Pro also enhances collaboration by enabling teams to standardize workflows and scale best practices across the organization. With enterprise-grade security and governance, it ensures safe deployment in complex environments. Its ability to persist through ambiguity and long tasks makes it highly effective for execution-heavy work. By reducing manual intervention and increasing speed, it allows teams to focus on higher-value activities. Ultimately, GPT-5.5 Pro enables businesses and professionals to operate at a significantly higher level of productivity and efficiency.
-
25
Holo4
H Company
Versatile AI models for seamless multi-platform task execution.
Holo4 is a family of agentic AI models developed by H Company for computer use and multi-step automation across desktop, web, mobile, terminal, MCP, and API environments. The series consists of Holo4 27B, a dense 27-billion-parameter model, and Holo4 35B-A3B, a 35-billion-parameter Mixture-of-Experts model with 3 billion active parameters. Rather than specializing exclusively in graphical interfaces or tool calling, Holo4 can click and type on screens, write and run its own code, and invoke MCP or API tools as different stages of a workflow require. The same model can therefore move between desktop applications, websites, Android applications, code sandboxes, and business APIs without switching to a separate model for each interface. H Company's Agentic Task Factory generated approximately 10,000 tasks across web applications, MCP servers, desktop software, and hybrid environments to support model development and evaluation. Holo4 underwent supervised fine-tuning on 127 billion tokens, with roughly three-quarters of that training data consisting of successful agentic trajectories covering desktop, web, MCP/API, and mobile tasks. Two reinforcement-learning experts were subsequently trained for desktop/web workflows and terminal/MCP/API workflows before being merged into the final generalist model. In H Company's evaluations, Holo4 27B scored 85.2% on OSWorld, 61.7% on OSWorld 2.0, 45.4% on AutomationBench, and 85.1% on AndroidWorld, although the company notes that reference-model results can use different harnesses and effort levels. Holo4 27B supports a 256K context window and is priced through the H Models API at $0.40 per million input tokens, $0.04 per million cached input tokens, and $3.00 per million output tokens. Holo4 35B-A3B also supports 256K context and is priced at $0.30 per million input tokens, $0.03 per million cached input tokens, and $2.00 per million output tokens.