开源
GPT-OSS-120B is an open-weight, 116.8B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized to run on a single H100 GPU with native MXFP4 quantization. The model supports configurable reasoning depth, full chain-of-thought access, and native tool use, including function calling, browsing, and structured output generation. It achieves near-parity with OpenAI o4-mini on core reasoning benchmarks. Note: While referred to as '120b' for simplicity, it technically has 116.8B parameters.
开源
The gpt-oss-20b model (technically 20.9B parameters) achieves near-parity with OpenAI o4-mini on core reasoning benchmarks, while running efficiently on a single 80 GB GPU. The gpt-oss-20b model delivers similar results to OpenAI o3‑mini on common benchmarks and can run on edge devices with just 16 GB of memory, making it ideal for on-device use cases, local inference, or rapid iteration without costly infrastructure. Both models also perform strongly on tool use, few-shot function calling, CoT reasoning (as seen in results on the Tau-Bench agentic evaluation suite) and HealthBench (even outperforming proprietary models like OpenAI o1 and GPT‑4o). Note: While referred to as '20b' for simplicity, it technically has 20.9B parameters.
开源
Safety model for policy screening, moderation, and risk-aware routing workflows
专有 多模态
Streaming speech-to-text model for low-latency transcript deltas from live audio
专有
The latest GPT-3.5 Turbo model with higher accuracy at responding in requested formats and a fix for a bug which caused a text encoding issue for non-English language function calls.
专有 多模态
GPT-4 is a large multimodal model capable of processing both image and text inputs and generating human-like text outputs. It demonstrates human-level performance on various professional and academic benchmarks.
专有
The latest GPT-4 model with improved performance, updated knowledge, and enhanced capabilities. It offers faster response times and more affordable pricing compared to previous versions.
专有 多模态
GPT-4.1 is OpenAI's latest and most advanced flagship model, significantly improving upon GPT-4 Turbo in performance across benchmarks, speed, and cost-effectiveness.
专有 多模态
GPT-4.1 mini provides a balance between intelligence, speed, and cost. It's a significant leap in small model performance, even beating GPT-4o in many benchmarks while reducing latency and cost.
专有 多模态
GPT-4.1 nano is OpenAI's fastest and cheapest model available in the GPT-4.1 family. It delivers exceptional performance at a small size with its 1 million token context window. Ideal for tasks like classification or autocompletion.
专有 多模态
GPT-4.5 is OpenAI's most advanced model, offering improved reasoning, coding, and creative capabilities with faster performance and longer context handling than GPT-4. It features enhanced instruction following, reduced hallucinations, and better factual accuracy.
专有 多模态
Omni-era GPT for multimodal chat, practical coding, and general assistants
专有 多模态
GPT-4o ('o' for 'omni') is a multimodal AI model that accepts text, audio, image, and video inputs, and generates text, audio, and image outputs. It matches GPT-4 Turbo performance on text and code, with improvements in non-English languages, vision, and audio understanding.
专有 多模态
GPT-4o ('o' for 'omni') is a multimodal AI model that accepts text, audio, image, and video inputs, and generates text, audio, and image outputs. It matches GPT-4 Turbo performance on text and code, with improvements in non-English languages, vision, and audio understanding.
专有 多模态
GPT model for general reasoning, writing, coding, and tool-assisted tasks
专有 多模态
GPT-4o mini is OpenAI's latest cost-efficient small model, designed to make AI intelligence more accessible and affordable. It excels in textual intelligence and multimodal reasoning, outperforming previous models like GPT-3.5 Turbo. With a context window of 128K tokens and support for text and vision, it offers low-cost, real-time applications such as customer support chatbots. Priced at 15 cents per million input tokens and 60 cents per million output tokens, it is significantly cheaper than its predecessors. Safety is prioritized with built-in measures and improved resistance to security threats.
专有 多模态
GPT-5 is our flagship model for coding, reasoning, and agentic tasks across domains. The best model for coding and agentic tasks with higher reasoning capabilities and medium speed.
专有 多模态
Chat-tuned GPT model for conversational assistance, writing, and tool workflows
专有
GPT-5 Codex has been trained specifically for conducting code reviews and finding critical flaws. When reviewing, it navigates your codebase and analyzes code patterns to identify potential security vulnerabilities, performance issues, and bugs.
专有 多模态
A faster, more cost-efficient version of GPT-5 for well-defined tasks. Great for well-defined tasks and precise prompts with high reasoning capabilities at reduced cost.
专有 多模态
GPT-5 nano is our fastest, cheapest version of GPT-5. It's great for summarization and classification tasks with average reasoning capabilities and very fast speed.
专有 多模态
Higher-accuracy GPT-5 tier for tough analysis, coding reviews, and planning
专有 多模态
Sharper GPT-5 generation for coding, product work, and tool-assisted tasks
专有 多模态
Chat-tuned GPT-5.1 for polished assistants, writing, and product conversations
专有 多模态
Codex GPT for repository edits, code review, and practical software agents
专有 多模态
Coding-optimized GPT model for repository edits, reviews, and agentic software work
专有 多模态
Coding-optimized GPT model for repository edits, reviews, and agentic software work
专有 多模态
Reliable GPT generation for broad coding, writing, and tool-assisted product work
专有 多模态
Chat-tuned GPT model for conversational assistance, writing, and tool workflows
专有 多模态
Code-specialist GPT for repository edits, reviews, and long-running software agents
专有 多模态
Higher-accuracy GPT-5.2 variant for tougher reasoning and review workflows
专有 多模态
Chat-tuned GPT model for conversational assistance, writing, and tool workflows
专有 多模态
Coding-optimized GPT model for repository edits, reviews, and agentic software work
专有 多模态
Agent-ready GPT for coding and computer-use workflows at a lower cost
专有 多模态
Strong small GPT for coding subagents, quick tool use, and high-volume work
专有 多模态
Cheapest GPT-5.4 lane for simple routing, extraction, and bulk automation
专有 多模态
More exact GPT-5.4 tier for demanding professional reasoning and agent tasks
专有 多模态
Default frontier GPT for coding, computer use, research, and knowledge work
专有 多模态
Compact GPT model for low-latency assistance and high-volume workloads
专有 多模态
Highest-accuracy GPT-5.5 tier for slower, precision-heavy reasoning and coding
新发布
专有 多模态
Cost-efficient GPT-5.6 model for fast, high-volume workloads
新发布
专有 多模态
Frontier GPT-5.6 model for complex professional work, coding, and agentic workflows
新发布
专有 多模态
Balanced GPT-5.6 model for capable, cost-efficient everyday work
专有 多模态
OpenAI image model for production generation, edits, and brand-safe visual workflows
专有 多模态
Image model for prompt-driven generation, editing, and visual design workflows
专有 多模态
Image model for prompt-driven generation, editing, and visual design workflows
新发布
专有 多模态
Realtime speech-to-speech model with configurable reasoning, tool use, and robust voice-agent behavior
专有
A research preview model focused on mathematical and logical reasoning capabilities, demonstrating improved performance on tasks requiring step-by-step reasoning, mathematical problem-solving, and code generation. The model shows enhanced capabilities in formal reasoning while maintaining strong general capabilities.
专有
o1-mini is a cost-efficient language model developed by OpenAI, designed for advanced reasoning tasks while minimizing computational resources.
专有
A research preview model focused on mathematical and logical reasoning capabilities, demonstrating improved performance on tasks requiring step-by-step reasoning, mathematical problem-solving, and code generation. The model shows enhanced capabilities in formal reasoning while maintaining strong general capabilities.
专有 多模态
o1-pro is OpenAI's advanced language model optimized for complex reasoning and specialized professional tasks, offering enhanced capabilities while maintaining high efficiency.
专有 多模态
OpenAI's most powerful reasoning model. o3 is a well-rounded and powerful model across domains. It sets a new standard for math, science, coding, and visual reasoning tasks. It also excels at technical writing and instruction-following. Use it to think through multi-step problems that involve analysis across text, code, and images.
专有 多模态
Research model for long-horizon investigation, synthesis, and analytical reports
专有
A smaller variant of O3, expected to offer enhanced multimodal capabilities, improved reasoning, and more efficient resource utilization compared to previous models while maintaining strong performance on core tasks.
专有 多模态
Version of o3 with more compute for better responses. The o3-pro model uses more compute to think harder and provide consistently better answers. Designed to tackle tough problems with advanced reasoning capabilities.
专有 多模态
o4-mini is OpenAI's latest small o-series model, optimized for fast, effective reasoning with exceptionally efficient performance in coding and visual tasks. It is faster and more affordable than o3.
专有 多模态
Research model for long-horizon investigation, synthesis, and analytical reports
开源 多模态
Open Whisper checkpoint for robust multilingual transcription and captioning
开源 多模态
Speech transcription model for accurate audio-to-text and captioning workflows