Browse all AI Gateway models
Every model available on Vercel AI Gateway, with API access, pricing, and a playground. 393 models · Page 3 of 7.
Search and filter all models →- OpenAIGPT 5.6 TerraGPT-5.6 Terra is a balanced GPT-5.6 model for everyday work, with performance comparable to the previous generation at half the cost.
- OpenAIGPT Image 1GPT Image 1 is OpenAI's new state-of-the-art image generation model. It is a natively multimodal language model that accepts both text and image inputs, and produces image outputs.
- OpenAIGPT Image 1 MiniA cost-efficient version of GPT Image 1. It is a natively multimodal language model that accepts both text and image inputs, and produces image outputs.
- OpenAIGPT Image 1.5GPT Image 1.5 is OpenAI's latest image generation model, with better instruction following and adherence to prompts.
- OpenAIGPT Image 2GPT Image 2 is OpenAI's state-of-the-art image generation model for fast, high-quality image generation and editing. It supports flexible image sizes and high-fidelity image inputs.
- OpenAIGPT Image 2.5 FlareGPT Image 2.5 Flare generates images from text and image inputs, optimized for fast, high-quality everyday image generation.
- OpenAIGPT Image 2.5 SunburstGPT Image 2.5 Sunburst generates and edits images from text and image inputs, with a focus on precise editing.
- OpenAIGPT OSS 120BExtremely capable general-purpose LLM with strong, controllable reasoning capabilities
- OpenAIGPT OSS 20BA compact, open-weight language model optimized for low-latency and resource-constrained environments, including local and edge deployments.
- OpenAIGPT OSS Safeguard 120BGPT OSS Safeguard 120B is OpenAI’s open-weight safety model for content moderation and guardrail enforcement. It helps teams evaluate content against custom safety policies and build scalable, adaptable protections for production AI applications.
- OpenAIGPT OSS Safeguard 20BOpenAI's first open weight reasoning model specifically trained for safety classification tasks. Fine-tuned from GPT-OSS, this model helps classify text content based on customizable policies, enabling bring-your-own-policy Trust & Safety AI where your own taxonomy, definitions, and thresholds guide classification decisions.
- OpenAIGPT-3.5 TurboOpenAI's most capable and cost effective model in the GPT-3.5 family optimized for chat purposes, but also works well for traditional completions tasks.
- OpenAIGPT-4 Turbogpt-4-turbo from OpenAI has broad general knowledge and domain expertise allowing it to follow complex instructions in natural language and solve difficult problems accurately. It has a knowledge cutoff of April 2023 and a 128,000 token context window.
- OpenAIGPT-4.1GPT 4.1 is OpenAI's flagship model for complex tasks. It is well suited for problem solving across domains.
- OpenAIGPT-4.1 miniGPT 4.1 mini provides a balance between intelligence, speed, and cost that makes it an attractive model for many use cases.
- OpenAIGPT-4.1 nanoGPT-4.1 nano is the fastest, most cost-effective GPT 4.1 model.
- OpenAIGPT-4oGPT-4o from OpenAI has broad general knowledge and domain expertise allowing it to follow complex instructions in natural language and solve difficult problems accurately. It matches GPT-4 Turbo performance with a faster and cheaper API.
- OpenAIGPT-4o miniGPT-4o mini from OpenAI is their most advanced and cost-efficient small model. It is multi-modal (accepting text or image inputs and outputting text) and has higher intelligence than gpt-3.5-turbo but is just as fast.
- OpenAIGPT-4o mini TranscribeGPT-4o mini Transcribe is a speech-to-text model that uses GPT-4o mini to transcribe audio. It offers improvements to word error rate and better language recognition and accuracy compared to original Whisper models. Use it for more accurate transcripts.
- OpenAIGPT-4o TranscribeGPT-4o Transcribe is a speech-to-text model that uses GPT-4o to transcribe audio. It offers improvements to word error rate and better language recognition and accuracy compared to original Whisper models. Use it for more accurate transcripts.
- OpenAIGPT-5GPT-5 is OpenAI's flagship language model that excels at complex reasoning, broad real-world knowledge, code-intensive, and multi-step agentic tasks.
- OpenAIGPT-5 miniGPT-5 mini is a cost optimized model that excels at reasoning/chat tasks. It offers an optimal balance between speed, cost, and capability.
- OpenAIGPT-5 nanoGPT-5 nano is a high throughput model that excels at simple instruction or classification tasks.
- OpenAIGPT-5 proGPT-5 pro uses more compute to think harder and provide consistently better answers. Since GPT-5 pro is designed to tackle tough problems, some requests may take several minutes to finish.
- OpenAIGPT-5-CodexGPT-5-Codex is a version of GPT-5 optimized for agentic coding tasks in Codex or similar environments.
- OpenAIGPT-5.1-CodexGPT-5.1-Codex is a version of GPT-5.1 optimized for agentic coding tasks in Codex or similar environments.
- OpenAIGPT-6 AstraGPT-6 Astra is OpenAI's most capable model for complex reasoning, coding, computer use, research, and document creation.
- OpenAIGPT-6 LunaGPT-6 Luna is OpenAI's efficient reasoning model for focused, high-volume tasks. It accepts text and images, generates text, and supports a 1,050,000-token context window with up to 128,000 output tokens. Configurable reasoning effort, structured outputs, prompt caching, and Responses API tools support cost-sensitive applications, coding tasks, and automated workflows.
- OpenAIGPT-6 Luna DecisionsGPT-6 Luna Decisions evaluates classification and scoring questions against shared text and image input using OpenAI's Decisions API. It supports predicate, choice, and score questions with probabilities and confidence values.
- OpenAIGPT-6 SolGPT-6 Sol is OpenAI's reasoning model for complex coding and agentic workflows. It accepts text and images, generates text, and supports a 1,050,000-token context window with up to 128,000 output tokens. Configurable reasoning effort, structured outputs, prompt caching, and Responses API tools make it suited to software development and multi-step automation.
- OpenAIGPT-6.1 SolGPT-6.1 Sol is OpenAI's reasoning model for complex coding, computer use, and professional work. It accepts text and images, generates text, and supports a 1,050,000-token context window with up to 128,000 output tokens. Configurable reasoning effort, structured outputs, prompt caching, and Responses API tools support multi-step workflows.
- OpenAIGPT-Live 1Our premier model for natural, expressive voice conversations with smooth interruption handling.
- OpenAIGPT-Realtime miniGPT-Realtime mini is capable of responding to audio and text inputs in realtime over WebRTC, WebSocket, or SIP connections.
- OpenAIGPT-Realtime-1.5GPT-Realtime-1.5 is our flagship audio model for voice agents and customer support.
- OpenAIgpt-realtime-2GPT Realtime 2 is our most capable realtime voice model. It supports speech-to-speech interactions with configurable reasoning effort, stronger instruction following, and more reliable tool use for complex voice-agent workflows.
- OpenAIgpt-realtime-2.1GPT-Realtime-2.1 updates GPT-Realtime-2 with improved alphanumeric recognition, silence and noise handling, and interruption behavior. It supports speech-to-speech interactions with configurable reasoning effort, instruction following, and tool use for complex voice-agent workflows.
- OpenAIgpt-realtime-whisperGPT Realtime Whisper is a streaming speech-to-text model for applications that need low-latency transcript deltas from live audio. It is designed for realtime use cases where developers need to tune latency and accuracy. GPT Realtime Whisper is priced by audio duration rather than text tokens.
- SpaceXAIGrok 4.1 Fast Non-Reasoning
- SpaceXAIGrok 4.1 Fast Reasoning
- SpaceXAIGrok 4.20 Beta Non-ReasoningGrok 4.20 Beta is the newest flagship model from xAI with industry-leading speed and agentic tool calling capabilities. It combines the lowest hallucination rate on the market with strict prompt adherance, delivering consistently precise and truthful responses.
- SpaceXAIGrok 4.20 Beta ReasoningGrok 4.20 Beta is the newest flagship model from xAI with industry-leading speed and agentic tool calling capabilities. It combines the lowest hallucination rate on the market with strict prompt adherance, delivering consistently precise and truthful responses.
- SpaceXAIGrok 4.20 Multi Agent BetaMultiple agents collaborate in parallel to perform deep research tasks.
- SpaceXAIGrok 4.20 Multi-AgentMultiple agents collaborate in parallel to perform deep research tasks.
- SpaceXAIGrok 4.20 Non-ReasoningGrok 4.20 Beta is the newest flagship model from xAI with industry-leading speed and agentic tool calling capabilities. It combines the lowest hallucination rate on the market with strict prompt adherence, delivering consistently precise and truthful responses.
- SpaceXAIGrok 4.20 ReasoningGrok 4.20 Beta is the newest flagship model from xAI with industry-leading speed and agentic tool calling capabilities. It combines the lowest hallucination rate on the market with strict prompt adherence, delivering consistently precise and truthful responses.
- SpaceXAIGrok 4.3Grok 4.3 is a new model matching the scale of Grok 4.20 with an improved architecture and a December 2025 knowledge cutoff.
- SpaceXAIGrok 4.5SpaceXAI's smartest model with frontier performance on coding, knowledge work, and STEM.
- SpaceXAIGrok 4.6Grok 4.6 builds on Grok 4.5 with a particular focus on long-running agents and more ambitious interactive and visual work. It stays with complex tasks across many steps, whether researching a topic, analyzing information, working across a codebase, or turning an idea into a polished application or work artifact.
- SpaceXAIGrok 4.7Grok 4.7 is SpaceXAI’s advanced AI model for coding and professional knowledge work, built to tackle complex, multi-hour tasks with improved self-verification and long-context handling. It strengthens software engineering, document creation, and presentation workflows while maintaining Grok 4.6’s speed and pricing.
- SpaceXAIGrok Build 0.1xAI's fast coding model trained specifically for agentic coding.
- SpaceXAIGrok ImagineState-of-the-art video generation across quality, cost, and latency. Grok Imagine is x.AI's most powerful video-audio generative model yet. Bring an image to life, start from a simple text prompt, or even refine a complex cinematic sequence.
- SpaceXAIGrok Imagine ImageGenerate high-quality images from text prompts with xAI's imagine API.
- SpaceXAIGrok Imagine Image 2.0
- SpaceXAIGrok Imagine Video 1.5
- SpaceXAIGrok Imagine Video 1.5 LiteGrok Imagine Video 1.5 Lite is xAI's lightweight video generation model for turning text prompts and images into clips with native audio. It supports videos from 1 to 15 seconds in 480p, 720p, and 1080p, with seven aspect ratios for formats including widescreen 16:9 and vertical 9:16.
- SpaceXAIGrok STTTranscribe audio to text in 25 languages with batch and streaming modes.
- SpaceXAIGrok TTSGenerate speech with 5 expressive voices, speech tags, and telephony codecs.
- SpaceXAIGrok Voice Think Fast 1.0Build real-time voice applications powered by Grok. Stream audio and text bidirectionally via WebSocket for voice assistants, phone agents, and interactive voice systems.
- SpaceXAIGrok Voice Think Fast 2.0Grok Voice Think Fast 2.0, our next-generation voice model with improved intelligence, transcription accuracy, and conversational capabilities.
- Tencent CloudHy3