Browse all AI Gateway models
Every model available on Vercel AI Gateway, with API access, pricing, and a playground. 393 models · Page 7 of 7.
Search and filter all models →- OpenAITTS-1TTS is a model that converts text to natural sounding spoken text.
- OpenAITTS-1 HDTTS is a model that converts text to natural sounding spoken text. The tts-1-hd model is optimized for high quality text-to-speech use cases.
- GoogleVeo 3.0Veo 3 is designed to handle a range of video generation tasks, from cinematic narratives to dynamic character animations. With Veo 3, you can create more immersive experiences by not only generating stunning visuals, but also audio like dialogue and sound effects.
- GoogleVeo 3.0 Fast GenerateVeo 3 Fast is a quicker and more cost effective version of Veo 3, allowing developers to create videos with sound while maintaining high quality and optimizing for speed and business use cases. Veo 3 Fast offers both text-to-video and image-to-video modalities.
- GoogleVeo 3.1Veo 3.1 is Google's state-of-the-art model for generating high-fidelity, 8-second 720p, 1080p or 4k videos featuring stunning realism and natively generated audio.
- GoogleVeo 3.1 Fast GenerateVeo 3.1 Fast is a specialized, high-speed variant of Google DeepMind’s Veo 3.1 text-to-video model, optimized for rapid generation of 8-second, high-fidelity videos. It is designed to create cinematic, 1080p, or 720p content with improved prompt adherence and native audio, making it ideal for creating quick, high-quality video clips, social media content, and ad creatives.
- GoogleVeo 3.1 Lite GenerateVeo 3.1 Lite Preview is a high-efficiency, developer-first video model providing high-fidelity video generation, editing, and cinematic control. It leverages the state-of-the-art Veo 3.1 model to democratize professional-grade video AI by offering a scalable, programmable interface for creators and enterprises.
- Voyage AI by MongoDBVoyage 3.5Voyage AI's embedding model optimized for general-purpose and multilingual retrieval quality.
- Voyage AI by MongoDBVoyage 3.5 LiteVoyage AI's embedding model optimized for latency and cost.
- Voyage AI by MongoDBVoyage 4Optimized for general-purpose and multilingual retrieval quality. All embeddings created with the 4 series are compatible with each other.
- Voyage AI by MongoDBVoyage 4 LargeThe best general-purpose and multilingual retrieval quality. All embeddings created with the 4 series are compatible with each other.
- Voyage AI by MongoDBVoyage 4 LiteOptimized for latency and cost. All embeddings created with the 4 series are compatible with each other.
- Voyage AI by MongoDBVoyage Code 2Voyage AI's embedding model optimized for code retrieval (17% better than alternatives). This is the previous generation of code embeddings models.
- Voyage AI by MongoDBVoyage Code 3Voyage AI's embedding model optimized for code retrieval.
- Voyage AI by MongoDBVoyage Finance 2Voyage AI's embedding model optimized for finance retrieval and RAG.
- Voyage AI by MongoDBVoyage Law 2Voyage AI's embedding model optimized for legal retrieval and RAG.
- Voyage AI by MongoDBVoyage Rerank 2.5A generalist reranker optimized for quality with instruction-following and multilingual support.
- Voyage AI by MongoDBVoyage Rerank 2.5 LiteA generalist reranker optimized for both latency and quality with instruction-following and multilingual support.
- Voyage AI by MongoDBVoyage Rerank 3A general-purpose reranker optimized for retrieval accuracy.
- Voyage AI by MongoDBVoyage Rerank 3 LiteA fast, cost-effective reranker optimized for latency-sensitive retrieval.
- Voyage AI by MongoDBvoyage-3-largeVoyage AI's embedding model with the best general-purpose and multilingual retrieval quality.
- Alibaba CloudWan v2.5 Text-to-Video Preview
- Alibaba CloudWan v2.6 Image-to-Video
- Alibaba CloudWan v2.6 Image-to-Video Flash
- Alibaba CloudWan v2.6 Reference-to-Video
- Alibaba CloudWan v2.6 Reference-to-Video Flash
- Alibaba CloudWan v2.6 Text-to-Video
- Alibaba CloudWan v2.7 Reference-to-Video
- Alibaba CloudWan v2.7 Text-to-Video
- Alibaba CloudWan v3.0 VideoAll-in-one video generation model supporting text-to-video, image-to-video, first/last-frame, and omni-modal reference-based generation (image, video, and audio references) with synchronized audio, at up to 30 seconds per clip.
- Alibaba CloudWan v3.0 Video PrimeWan 3.0 Video Prime is Alibaba’s high-speed, all-in-one AI video generation model for creating polished clips from text, images, video, and audio references. It produces videos up to 30 seconds long with synchronized dialogue, music, and sound effects, while accelerated generation makes it ideal for rapid creative iteration and production workflows.
- OpenAIWhisperWhisper is a general-purpose speech recognition model, trained on a large dataset of diverse audio. You can also use it as a multitask model to perform multilingual speech recognition as well as speech translation and language identification.
- Topaz LabsWonder 3.5Wonder 3.5 is Topaz Labs’ image enhancement model for generative upscaling and detail recovery. It repairs and enhances an existing JPEG, PNG, or TIFF image, with controls for enhancement strength, output dimensions, and grain. An input image is required, making it suited to restoring low-quality photos and refining images for larger displays or print.