AI Model Directory
Meet the underlying models behind the tools: vendors, modalities, and which tools use them. Browse by vendor or modality, then verify against official docs on the tool pages.
GPT-4o
1 toolsOpenAI
OpenAI's flagship multimodal model with text, image, and audio in one, low latency and balanced multilingual performance — the default engine of ChatGPT and many third-party tools.
GPT-4
7 toolsOpenAI
OpenAI's fourth-generation model, known for reasoning and writing quality, used as the default engine by many writing, coding, and knowledge tools.
GPT-3.5
1 toolsOpenAI
OpenAI's earlier mainstream model — fast and cheap, common for latency-sensitive, moderate-precision chat and text tasks.
DALL·E 3
2 toolsOpenAI
OpenAI's third text-to-image model with accurate prompt understanding and clean text rendering, built into ChatGPT and Bing Image Creator.
DALL·E 2
1 toolsOpenAI
OpenAI's second text-to-image model with editing and variations, which commercialized diffusion-based image generation.
DALL·E 系列
1 toolsOpenAI
The OpenAI text-to-image family from DALL·E 2 to DALL·E 3, serving conversational and standalone image creation.
Sora
1 toolsOpenAI
OpenAI's video generation model producing coherent multi-shot HD video from text or images — the reference model for text-to-video.
Claude 系列
19 toolsAnthropic
Anthropic's Claude family, known for long context, writing, and coding; widely adopted by developer and writing tools.
Claude 3.5 Sonnet
1 toolsAnthropic
The workhorse Claude 3.5 model with balanced reasoning, writing, and coding — a default in Cursor, Windsurf, and similar tools.
Claude 3.5
2 toolsAnthropic
The Claude 3.5 generation, strong at long documents and code comprehension.
Claude 3 Opus
1 toolsAnthropic
Claude 3's flagship for the hardest reasoning and creative tasks.
Claude 3 Haiku
1 toolsAnthropic
Claude 3's lightweight model — low latency and cost for high-frequency chat and classification.
Claude 3
1 toolsAnthropic
The Claude 3 generation that introduced vision, spanning Haiku, Sonnet, and Opus.
Gemini 系列
2 toolsGoogle's natively multimodal family across Ultra, Pro, Flash, and Nano, with standout long context and deep Search/Workspace integration.
Gemini Pro
1 toolsGemini's workhorse for general chat and creation with balanced text and image understanding.
Gemini Flash
1 toolsGemini's fast tier serving high-volume requests at low latency and cost.
Gemini Ultra
1 toolsGemini's flagship for the hardest reasoning and creative tasks.
Veo 3
1 toolsGoogle DeepMind
Google DeepMind's video model with high resolution, camera-language control, and integrated audio — top-tier quality and consistency.
Llama 系列
1 toolsMeta
Meta's open-weight model family and the de facto base for local and self-hosted deployments.
Grok 4
1 toolsxAI
xAI's fourth-generation model with stronger reasoning and realtime information, deeply tied to X data.
Grok 3
1 toolsxAI
xAI's third-generation model for realtime Q&A in a direct style.
Mistral Large
1 toolsMistral
Mistral's flagship with strong multilingual and coding ability, popular for European enterprise deployments.
Mistral Small
1 toolsMistral
Mistral's efficient lightweight model for local and edge deployment.
DeepSeek V3
1 toolsDeepSeek
DeepSeek's open flagship rivaling top closed models in reasoning and code, with very competitive API pricing and wide developer adoption.
DeepSeek V4
1 toolsDeepSeek
DeepSeek's next generation pushing deeper reasoning and longer context for professional development and research.
DeepSeek V4 Flash
1 toolsDeepSeek
The fast variant of DeepSeek V4 serving high-volume calls at lower cost and latency.
DeepSeek 系列
1 toolsDeepSeek
DeepSeek's open model family; the V3/R1 line is influential in reasoning and code, adopted widely in China and abroad.
Qwen Max
1 tools阿里巴巴
Qwen's flagship with strong Chinese understanding and multimodal ability, widely deployed in enterprises.
Qwen Plus
1 tools阿里巴巴
Qwen's balanced tier for most Chinese chat and writing scenarios.
Qwen Turbo
1 tools阿里巴巴
Qwen's high-throughput tier for large-scale text processing at low cost.
Wan 2.x(通义万相)
1 tools阿里巴巴
Alibaba's open video generation family, top-tier among open models for text- and image-to-video, widely used in Chinese creator communities.
海螺视频模型
1 toolsMiniMax
MiniMax's video model with natural motion and realistic physics, understanding Chinese instructions well.
Kimi 模型
1 tools月之暗面
Moonshot's long-context models known for ultra-long document reading and file parsing.
SDXL
10 toolsStability AI
Stability AI's open flagship text-to-image model with a huge local-deployment and fine-tuning ecosystem; the base model for Civitai and ComfyUI communities.
Stable Diffusion 1.5
4 toolsStability AI
The classic Stable Diffusion version — compact with a mature ControlNet ecosystem, the default entry point for local deployment.
SD Turbo
1 toolsStability AI
Stability AI's realtime text-to-image model with one-step generation for interactive canvases.
SVD(Stable Video Diffusion)
1 toolsStability AI
Stable Video Diffusion, the reference open model for image-to-video with broad community fine-tuning.
Stable Audio 2.x
1 toolsStability AI
Stability AI's music and sound model generating commercially usable music, samples, and effects.
FLUX 系列
7 toolsBlack Forest Labs
Black Forest Labs' image models with leading quality and text rendering; Pro, Dev, and Schnell cover commercial to open self-hosted.
Flux Pro
1 toolsBlack Forest Labs
FLUX's commercial flagship with the highest image quality for professional design and marketing.
Flux Dev
1 toolsBlack Forest Labs
FLUX's developer tier with open weights for local fine-tuning — the main open self-hosted choice.
Flux Schnell
1 toolsBlack Forest Labs
FLUX's fast tier under Apache 2.0 with one-step generation for realtime and large-scale output.
Real-ESRGAN
1 tools开源社区
The open-source image super-resolution model running fully local and free — the engine inside upscalers like Upscayl.
Midjourney V6
1 toolsMidjourney
Midjourney V6, the industry benchmark for image quality and artistic style in creative design and concept art.
Midjourney V5
1 toolsMidjourney
Midjourney V5 with a big realism jump and a massive community corpus.
Recraft V3
2 toolsRecraft
Recraft V3 leads in vector generation and brand-style control for designer workflows.
Ideogram 2.0
1 toolsIdeogram
Ideogram 2.0 with reliable text rendering for posters, logo concepts, and typographic images.
Ideogram 1.0
1 toolsIdeogram
Ideogram 1.0 pioneered dependable text rendering in text-to-image.
Leonardo Phoenix
1 toolsLeonardo AI
Leonardo AI's next-gen foundation model with much better quality and prompt adherence for game and creative assets.
Leonardo Diffusion
1 toolsLeonardo AI
Leonardo AI's earlier diffusion line with rich style presets and a large creative community.
Firefly 3
1 toolsAdobe
Adobe's third Firefly model, deeply integrated into Photoshop and Illustrator with clear commercial licensing.
Firefly 2
1 toolsAdobe
Adobe's second Firefly model with mature generative fill and text effects inside design workflows.
Mystic
1 toolsFreepik
Freepik's high-fidelity image model producing richly detailed output, bundled with its asset library and licensing.
Magnific Upscaler
1 toolsMagnific AI
Magnific's upscaler known for hallucinated detail reconstruction, turning low-res images into high-detail artwork.
Playground v2
1 toolsPlayground
Playground's canvas model combining generation, inpainting, and layer editing.
Krea Realtime
1 toolsKrea
Krea's realtime model rendering sketches live on canvas for extremely fast concept iteration.
Runway Gen-3 Alpha
1 toolsRunway
Runway Gen-3 Alpha with strong text-to-video quality and motion control — a pro video staple.
Runway Gen-2
1 toolsRunway
Runway Gen-2, the milestone that popularized text-to-video with a mature tutorial ecosystem.
Pika 2.0
1 toolsPika
Pika 2.0 with effect templates and character-consistent videos, common for social shorts.
Pika 1.0
1 toolsPika
Pika 1.0, a lightweight and approachable model popular for creative social video.
Dream Machine 1.5
1 toolsLuma AI
Luma's video model known for natural physics and cinematic shots, with keyframe control and loops.
Luma Capture
1 toolsLuma AI
Luma's capture and reconstruction tech turning video and photos into 3D scenes and Gaussian splats.
Mochi 1
1 toolsGenmo
Genmo's open video model deployable locally — a milestone for open video generation.
LTX Video
1 toolsLTX Studio
LTX Studio's video model integrating storyboards, character consistency, and shot control for film narrative.
HeyGen AI
1 toolsHeyGen
HeyGen's avatar model turning scripts into realistic talking-head video with strong translation and lip-sync.
VEED AI
1 toolsVEED
VEED's in-browser AI covering auto captions, AI voiceover, noise removal, and avatars.
Suno V4
1 toolsSuno
Suno V4 generates complete songs with vocals in one pass, with natural structure and strong audio quality.
Suno V3
1 toolsSuno
Suno V3 was the first text-to-song model to reach publishable quality, sparking mainstream AI music.
Udio V2
1 toolsUdio
Udio V2 excels at vocal songs and full tracks across genres, with regional regeneration and fine editing.
ElevenLabs TTS
1 toolsElevenLabs
ElevenLabs' TTS family umbrella — industry benchmark for realism and emotional expression.
Eleven Multilingual v2
1 toolsElevenLabs
ElevenLabs' multilingual voice model with standout cross-lingual consistency across 29 languages.
Eleven Turbo v2
1 toolsElevenLabs
ElevenLabs' low-latency voice model for realtime conversation and live voiceover.
Murf Voice
1 toolsMurf
Murf's voiceover library with 120+ voices in 20+ languages, common in corporate training and explainers.
LOVO Genny
1 toolsLOVO
LOVO's Genny voice model with 500+ voices and expressive emotion, paired with video editing.
PlayHT 2.0
1 toolsPlay.ht
PlayHT 2.0's low-latency streaming voice model with a multilingual voice library.
Speechify Voice
1 toolsSpeechify
Speechify's reading model turning documents, web pages, and books into natural speech for learning and accessibility.
Kits Voice
1 toolsKits.AI
Kits.AI's voice-model training for cloning your own voice or using licensed artist voices in demos.
Adobe Enhance
1 toolsAdobe
Adobe Podcast's speech enhancement model turning noisy recordings into studio-quality audio in one click, free to use.
Meshy-2
1 toolsMeshy
Meshy-2, the second generation with better text-to-3D and image-to-3D quality for game and e-commerce assets.
Meshy-1
1 toolsMeshy
Meshy-1, the first generation making text-to-3D approachable.
Tripo SR
1 toolsTripo AI
Tripo AI's 3D model with fast image- and text-to-3D and steadily improving topology.
Rodin Gen-2
1 toolsDeemos
Deemos' high-fidelity 3D model focused on photorealistic character and portrait assets for film and virtual humans.
Spline AI
1 toolsSpline
Spline's in-browser 3D design AI generating scenes and motion from prompts, embeddable on web pages.
Grammarly AI
1 toolsGrammarly
Grammarly's writing assistant model covering grammar, tone, and plagiarism across browser and mobile keyboard.
Jasper AI
1 toolsJasper
Jasper's marketing content model with brand-voice training and multi-channel templates for enterprise marketing automation.
Copy.ai 模型
1 toolsCopy.ai
Copy.ai's GTM workflow models automating sales copy, lead nurturing, and content production.
Writesonic 模型
1 toolsWritesonic
Writesonic's all-purpose writing models covering SEO articles, ads, and chatbot building.
Rytr AI
1 toolsRytr
Rytr's lightweight writing model with 40+ use cases and tone templates at an affordable price.
Consensus AI
1 toolsConsensus
Consensus' academic Q&A model answering from peer-reviewed papers with findings and citation-quality markers.
Character AI
1 toolsCharacter.AI
Character.AI's roleplay model with a huge community-built character ecosystem.
Replit AI
1 toolsReplit
Replit's coding models building runnable, deployable apps from prompts.
CodiumAI / Qodo 模型
1 toolsQodo
Qodo's test-generation and code-review models analyzing code behavior to produce boundary cases.
ClickUp Brain
1 toolsClickUp
ClickUp Brain's project-management models auto-writing tasks, summarizing docs, and answering status questions.
Zapier AI
1 toolsZapier
Zapier's automation AI chaining thousands of apps with AI steps and smart decisions.
Make AI
1 toolsMake
Make's visual-automation AI for complex branching and data-transformation scenarios.
Gamma AI
1 toolsGamma
Gamma's presentation model generating polished decks and sites from outlines or documents.
Beautiful AI
1 toolsBeautiful.ai
Beautiful.ai's smart-layout model auto-aligning and reflowing slides as content changes.
Figma AI
1 toolsFigma
Figma AI's design models generating mockups from prompts and auto-renaming layers and content.
Descript AI
1 toolsDescript
Descript's editing model that pioneered transcript-driven audio and video editing.
Canva AI
1 toolsCanva
Canva's built-in AI covering text-to-image, smart editing, copy, and presentations inside the design-collaboration platform.
Fliki Voice
1 toolsFliki
Fliki's voiceover library with 2,000+ AI voices in 80+ languages for text-to-video.
Kayra
1 toolsNovelAI
NovelAI's story-continuation model for long-form fiction with strong style and character consistency.
NAI Diffusion
1 toolsNovelAI
NovelAI's anime-style image model with good character consistency and a strong fan-fiction community reputation.
Captions AI
1 toolsCaptions
Captions' talking-head video models covering auto captions, eye-contact correction, and background replacement.
Related models
What model pages help you with
Tool pages describe product capabilities; model pages answer what engine sits underneath: how versions from the same vendor differ, and whether your shortlisted tools share the same base. Checking the base first narrows the field faster than comparing tool-level features alone.