AI Models Directory
Explore 65 AI models with standardized fact sheets
Midjourney V7
Midjourney, Inc.
Basic $10/month, Standard $30/month, Pro $60/month, Mega $120/month. Annual 20% discount. No free tier.
Stable Diffusion 3.5
Stability AI
Free for non-commercial and <$1M revenue. API: $0.065 per image (Large). Enterprise license for $1M+ revenue companies.
DALL-E 3
OpenAI
API: $0.04-$0.12 per image (standard/HD, various resolutions). ChatGPT Plus: $20/month includes access.
OpenAI GPT-5.2 Series
OpenAI
GPT-5.2 operates under a strict proprietary license, with access provided exclusively through OpenAI's API and its ChatGPT consumer products. The API pricing for the standard GPT-5.2 model is set at $1.75 per million input tokens and $14.00 per million output tokens. OpenAI offers significant cost-saving mechanisms, including a 90% discount for cached inputs ($0.175 per million tokens) and a 50% discount for asynchronous tasks via the Batch API. The flagship GPT-5.2 Pro model is priced at a sig
Make.com
Make (formerly Integromat)
Free: 1,000 credits/month. Core: $9/month (10,000 credits). Pro: $16/month (10,000 credits). Teams: $29/month (10,000 credits). Enterprise: Custom. Credit-based system.
Runway Gen-3 Alpha
Runway ML
Free: 125 credits (one-time). Standard: $12/month (625 credits). Pro: $28/month (2,250 credits). Unlimited: $76/month (2,250 + unlimited relaxed). Enterprise: Custom. Gen-3 Alpha: 10 credits/sec.
Figma AI
Figma Inc.
Starter: Free (150 AI credits/day). Professional: $16/month (3,000 AI credits). Organization: $55/month (3,500 credits). Enterprise: $90/month (4,250 credits). Figma AI module: $16.
Udio AI
Udio
Free: 10 credits/day + 100 monthly. Standard: $10/month ($96 annual, 1,200 credits). Pro: $30/month ($288 annual, 4,800 credits). Commercial use rights included.
Cursor AI
Anysphere
Hobby: Free (limited). Pro: $20/month (500 premium requests + $20 API credit). Ultra: $200/month. Teams: $40/user/month. Enterprise: Custom.
ChatGPT Enterprise
OpenAI
Custom pricing (~$60/user/month reported, 150 user minimum, 12-month contract). Estimated $108,000 annual minimum.
Google Gemini 3 Pro
Gemini 3 Pro is a proprietary model available through Google's developer APIs. During its preview phase, Google has implemented a tiered pricing structure based on the length of the context used. For context lengths up to 200,000 tokens, the API pricing is $2.00 per million input tokens and $12.00 per million output tokens. For contexts exceeding 200,000 tokens, the price increases to $4.00 per million input tokens and $18.00 per million output tokens. Google has a history of reducing prices on
Leonardo.AI
Leonardo.AI (acquired by Canva July 2024)
Free: 150 daily tokens. Apprentice: $12/month (8,500 tokens). Artisan: $30/month (25,000 tokens + unlimited relaxed). Maestro: $60/month (60,000 tokens + unlimited). API: $9-$299/month.
Pika Labs
Pika Labs
Free: 80-300 credits/month. Basic: $10/month (700 credits). Pro: $35/month (2,300 credits + unlimited chill). Fancy: $95/month (6,000 credits). Credit-based consumption.
GitHub Copilot
GitHub (Microsoft)
Free: 2,000 completions + 50 chats/month. Pro: $10/month (unlimited completions). Pro+: $39/month (1,500 premium requests). Business: $19/user/month. Enterprise: $39/user/month.
Canva AI (Magic Studio)
Canva
Free: Limited AI (~50 uses). Pro: $15/month ($120 annual, ~500 AI uses). Teams: $10/user/month (min 3, pooled AI). Business: $16.67/month ($200 annual). Enterprise: Custom.
Anthropic Claude 4.5 Series
Anthropic
The Claude 4.5 models are governed by a proprietary license and are accessible via Anthropic's API and through major cloud platforms. The pricing is structured per million tokens (MTok). For Claude Opus 4.5, the rate is $5.00 per MTok for input and $25.00 per MTok for output. For Claude Sonnet 4.5, the pricing for its standard 200K context window is $3.00 per MTok for input and $15.00 per MTok for output. When using the extended 1M context window, the rates for Sonnet 4.5 increase to $6.00 per
Bolt.new
StackBlitz
Free: 300K daily/1M monthly tokens. Pro: $25/month (10M tokens). Teams: $30/member/month (10M+ shared). Enterprise: Custom. 10% discount yearly.
Zapier AI
Zapier Inc.
Free: 100 tasks/month. Professional: $19.99/month (750 tasks). Team: $69/month (2,000 tasks). Enterprise: Custom. Task-based pricing scales with usage.
Replit AI
Replit
Starter: Free (limited AI). Core: $20/month ($25 credits). Teams: $35/user/month ($40 credits). Enterprise: Custom. Agent: $0.25/checkpoint. Assistant: $0.05/edit.
V0 by Vercel
Vercel
Free: $5 credits/month. Premium: $20/month ($20 credits + $2 daily). Team: $30/user/month. Business: $100/user/month. Enterprise: Custom. Token-based: $0.50-$17.50 per 1M tokens.
Notion AI
Notion Labs
Free: Limited AI trial (20 responses). Plus: $12/month ($10 annual). Business: $24/month ($20 annual, includes full AI). Enterprise: Custom (includes full AI).
HeyGen
HeyGen
Free: 3 videos/month (3 min, 720p). Creator: $29/month (unlimited 30 min, 1080p). Team: $39/seat/month (4K, 2 seats min). Enterprise: Custom pricing.
Adobe Firefly
Adobe Inc.
Free: Limited use. Standard: $9.99/month (2,000 credits, ~20 videos). Pro: $29.99/month (~70 videos). Premium: TBA (~500 videos). Unlimited image gen until Jan 15, 2026.
Jasper AI
Jasper AI
Creator: $49/month ($39 annual). Teams: $125/month ($99 annual). Business: Custom pricing. 7-day free trial. Unlimited words on all plans.
Copy.ai
Copy.ai
Chat: $29/month ($24 annual, 5 seats). Growth: $1,000/month (75 seats). Expansion: $2,000/month (150 seats). Scale: $3,000/month (200 seats). Enterprise: Custom.
Suno AI
Suno, Inc.
Free: 50 credits/day (~10 songs, non-commercial). Pro: $8-$10/month (~500 songs). Premier: $24-$30/month (~2,000 songs). Commercial licensing with paid plans.
Grammarly
Grammarly Inc.
Free: $0. Pro: $12/month ($144 annual), $20/month (quarterly), $30/month (monthly). Enterprise: Custom pricing. 25% discount often available.
ElevenLabs
ElevenLabs
Free: 10K credits/month. Starter: $5/month (30K credits). Creator: $11-$22/month (100K credits). Pro: $99/month (500K credits). Scale: $330/month (2M credits). Business: $1,320/month (11M credits). Enterprise: Custom.
Mistral AI Devstral 2 and Large 2
Mistral AI
Mistral's licensing strategy is nuanced. Mistral Large 2 is released under the Mistral Research License, which permits non-commercial use; commercial self-deployment requires a separate license. Its API pricing is $2.00 per million input tokens and $6.00 per million output tokens. In a significant move for the open-source community, Devstral 2 is released under a modified MIT license, making it fully open-source and permissively licensed for commercial use. During an initial free period, its AP
xAI Grok-4 Series
xAI
All Grok-4 models are proprietary and accessed via the xAI API. The pricing structure clearly reflects the trade-off between intelligence and cost. The premium models, Grok-4 and Grok-4.1, are priced at $3.00 per million input tokens and $15.00 per million output tokens. The highly cost-efficient "Fast" variants, Grok 4 Fast and Grok-4.1 Fast, are priced at just $0.20 per million input tokens and $0.50 per million output tokens, making them exceptionally competitive for high-throughput automati
Alibaba Qwen Series (Qwen2.5 and Qwen3)
Alibaba
A major strength of the Qwen series is its licensing. All the mentioned models are released under the permissive Apache 2.0 license, which allows for broad commercial use without significant restrictions. This has fostered wide adoption. API access through providers like DeepInfra and Novita is very affordable. For instance, the Qwen3 32B model starts at $0.10 per million input tokens and $0.30 per million output tokens. The MoE variant, Qwen3 30B A3B, is priced similarly, highlighting the cost
DeepSeek AI Models (R1 and V3.2)
DeepSeek AI
All the major DeepSeek models, including R1 and the V3.2 series, are released under the highly permissive MIT license, allowing for unrestricted commercial use. This has made them very popular among developers and startups. The pricing for API access reflects the models' efficiency. The standard DeepSeek-V3.2 model is priced at just $0.28 per million input tokens and $0.42 per million output tokens through DeepSeek's own API, with even lower costs for cached inputs. The sparse attention mechani
Google Gemma 3 Series
All Gemma 3 models are released with open weights under the Gemma license, which permits commercial use and distribution. For managed API access, providers like DeepInfra and Novita offer extremely competitive pricing. For example, the Gemma 3 27B model is priced at approximately $0.10 per million input tokens and $0.20 per million output tokens. The smaller 4B model is even cheaper, at $0.02 per million input tokens and $0.04 per million output tokens. This aggressive pricing makes the Gemma 3
Meta Llama 4 Series
Meta
The models are released under the Llama 4 Community License Agreement, which allows for broad use, modification, and distribution, including for commercial purposes, with some restrictions. As open-weight models, they can be self-hosted for free (excluding hardware costs). For those preferring managed access, a vibrant ecosystem of third-party API providers has emerged. Pricing is exceptionally competitive due to the models' computational efficiency. For Llama 4 Maverick, API input costs range
Lovable
Lovable (Stockholm, Sweden)
Lovable is an AI-powered full-stack app builder that enables non-technical users, founders, and designers to create production-ready web applications using natural language. Built on React, Supabase, and Tailwind CSS, it handles frontend, backend, database, authentication, and deployment. It features an autonomous Agent Mode, Visual Edits for real-time UI changes, and GitHub sync for code ownership.
Perplexity AI
Perplexity AI
Perplexity AI is an AI-powered answer engine and research platform that has evolved into a comprehensive agentic system in 2026. It features Perplexity Computer (autonomous multi-step agent routing 20+ frontier models), Deep Research for structured multi-source analysis, Model Council for parallel model comparison, and enterprise integrations with Microsoft 365 and Teams.
Google NotebookLM
Google NotebookLM is an AI-powered research and note-taking tool powered by Gemini 3.5 and the Antigravity agent-first IDE. It enables users to upload documents and build knowledge bases, generating Audio Overviews, Video Overviews, flashcards, quizzes, and structured outputs (PDF, Word, Excel, PowerPoint, CSV) from source material. Major 2026 updates added a secure cloud computer for code execution and agentic source discovery.
Windsurf (by Codeium)
Codeium
Windsurf is an AI-native IDE built on VS Code by Codeium, featuring the Cascade agentic system for autonomous multi-file editing, terminal integration, and codebase-aware task execution. In 2026 it evolved into Devin Desktop, a unified command center for managing local and cloud-based coding agents. It supports multiple frontier models including Claude, GPT-5.4, and Codeium's proprietary SWE-1.5.
Perplexity Comet
Perplexity AI
Perplexity Comet is an AI-native web browser built on Chromium (desktop) and WebKit (iOS) that integrates an AI assistant directly into the browsing experience. Launched in July 2025 and made free to all users in October 2025, it features agentic browsing, Deep Research mode for generating structured outputs, voice interaction, cross-tab summarization, and Background Assistants for autonomous multi-step task execution.
Devin Desktop
Cognition AI
Devin Desktop (formerly Windsurf IDE, rebranded June 2, 2026) is an autonomous AI coding agent platform by Cognition AI. It functions as an 'Agent Command Center' with a Kanban-style interface for managing multiple parallel AI agents. Built on the proprietary SWE-1.6 model, it supports autonomous task delegation, multi-repo workflows, and the open Agent Client Protocol (ACP) for third-party agent integration.
SellerClaw
SellerAI
SellerClaw is an autonomous AI agent platform for e-commerce store management, using a supervisor-led multi-agent architecture to handle product sourcing, listing creation, ad management, pricing, order fulfillment, and customer support. Supports Shopify, eBay, and Amazon with native integrations for Meta Ads, Google Ads, and Google Merchant Center. Launched June 5, 2026 as Product Hunt #1 Product of the Day.
n8n
n8n GmbH
n8n is an open-source, self-hostable workflow automation platform with native AI agent capabilities. It uses execution-based billing (one execution = one complete workflow run, regardless of node count) rather than per-step pricing, making it significantly more cost-effective than Zapier or Make.com for complex multi-step automations. The Community Edition is free and self-hosted with unlimited executions.
BrowserAct
BrowserAct
BrowserAct is a browser automation layer designed specifically for AI agents, enabling reliable persistent browser workflows that handle authentication, CAPTCHAs, and anti-bot protections. It provides session persistence, human-in-the-loop remote assist for blockers, stealth execution, and structured data extraction. Launched June 25, 2026 as Product Hunt #1 Product of the Day.
Claude Code
Anthropic
Claude Code is Anthropic's agentic coding tool that operates within the terminal, IDEs (VS Code, JetBrains), and a dedicated desktop application. Generally released May 2025 and actively maintained through June 2026, it understands entire codebases, executes multi-file edits, manages git workflows, and supports sub-agent spawning (up to 5 levels deep) via natural language commands.
Manus AI
Butterfly Effect
Manus is an autonomous AI agent platform that decomposes complex, multi-step goals into subtasks and executes them within a sandboxed virtual computer environment. It uses a multi-agent architecture with specialized sub-agents for web browsing, code execution, file management, and task planning. A Meta acquisition attempt was blocked by Chinese regulators in April 2026; the company continues to operate independently from Singapore.
Mem0
Mem0 (Y Combinator)
Mem0 is an AI memory infrastructure platform providing persistent, intelligent memory for AI agents and applications. It compresses chat history into compact, distilled memories to optimize token usage and latency, and supports deployment across Kubernetes, private cloud, and air-gapped environments. SOC 2 Type 1 certified and HIPAA compliant.
OpenClaw
OpenClaw (Open Source / OneClaw managed hosting)
OpenClaw is an MIT-licensed open-source framework for self-hosting personal AI assistants with multi-channel messaging support (WhatsApp, Telegram, Slack, Discord, Signal, iMessage, and more). It features intelligent model routing (ClawRouters) for 40-60% API cost reduction, a Skill System for custom plugins, and OpenClaw-RL for training personalized agents through reinforcement learning from user feedback.
Granola
Granola
Granola is a bot-free AI meeting notes platform that captures system audio locally without joining meetings as a visible participant. It enhances user-taken rough notes post-meeting using the captured transcript. The platform raised a $125M Series C in 2026 and supports MCP integration for connecting meeting context to other AI tools like Claude and ChatGPT.
Firecrawl
Firecrawl (Y Combinator)
Firecrawl is an open-source web data infrastructure layer for AI agents, providing a unified API to search, scrape, and crawl websites and convert them into LLM-ready structured data. It achieves 96% web coverage with 3.4s latency and 93% fewer input tokens compared to raw HTML, and is SOC 2 Type 2 certified.
Upstream
Upstream (Y Combinator)
Upstream is an AI-native email platform designed for both humans and AI agents to collaborate within the same inbox. Founded by former Asana and Algolia product leaders, it enables AI agents to triage, draft, and prioritize email while requiring human approval before sending. Raised $3M pre-seed from Y Combinator, Connect Ventures, and Roosh Ventures. Launched June 2026.
Kiro
Amazon Web Services (AWS)
Kiro is an agentic AI coding platform from AWS that launched internationally on May 7, 2026, as the official successor to Amazon Q Developer. It introduces a 'spec-driven development' methodology requiring formal requirements, design, and task artifacts before code generation, using an SMT solver to mathematically verify specifications for logical contradictions. Built on Code OSS and integrated with Amazon Bedrock.
Runway Gen-4.5
Runway
Runway Gen-4.5 is the current flagship video generation model from Runway, succeeding Gen-3 Alpha. As of July 2026, Runway has expanded the platform with Agent Skills for automated advertising campaign creation and integration of third-party models including Gemini Omni Flash for high-speed video generation.
Probably
Probably
Probably is an enterprise AI reliability platform that raised $9 million in June 2026 to build deterministic validation systems for AI workflows. Rather than building a new foundation model, it provides a validation and orchestration layer that coordinates smaller, locally-run models to achieve extremely high accuracy (targeting 99.99%) by reducing ambiguity and verifying outputs.
FLUX.2 Pro
Black Forest Labs
FLUX.2 is a family of 32-billion-parameter AI image generation and editing models by Black Forest Labs, released in November 2025. Built on a latent flow matching architecture integrating a Mistral-3 vision-language model with a rectified flow transformer, it delivers state-of-the-art photorealism, precise text rendering, multi-reference consistency, and 4K resolution output. The Pro variant is the flagship commercial model balancing speed and visual fidelity.
Kimi K2.7-Code
Moonshot AI
Kimi K2.7-Code is an open-weight coding-focused language model released by Moonshot AI on June 12, 2026, under a Modified MIT license. It is optimized for agentic coding workflows with 30% lower reasoning-token usage than its predecessor K2.6, and serves as the engine behind the Kimi Code terminal-first coding agent.
MinerU2.5-Pro
OpenDataLab (Shanghai AI Laboratory)
MinerU2.5-Pro is a 1.2-billion-parameter vision-language model specialized for high-performance document parsing and understanding. It achieves state-of-the-art results on the OmniDocBench v1.6 benchmark (95.69 overall score) through data-centric engineering rather than architectural changes, processing complex layouts with text, tables, formulas, and images.
MiniMax M3
MiniMax AI
MiniMax M3 is a frontier-tier open-weight multimodal agent model with 428B total parameters (23B activated) and a 1M token context window. Released May 31, 2026, it features MiniMax Sparse Attention (MSA) for 9x prefill and 15x decode speedup over its predecessor, and is specialized for coding, agentic tasks, and complex reasoning.
GPT Image 2
OpenAI
GPT Image 2 is OpenAI's flagship image generation model released in April 2026, succeeding GPT Image 1.5 (December 2025). It integrates a reasoning model into the image generation pipeline, offering advanced photorealism, improved prompt adherence, and multilingual text rendering. It replaced DALL-E 3 (deprecated May 12, 2026) as OpenAI's primary image generation offering.
Gemini 3.5 Flash
Google DeepMind
Gemini 3.5 Flash is Google's high-efficiency multimodal language model released May 19, 2026 at Google I/O 2026. It features a 1M token context window and native multimodality (text, image, video, audio), designed to outperform previous Pro-tier models in coding and agentic tasks at approximately 40% lower cost than Gemini 3.1 Pro for those use cases.
NVIDIA Cosmos 3
NVIDIA
NVIDIA Cosmos 3 is an open frontier foundation model for physical AI, built on a Mixture-of-Transformers (MoT) architecture. It unifies text, image, video, ambient sound, and action generation within a single omnimodal system designed for robotics, autonomous vehicles, and physical environment simulation. Available in Nano (16B) and Super (64B) variants.
Claude Opus 4.8
Anthropic
Claude Opus 4.8 is Anthropic's flagship large language model released May 28, 2026, achieving top rankings on SWE-bench Pro (69.2%) and the Artificial Analysis GDPval-AA leaderboard (1890 Elo). It introduces Dynamic Workflows for orchestrating hundreds of parallel subagents and Effort Control for adjustable reasoning depth, with a 1M token context window by default.
Google Gemma 4
Google DeepMind
Gemma 4 is Google DeepMind's fourth generation of open-weight multimodal AI models, released March 31, 2026 under the Apache 2.0 license. The family includes four variants (E2B, E4B, 26B MoE, 31B Dense) optimized for deployment from edge devices to enterprise servers. The 31B Dense model ranks #3 among open models on Arena AI and achieves 89.2% on AIME 2026 mathematics benchmark.
Google Gemini 3.1 Pro
Google DeepMind
Gemini 3.1 Pro is Google DeepMind's flagship reasoning model released February 19, 2026, as a mid-cycle update to the Gemini 3 series. It achieves 77.1% on ARC-AGI-2 (more than double its predecessor), supports a 1-million-token context window, and introduces three-tier 'Thinking' levels for balancing speed and computational depth. It uses a Transformer-based Mixture-of-Experts architecture.
ZAYA1-8B
Zyphra
ZAYA1-8B is a reasoning-focused mixture-of-experts (MoE) language model from Zyphra with 8.4 billion total parameters and only 760 million active parameters per token. It was trained entirely on AMD Instinct MI300X hardware and achieves competitive reasoning performance at a fraction of the compute cost of larger models.
Claude Sonnet 5
Anthropic
Claude Sonnet 5 is Anthropic's most agentic Sonnet-class model, released June 30, 2026. It can make plans, use tools like browsers and terminals, and run autonomously at a level previously requiring larger and more expensive models. Performance is close to Claude Opus 4.8 at lower prices.