Directory

AI Models

The models worth knowing across five modalities — each with sourced specs, current pricing and an honest verdict. Filter by type below.

Meta open LLMs

Muse Glimmer

Meta's first Apache 2.0 model — a 30B dense multimodal agent optimized for local deployment on consumer GPUs, with DFlash speculative decoding for fast inference.

Free (open weights) / hosted via partners
Meta LLMs

Muse Spark 1.2

Meta's coding-focused flagship — co-trained with Muse Code terminal agent, TB 2.1 82.9%, $1.25/$4.25 standard pricing, 1M context, open weights announced.

$1.25 / $4.25 per 1M tokens (input / output)
Alibaba LLMs

Qwen3.8-Max

Alibaba's new flagship — 2.4T-parameter sparse MoE (95B active), multimodal, top scores on agentic/research benchmarks, $2/$6 pricing, with open weights promised.

$2 / $6 per 1M tokens (input / output)
Anthropic LLMs

Claude Opus 5

Anthropic's July 2026 flagship — near-Fable intelligence at Opus pricing ($5/$25), SOTA on Frontier-Bench, ARC-AGI-3 and OSWorld 2.0, with thinking on by default.

$5 / $25 per 1M tokens (input / output)
Google LLMs

Gemini 3.6 Flash

Google's July 2026 workhorse model — 17% fewer output tokens than 3.5 Flash, cheaper at $1.50/$7.50, optimized for agentic coding and knowledge work.

$1.50 / $7.50 per 1M tokens (input / output)
Moonshot AI open LLMs

Kimi K3

The world's first open 3T-class frontier model — 2.8T parameters (104B active MoE), 1M context, native vision, open weights under a revenue-tiered custom license.

$3 / $15 per 1M tokens (input / output)
xAI LLMs

Grok 4.5

xAI's July 2026 coding-and-agent flagship — 500K context, $2/$6 per 1M, default in Grok Build and Cursor, trained alongside Cursor.

$2 / $6 per 1M tokens (input / output)
Anthropic LLMs

Claude Sonnet 5

Anthropic's most agentic Sonnet yet — near-Opus quality for coding, agents and knowledge work at $2/$10 introductory pricing, and the new default model on claude.ai.

$2 / $10 per 1M tokens (intro, through Aug 31) · $3 / $15 standard
OpenAI LLMs

GPT-5.6

OpenAI's GPT-5.6 family — flagship Sol, balanced Terra and fast Luna — with max/ultra reasoning modes and tiered pricing. Generally available since July 9, 2026.

Sol $5/$30 · Terra $2.50/$15 · Luna $1/$6 per 1M
Sakana AI LLMs

Sakana Fugu

Tokyo-based Sakana AI's June 2026 multi-agent orchestration system — shipped as a single OpenAI-compatible model that dynamically routes each task to a swappable pool of frontier and open models.

Fugu Ultra API $5 / $30 per 1M tokens (input / output)
Z.ai open LLMs

GLM-5.2

Z.ai's open-weight long-horizon flagship — a 753B-param MoE with a stable 1M context under an MIT license, beating GPT-5.5 on several coding benchmarks.

~$1.40 / $4.40 per 1M tokens (input / output)
Moonshot AI open LLMs

Kimi K2.7 Code

Moonshot's open-weight agentic coding flagship — a 1T-param MoE (32B active) with 256K context and native image/video input, thinking ~30% leaner than K2.6.

~$0.95 / $4.00 per 1M tokens (input / output)
Anthropic LLMs

Claude Fable 5

Anthropic's first Mythos-class model for general use — state-of-the-art on nearly every benchmark, with safety classifiers that fall back to Opus 4.8 on sensitive topics. Suspended by US export controls June 12–30, back globally since July 1.

$10 / $50 per 1M tokens (input / output)
Anthropic LLMs

Claude Opus 4.8

The June 2026 overall intelligence leader and top model for agentic coding, with the most careful hallucination calibration of any frontier model.

$5 / $25 per 1M tokens (input / output)
DeepSeek open LLMs

DeepSeek V4-Pro

The open-weight value champion — near-frontier coding (SWE-V ~85%) under an MIT license with a 1M-token context at a fraction of closed-model prices.

~$1.74 / $3.48 per 1M tokens (input / output)
OpenAI LLMs

GPT-5.5

OpenAI's unified reasoning-and-chat flagship — the strongest model for terminal-native agentic work, natively omnimodal across text, image, audio and video.

$5 / $30 per 1M tokens (input / output)
Google LLMs

Gemini 3.1 Pro

The highest-value closed frontier model — top-tier reasoning and native multimodality at roughly $2 per 1M input tokens with a 1M-token context.

$2 / $4–12 per 1M tokens (input / output)
NVIDIA open LLMs

Nemotron 3.5 Lightning

NVIDIA's highest-efficiency open model for long-running agentic AI workloads — built by the Nemotron Coalition, available with NeMo Switchyard for intelligent multi-model routing.

Free (open weights) / hosted via partners
Thinking Machines Lab open LLMs

Inkling

Mira Murati's first model — a 975B/41B-active open-weight MoE with Apache 2.0 license, native text/image/audio, 1M context, designed as a base for fine-tuning.

$1.87 / $4.68 per 1M tokens (64K ctx) · $3.74 / $9.36 (256K ctx)
MiniMax open LLMs

MiniMax M3

The first open-weight model to combine frontier coding, a 1M-token context and native multimodality — a 428B MoE (22B active) built on sparse attention.

~$0.60 / $2.40 per 1M tokens (input / output)
DeepReinforce open LLMs

Ornith-1.0

Ornith-1.0 is DeepReinforce's open-source, self-improving LLM family for agentic coding — from a 9B edge model to a 397B MoE that rivals Claude Opus 4.7 on SWE-Bench.

Open weight (free to self-host)
Alibaba LLMs

Qwen3.7-Max

Alibaba's proprietary "Agent Frontier" flagship — top knowledge scores and long-horizon autonomy across a 1M-token context.

Pay-as-you-go via Alibaba Cloud Model Studio
xAI LLMs

Grok 4.3

xAI's budget-frontier model — strong agentic tool calling and a self-reported low hallucination rate inside a 1M-token window at $1.25 / $2.50 per 1M.

$1.25 / $2.50 per 1M tokens (input / output)
Black Forest Labs open Image

FLUX.2

The open-weight champion for photorealism, text rendering and developer workflows — top quality at a few cents per image via API or self-hosting.

~$0.04–0.08 per image (hosted API)
Google Image

Nano Banana 2

Google's Gemini 3 Pro Image model — 4K photorealism free inside the Gemini app, with near-perfect text rendering and conversational editing.

Free in the Gemini app · API via Gemini (usage-based)
Midjourney Image

Midjourney v8.1

The industry leader for raw artistic quality and aesthetic polish — a subscription-only image generator beloved by designers and concept artists.

$10–120 / month (subscription)
xAI Image

Grok Imagine Image 2.0

xAI's editing-first image model — Arena #2 in both T2I and image editing behind GPT-Image-2, with magic wand, segmentation, multi-reference compositing and templates.

Free on grok.com; API pricing TBD
OpenAI Image

GPT Image 2

OpenAI's image model — the strongest at prompt instruction-following and conversational editing, native to the ChatGPT and OpenAI API stack.

~$0.07 / image (standard), ~$0.13 / image (HD)
Boogu open Image

Boogu-Image 0.1

Apache-2.0 open-source image generation and editing family (Base / Turbo / Edit) with strong bilingual text rendering and near closed-source quality — trained on roughly 10× less data and fully self-hostable.

Free (Apache-2.0 open weights, self-host)
Recraft Image

Recraft V4

The design-and-brand specialist — vector/SVG output, brand-style consistency and strong typography built for professional design work.

Free tier + paid plans (from ~$12 / month)
Ideogram Image

Ideogram 3.0

The specialist for typography and logo design — the most reliable model for correct, legible text inside generated images.

Free tier + ~$16 / month (Plus)
Google Image

Imagen 4

Google's photorealism leader — best-in-class typography and people, available cheaply through Gemini Advanced or per-image on Vertex AI.

~$0.04 / image (Vertex AI) or included with Gemini Advanced ($20/mo)
Stability AI open Image

Stable Diffusion 3.5

The open-weights standard for local, private and customizable image generation — run it on your own GPU at no per-image cost, or via Stability's API.

Free (self-hosted) · API from $0.03–0.08 / image
MiniMax open Video

MiniMax H3

The first major open-weight video model with native stereo audio — a 33B omni-modal transformer generating 4–15s clips at up to 2K resolution from text, image, video and audio references.

API pricing varies by resolution; open weights available
Black Forest Labs Video

FLUX 3

BFL's unified multimodal model — video (up to 20s with native audio), image, and action prediction from one set of weights, built on Self-Flow. Video GA as of August 4, 2026.

$0.06–0.53 / second (varies by mode and resolution)
ByteDance Video

Seedance 2.5

ByteDance's long-sequence video model — native 30-second single-shot clips, native 4K, up to 50 multimodal reference assets, phoneme-level lip-sync and one-pass synced audio.

Not public yet (enterprise beta; general availability early July 2026)
PixVerse Video

PixVerse V6

The best free pick for short-form video — multi-shot character consistency, native audio and 20+ cinematic lens controls, with daily free credits.

Free daily credits · $10–$199 / month · API ~$0.025–$0.115 / sec
Kuaishou Video

Kling 3.0

A top value-for-performance video model — excellent human motion and face consistency, native 4K output and multilingual audio.

~$0.067 / sec (Standard) to ~$0.17 / sec (Pro)
Google Video

Veo 3.1

Google's cinematic video model with native audio — the quality leader for polished, sound-synced clips, billed per second on the Gemini API.

$0.40 / sec (1080p std, with audio) · $0.05–0.12 / sec (Fast/Lite)
Runway Video

Runway Gen-4.5

The professional's video model — rich creative controls like motion brush and camera moves, built into a mature editorial toolset.

Credit-based (~$0.10–0.25 / sec effective)
Alibaba Video

Wan 3.0

Alibaba's unified video generation model — 30-second single-pass clips, multimodal inputs including documents and web pages, $0.05–$0.20/sec API pricing.

$0.05–$0.20 per second (by resolution)
Sand.ai open Video

MAGI-2 Preview

Sand.ai's open-source 114B MoE video model — generates 10-second clips with synchronized audio using just 6B active parameters per token. Apache 2.0.

Free (open weights)
xAI Video

Grok Imagine Video 1.5

xAI's dedicated video generation model — native 1080p, 6–15 seconds with synced audio, image/voice references, built on the Aurora autoregressive engine.

$0.08 / second (API)
Alibaba Video

Wan 2.7

Alibaba's flagship video model — 1080p with native audio sync, first/last-frame and multi-scene controls, from the Wan series whose open-source lineage made it the community base.

~$6 / minute (cloud), a ~33% cut versus Wan 2.5/2.6
ByteDance Video

Seedance 2.0

A multimodal-control video model built for e-commerce and reference-heavy jobs — strong image-to-video fit, motion physics and multi-asset input.

~$0.022–0.14 / sec (varies by provider)
MiniMax Video

Hailuo 2.3

The budget speed champion of AI video — the lowest cost per second and fast turnaround, ideal for drafts and high-volume social content.

~$0.025 / sec
OpenAI Video

Sora 2

OpenAI's video model known for physics realism and synchronized audio, available through ChatGPT and the API.

Via ChatGPT Plus/Pro · ~$0.15 / sec (third-party API routes)
Suno Audio

Suno v5.5

The leading AI music generator for full, radio-ready songs — natural vocals, real song structure, voice cloning and 30+ genres from a single text prompt.

Free tier · paid Pro and Premier plans
ACE Studio / StepFun open Audio

ACE-Step

The open-source full-song generator — free, self-hostable music with vocals that runs on a single consumer GPU.

Free (open source, self-hosted)
ElevenLabs Audio

ElevenLabs v3

The industry-standard text-to-speech model — the most expressive, emotionally rich voice synthesis, with 70+ languages and multi-speaker dialogue.

~$0.12 / minute (v3) · Flash/Turbo ~$0.06 / min
ElevenLabs Audio

ElevenLabs Music v2

The commercially safest AI music generator — built with label and publisher partnerships, with section-by-section editing, a full API and studio-grade exports.

$11 / mo (30 tracks) to $99 / mo (Scale, API)
Google Audio

Lyria 3

Google DeepMind's music generation model — high-fidelity 48kHz audio up to three minutes, available through Vertex AI and the Gemini API.

Via Vertex AI / Gemini API (usage-based)
MiniMax Audio

MiniMax Speech 2.8

A top-ranked text-to-speech model — expressive, multilingual voice synthesis with fast voice cloning and a large built-in voice library.

Usage-based via MiniMax platform (varies)
Stability AI Audio

Stable Audio 2.5

Stability's music-and-sound model — fast, commercially-cleared instrumental tracks and sound design with an enterprise-friendly API.

Via Stability AI subscription / API (usage-based)
Udio Audio

Udio v4

The audio-quality leader among AI music tools — 48kHz rendering with the cleanest instrument separation, for film and brand-grade music.

~$10 / mo (Standard) · ~$30 / mo (Pro)
Meshy 3D

Meshy 5

The most reliable all-rounder for AI 3D — best-in-class PBR textures, fast iteration, and a deep plugin ecosystem for Blender and Unity.

Free tier + $20 / mo (Pro) · $60 / mo (Studio)
Hyper3D (Deemos) 3D

Rodin Gen-2

The highest-fidelity AI 3D generator — production-ready geometry with clean quad topology, the top pick for hero assets and photorealism.

$30 / mo (Creator) + credits · free to generate, pay to download
Tencent open 3D

Hunyuan3D 2.1

The leading open-source 3D generator — fully open weights and training code with a production-ready PBR texture pipeline, self-hostable for free.

Free (self-hosted) · paid hosted APIs available
Hitem3D 3D

Hitem3D

The ultra-detailed specialist — high-resolution 3D models for miniatures, e-commerce and product visualization.

Free tier + paid plans
Luma AI 3D

Luma Genie

A free, accessible text-to-3D generator from Luma AI — a fast, no-cost way to turn prompts into 3D models via web and Discord.

Free
Sloyd 3D

Sloyd AI

Parametric, game-ready 3D — instant low-poly assets with clean topology, tuned for real-time engines.

Free tier + paid plans
Microsoft open 3D

TRELLIS 2

Microsoft's open-source 3D generator known for high-quality Gaussian-splatting visuals and flexible output representations from images or text.

Free (open source, self-hosted)
Tripo AI 3D

Tripo 3.0

The speed-and-value leader for game-ready 3D — clean low-poly topology and auto-rigging at the lowest cost per asset.

From ~$12 / mo (Pro) · free tier with 300 credits/mo