Find, compare and review AI models

Reviuws tracks 34 AI models across 11 providers with specs, pricing per million tokens, benchmark scores and community reviews.

  • SDXL — stability: The go-to open image model. Huge LoRA and community ecosystem.
  • Runway Gen-4 — runway: Runway's production-grade video model — strong controls for creators.
  • Whisper Large v3 — openai: Open-source multilingual speech-to-text. Robust across accents and noise.
  • text-embedding-3-large — openai: High-quality text embeddings for retrieval and semantic search.
  • text-embedding-3-small — openai: Cost-efficient text embeddings for high-volume RAG.
  • GPT Image 2 — openai: OpenAI's state-of-the-art image generation and editing model, listed in the current API model catalog.
  • Grok 4.3 — xai: Cheaper long-context Grok with a 1M-token window. $1.25/$2.50 per 1M tokens under 200k prompt tokens.
  • Claude Fable 5 — anthropic: Anthropic's most capable widely released model, built for long-running agents. 1M-token context, 128k max output, always-on adaptive thinking. $10/$50 per 1M to
  • Claude Opus 5 — anthropic: Anthropic's recommended default for complex agentic coding and enterprise work. 1M context, 128k max output, $5/$25 per 1M tokens.
  • Claude Sonnet 5 — anthropic: The best Claude balance of speed and intelligence. 1M context, $3/$15 per 1M tokens list price (introductory $2/$10 through August 31, 2026).
  • Claude Haiku 4.5 — anthropic: Claude Haiku 4.5 is Anthropic's fastest and most cost-effective model, optimized for near-instantaneous response times. It delivers exceptional performance on h
  • GPT-5.6 Sol — openai: GPT-5.6 Sol represents OpenAI's absolute pinnacle of machine intelligence, specifically engineered to tackle ultra-complex reasoning challenges and autonomous a
  • GPT-5.6 Terra — openai: Balances intelligence and cost in the GPT-5.6 family. 1.05M context, $2.50/$15 per 1M tokens.
  • GPT-5.6 Luna — openai: GPT-5.6 tuned for cost-sensitive, high-volume workloads. 1.05M context, $1/$6 per 1M tokens.
  • GPT-5.5 — openai: GPT-5.5 represents OpenAI's next-generation frontier model, specifically optimized for highly complex reasoning, advanced coding tasks, and multi-step analytica
  • GPT-5.4 mini — openai: GPT-5.4 mini is OpenAI's highly efficient and cost-effective model optimized for high-volume automated tasks. It excels in driving specialized coding sub-agents
  • GPT-5.4 nano — openai: GPT-5.4 nano is OpenAI's highly optimized, lightweight model designed for high-throughput, low-latency text processing tasks. Positioned as the most cost-effect
  • GPT-Realtime 2.1 — openai: OpenAI's current realtime speech-to-speech model with reasoning and tool use.
  • Sora 2 — openai: OpenAI's text-to-video model with synchronized audio, billed per second of generated video.
  • Gemini 3.6 Flash — google: Google's most intelligent model built for speed, with strong search grounding. $1.50/$7.50 per 1M tokens; free tier available.
  • Gemini 3.5 Flash — google: Google's Gemini 3.5 Flash is a highly optimized, cost-effective model designed specifically for high-speed, high-volume conversational tasks and multi-turn agen
  • Gemini 3.5 Flash-Lite — google: Google's most cost-efficient GA model for high-volume agentic tasks, translation and data processing. $0.30/$2.50 per 1M tokens.
  • Gemini 3.1 Pro (Preview) — google: Gemini 3.1 Pro is a highly advanced multimodal reasoning model from Google, specifically optimized for complex chat interactions and deep analytical coding. Equ
  • Nano Banana 2 (Gemini 3.1 Flash Image) — google: Gemini 3.1 Flash Image — fast image generation and editing. Roughly $0.067 per 1K image, $0.151 per 4K image.
  • Veo 3.1 — google: Google's latest video model with native audio. Standard $0.40/sec at 720p–1080p; Fast from $0.10/sec.
  • Gemini Embedding 2 — google: Google's first multimodal embedding model — text, image, video, audio and PDFs in one space. $0.20 per 1M text input tokens.
  • Grok 4.5 — xai: xAI's frontier model for coding and agentic work. 500k context, $2/$6 per 1M tokens under 200k prompt tokens ($4/$12 above).
  • Mistral Large 3 — mistral: Mistral Large 3 is Mistral AI's premier European frontier model, engineered to deliver top-tier reasoning, advanced coding capabilities, and highly sophisticate
  • DeepSeek V4 Pro — deepseek: Mixture-of-experts model with 1.6T total and 49B active parameters, hybrid compressed attention, and a 1M-token context window.
  • Qwen3.5-397B-A17B — alibaba: Alibaba's first Qwen3.5 release: a 397B-parameter MoE with 17B active parameters, Apache 2.0 licensed, hosted as Qwen3.5-Plus on Model Studio.
  • Llama 4 Maverick — meta: Meta's natively multimodal MoE with 400B total and 17B active parameters, under the Llama 4 Community License.
  • Llama 4 Scout — meta: Efficient Llama 4 MoE: 109B total / 17B active parameters with an industry-leading 10M-token context window.
  • FLUX.2 [pro] — black-forest-labs: FLUX 2 Pro by Black Forest Labs is a top-tier text-to-image model engineered engineered specifically for elite-level photorealism and precise visual rendering.
  • FLUX.2 [dev] — black-forest-labs: FLUX 2 Dev is an advanced open-weights image generation model developed by Black Forest Labs, designed to deliver state-of-the-art visual quality and prompt adh