Find, compare and review AI models

Reviuws tracks 137 AI models across 35 providers with specs, pricing per million tokens, benchmark scores and community reviews.

  • Gemini 3 Pro — google: Google's most capable model at launch, with state-of-the-art reasoning and deep multimodal understanding across text, image, audio and video. It powers the Gemi
  • Veo 3.1 — google: Google's Veo 3.1 is a cutting-edge generative AI model that produces high-quality video with native, synchronized audio. Supporting resolutions up to 1080p, it
  • GPT-4o — openai: GPT-4o is OpenAI's natively multimodal model supporting text, image and audio inputs with fast, low-latency responses. It powers ChatGPT and the API with a 128k
  • o3 — openai: o3 is a reasoning-focused model that uses extended chain-of-thought to solve complex math, science and coding problems. It supports tool use during reasoning. I
  • Claude Haiku 4.5 — anthropic: Claude Haiku 4.5 is Anthropic's fastest and most cost-effective model, optimized for near-instantaneous response times. It delivers exceptional performance on h
  • DeepSeek-R1 — deepseek: DeepSeek-R1 is a reasoning-focused model trained with large-scale reinforcement learning, achieving performance comparable to OpenAI's o1 on math and coding ben
  • Claude Sonnet 4.5 — anthropic: Claude 4.5 Sonnet is Anthropic's premier mid-tier model designed to deliver elite-level reasoning, coding, and comprehension capabilities. It excels at deeply u
  • Gemini 2.5 Flash — google: Gemini 2.5 Flash balances speed, cost and reasoning quality with a configurable thinking budget. It supports a 1M token context and full multimodal input. It's
  • Claude Opus 4.5 — anthropic: Claude 4.5 Opus is Anthropic's flagship model designed to tackle the most demanding cognitive tasks, offering unmatched depth in reasoning and precision. It is
  • Sora 2 — openai: Sora 2 is OpenAI's cutting-edge text-to-video generator that seamlessly integrates high-fidelity synchronized audio directly into its video outputs. Building up
  • Grok 4 — xai: Grok 4 is xAI's flagship conversational AI, engineered for advanced analytical reasoning and deep conceptual synthesis. Building on its predecessor's strengths,
  • Mistral Medium 3 — mistral: An enterprise-focused multimodal model that delivers frontier-level coding, reasoning and vision quality at a fraction of flagship pricing. Mistral positions it
  • GPT-5 — openai: OpenAI's flagship unified model that combines fast responses with built-in deep reasoning, replacing the separate GPT-4o and o-series split. It set new state-of
  • Llama 4 Maverick — meta: Llama 4 Maverick is a state-of-the-art natively multimodal Mixture-of-Experts (MoE) model developed by Meta, engineered to excel in high-performance chat and co
  • Kimi K2 — moonshot ai: Kimi K2 is a 1T parameter (32B active) mixture-of-experts model trained for agentic tool-use and coding, released with open weights under a modified MIT license
  • FLUX.1.1 [pro] — black forest labs: FLUX1.1 [pro] is a 12B parameter diffusion transformer producing high-fidelity images from text prompts with improved speed over the original FLUX.1. It's avail
  • Midjourney V7 — midjourney: Midjourney's first new image model in nearly a year, with smarter prompt understanding, richer textures and better coherence for hands and objects. It made mode
  • DeepSeek-V3.2 — deepseek: Open-weight successor to V3.2-Exp that introduces DeepSeek Sparse Attention for efficient long-context reasoning and agentic work. A high-compute Speciale varia
  • Qwen3-Max — alibaba: Alibaba's largest Qwen release, a trillion-parameter-scale mixture-of-experts model with a production thinking mode for coding, reasoning and agentic tasks. It
  • GPT-5.5 — openai: GPT-5.5 represents OpenAI's next-generation frontier model, specifically optimized for highly complex reasoning, advanced coding tasks, and multi-step analytica
  • Claude Fable 5.1 — anthropic: Anthropic's September 2026 frontier refresh; 52.6% on Terminal-Bench Science, with a 75% cache-read price cut.
  • SDXL — stability: Stable Diffusion XL (SDXL) is the definitive open-source standard for text-to-image generation, renowned for its massive community-driven ecosystem. Utilizing a
  • Runway Gen-4 — runway: Runway Gen-4 represents the pinnacle of production-grade AI video generation, offering filmmakers and creators unprecedented control over physics, camera moveme
  • Whisper Large v3 — openai: Whisper Large v3 is OpenAI's state-of-the-art open-source speech recognition model, trained on millions of hours of diverse audio data to deliver industry-leadi
  • text-embedding-3-small — openai: Text-embedding-3-small is OpenAI's highly efficient and cost-effective embedding model designed to convert textual data into numerical vectors. It offers a sign
  • o4-mini — openai: o4-mini delivers strong reasoning performance at lower cost and latency than o3. It supports tool use and is optimized for math and coding. It's designed for hi
  • Gemini 2.5 Pro — google: Gemini 2.5 Pro is Google's flagship model with native multimodality and a 1M token context window, featuring built-in 'thinking' for complex reasoning. It leads
  • Gemini 2.5 Flash-Lite — google: Gemini 2.5 Flash-Lite is optimized for high-throughput, latency-sensitive tasks at the lowest cost in the Gemini 2.5 family. It retains a 1M token context windo
  • text-embedding-004 — google: text-embedding-004 generates vector representations for text used in search, clustering and classification. It's accessible via the Gemini API. It supports task
  • Claude Opus 4.1 — anthropic: Claude Opus 4.1 is Anthropic's top-tier model, excelling at agentic coding, complex reasoning and long-horizon tasks. It supports a 200k context window and visi
  • Claude Haiku 4.5 — anthropic: Claude Haiku 4.5 is optimized for speed and affordability while retaining solid reasoning ability inherited from the Claude 4 family. It supports a 200k context
  • Mistral Large 2 — mistral: Mistral Large 2 is a 123B parameter dense model with strong multilingual, coding and reasoning capabilities and a 128k context window. It's available under a re
  • Codestral 2508 — mistral: Codestral is fine-tuned specifically for code generation, completion and understanding across 80+ programming languages, with a 256k context window. It supports
  • Claude Mythos 5.1 — anthropic: Trusted-access counterpart to Fable 5.1 with elevated cyber capability, gated under Anthropic's access program.
  • Whisper-1 — openai: Whisper is a general-purpose speech recognition model trained on diverse multilingual audio. It's available open-source and via OpenAI's API. It handles transcr
  • Gemma 3 — google: Gemma 3 is a family of open-weight models (1B-27B) with multimodal and multilingual support, derived from Gemini research. It supports a 128k context window. It
  • Claude Sonnet 4.5 — anthropic: Claude Sonnet 4.5 offers strong coding and agentic capabilities at a mid-tier price point, with a 200k context window (1M in beta). It's positioned as Anthropic
  • Llama 3.3 70B — meta: Llama 3.3 70B delivers performance comparable to Llama 3.1 405B at a fraction of the size, focused purely on text tasks. It supports a 128k context window and m
  • Mistral Small 3.2 — mistral: Mistral Small 3.2 is a 24B parameter open-weight model with multimodal support, tuned for instruction following and reduced repetition errors. It offers a 128k
  • DeepSeek-V3.1 — deepseek: DeepSeek-V3.1 is a 671B parameter (37B active) mixture-of-experts model combining fast and thinking modes in one model. It offers strong coding and reasoning at
  • Qwen2.5-VL-72B — alibaba: Qwen2.5-VL-72B provides advanced visual understanding including document parsing, video comprehension and object grounding. It supports a 128k context window an
  • Deepgram Nova-3 — deepgram: Deepgram's most accurate speech-to-text model, tuned for noisy, multi-speaker enterprise audio. It extends the Nova line's lead in real-time transcription laten
  • Gemini 3.8 Flash — google: Google's September 2026 workhorse for coding and agents, with a defenders-only Cyber variant.
  • Gemini 3.8 Flash Cyber — google: Cyber-capability tier of Gemini 3.8 Flash, available to vetted security teams only.
  • Qwen3-235B-A22B — alibaba: Qwen3-235B-A22B is a mixture-of-experts model with 235B total/22B active parameters, supporting seamless switching between thinking and non-thinking modes. It s
  • Qwen 3 235B — alibaba: Alibaba's Qwen 3 235B is an exceptionally powerful open-weight model optimized for elite multilingual understanding and advanced programming tasks. Built on a m
  • Gemini 3.7 Flash — google: Google's most intelligent workhorse model of August 2026 for coding and agents, with introductory pricing.
  • Gemini 3.8 Live — google: Google's most advanced live-dialogue voice model, with an Extended Thinking variant for multi-step reasoning mid-conversation.
  • Muse Spark 1.3 — meta: Meta's low-cost September 2026 model at a blended price near $0.10 per million tokens.
  • DeepSeek-V4.1-Flash — deepseek: DeepSeek's September 2026 release cutting agent memory costs fourfold over V4.
  • GPT-5.6 Terra — openai: GPT-5.6 Terra is a balanced powerhouse within OpenAI's latest model family, engineered to deliver a cost-effective blend of advanced intelligence and efficiency
  • Claude Fable 5 — anthropic: Claude Fable 5 is Anthropic's premier model designed to power highly sophisticated, autonomous, long-running agents. It features a massive 1M-token context wind
  • GPT-5.6 Sol — openai: GPT-5.6 Sol represents OpenAI's absolute pinnacle of machine intelligence, specifically engineered to tackle ultra-complex reasoning challenges and autonomous a
  • GPT-5.4 mini — openai: GPT-5.4 mini is OpenAI's highly efficient and cost-effective model optimized for high-volume automated tasks. It excels in driving specialized coding sub-agents
  • GPT-5.4 nano — openai: GPT-5.4 nano is OpenAI's highly optimized, lightweight model designed for high-throughput, low-latency text processing tasks. Positioned as the most cost-effect
  • Gemini 3.1 Pro (Preview) — google: Gemini 3.1 Pro is a highly advanced multimodal reasoning model from Google, specifically optimized for complex chat interactions and deep analytical coding. Equ
  • Nano Banana 2 (Gemini 3.1 Flash Image) — google: Nano Banana 2 (Gemini 3.1 Flash Image) is Google's highly efficient model designed for rapid image generation and editing. It offers an exceptionally cost-effec
  • Qwen3.8-Max — alibaba: Alibaba's 2.4T-parameter vision-language flagship (95B active) — first Max-tier with downloadable weights under a custom licence.
  • Qwen3.8-27B — alibaba: Dense 27B open-weight Qwen3.8 release for self-hosting and fine-tuning.
  • Qwen3.8-Flash-Next — alibaba: Multimodal 125B MoE with just 6B active parameters per token — an early preview of the Qwen4 architecture, open-weighted with FP8 checkpoint.