Find, compare and review AI models

Reviuws tracks 152 AI models across 39 providers with specs, pricing per million tokens, benchmark scores and community reviews.

  • Gemini 3 Pro — google: Google's most capable model at launch, with state-of-the-art reasoning and deep multimodal understanding across text, image, audio and video. It powers the Gemi
  • Veo 3.1 — google: Google's Veo 3.1 is a cutting-edge generative AI model that produces high-quality video with native, synchronized audio. Supporting resolutions up to 1080p, it
  • GPT-4o — openai: GPT-4o is OpenAI's natively multimodal model supporting text, image and audio inputs with fast, low-latency responses. It powers ChatGPT and the API with a 128k
  • o3 — openai: o3 is a reasoning-focused model that uses extended chain-of-thought to solve complex math, science and coding problems. It supports tool use during reasoning. I
  • Claude Haiku 4.5 — anthropic: Claude Haiku 4.5 is Anthropic's fastest and most cost-effective model, optimized for near-instantaneous response times. It delivers exceptional performance on h
  • DeepSeek-R1 — deepseek: DeepSeek-R1 is a reasoning-focused model trained with large-scale reinforcement learning, achieving performance comparable to OpenAI's o1 on math and coding ben
  • Claude Sonnet 4.5 — anthropic: Claude 4.5 Sonnet is Anthropic's premier mid-tier model designed to deliver elite-level reasoning, coding, and comprehension capabilities. It excels at deeply u
  • Gemini 2.5 Flash — google: Gemini 2.5 Flash balances speed, cost and reasoning quality with a configurable thinking budget. It supports a 1M token context and full multimodal input. It's
  • Claude Opus 4.5 — anthropic: Claude 4.5 Opus is Anthropic's flagship model designed to tackle the most demanding cognitive tasks, offering unmatched depth in reasoning and precision. It is
  • Sora 2 — openai: Sora 2 is OpenAI's cutting-edge text-to-video generator that seamlessly integrates high-fidelity synchronized audio directly into its video outputs. Building up
  • Grok 4 — xai: Grok 4 is xAI's flagship conversational AI, engineered for advanced analytical reasoning and deep conceptual synthesis. Building on its predecessor's strengths,
  • Mistral Medium 3 — mistral: An enterprise-focused multimodal model that delivers frontier-level coding, reasoning and vision quality at a fraction of flagship pricing. Mistral positions it
  • GPT-5 — openai: OpenAI's flagship unified model that combines fast responses with built-in deep reasoning, replacing the separate GPT-4o and o-series split. It set new state-of
  • Llama 4 Maverick — meta: Llama 4 Maverick is a state-of-the-art natively multimodal Mixture-of-Experts (MoE) model developed by Meta, engineered to excel in high-performance chat and co
  • Kimi K2 — moonshot ai: Kimi K2 is a 1T parameter (32B active) mixture-of-experts model trained for agentic tool-use and coding, released with open weights under a modified MIT license
  • FLUX.1.1 [pro] — black forest labs: FLUX1.1 [pro] is a 12B parameter diffusion transformer producing high-fidelity images from text prompts with improved speed over the original FLUX.1. It's avail
  • Midjourney V7 — midjourney: Midjourney's first new image model in nearly a year, with smarter prompt understanding, richer textures and better coherence for hands and objects. It made mode
  • DeepSeek-V3.2 — deepseek: Open-weight successor to V3.2-Exp that introduces DeepSeek Sparse Attention for efficient long-context reasoning and agentic work. A high-compute Speciale varia
  • Qwen3-Max — alibaba: Alibaba's largest Qwen release, a trillion-parameter-scale mixture-of-experts model with a production thinking mode for coding, reasoning and agentic tasks. It
  • GPT-5.5 — openai: GPT-5.5 represents OpenAI's next-generation frontier model, specifically optimized for highly complex reasoning, advanced coding tasks, and multi-step analytica
  • Claude Fable 5.1 — anthropic: Anthropic's September 2026 frontier refresh; 52.6% on Terminal-Bench Science, with a 75% cache-read price cut.
  • Claude Opus 5.5 — anthropic: Leading Anthropic model for long-horizon agentic coding at $4/$20 per million tokens.
  • Grok 4.7 — xai: Frontier xAI coder with 500k context at Grok 4.6 pricing.
  • SDXL — stability: Stable Diffusion XL (SDXL) is the definitive open-source standard for text-to-image generation, renowned for its massive community-driven ecosystem. Utilizing a
  • Runway Gen-4 — runway: Runway Gen-4 represents the pinnacle of production-grade AI video generation, offering filmmakers and creators unprecedented control over physics, camera moveme
  • Whisper Large v3 — openai: Whisper Large v3 is OpenAI's state-of-the-art open-source speech recognition model, trained on millions of hours of diverse audio data to deliver industry-leadi
  • text-embedding-3-small — openai: Text-embedding-3-small is OpenAI's highly efficient and cost-effective embedding model designed to convert textual data into numerical vectors. It offers a sign
  • o4-mini — openai: o4-mini delivers strong reasoning performance at lower cost and latency than o3. It supports tool use and is optimized for math and coding. It's designed for hi
  • Gemini 2.5 Pro — google: Gemini 2.5 Pro is Google's flagship model with native multimodality and a 1M token context window, featuring built-in 'thinking' for complex reasoning. It leads
  • Gemini 2.5 Flash-Lite — google: Gemini 2.5 Flash-Lite is optimized for high-throughput, latency-sensitive tasks at the lowest cost in the Gemini 2.5 family. It retains a 1M token context windo
  • text-embedding-004 — google: text-embedding-004 generates vector representations for text used in search, clustering and classification. It's accessible via the Gemini API. It supports task
  • Claude Opus 4.1 — anthropic: Claude Opus 4.1 is Anthropic's top-tier model, excelling at agentic coding, complex reasoning and long-horizon tasks. It supports a 200k context window and visi
  • Claude Haiku 4.5 — anthropic: Claude Haiku 4.5 is optimized for speed and affordability while retaining solid reasoning ability inherited from the Claude 4 family. It supports a 200k context
  • Mistral Large 2 — mistral: Mistral Large 2 is a 123B parameter dense model with strong multilingual, coding and reasoning capabilities and a 128k context window. It's available under a re
  • Codestral 2508 — mistral: Codestral is fine-tuned specifically for code generation, completion and understanding across 80+ programming languages, with a 256k context window. It supports
  • Whisper-1 — openai: Whisper is a general-purpose speech recognition model trained on diverse multilingual audio. It's available open-source and via OpenAI's API. It handles transcr
  • GPT-6 Sol — openai: Mid-tier GPT-6 for coding and agents at $2/$10 per million tokens.
  • Gemma 3 — google: Gemma 3 is a family of open-weight models (1B-27B) with multimodal and multilingual support, derived from Gemini research. It supports a 128k context window. It
  • GPT-6 Luna — openai: Cheapest GPT-6 tier for high-volume work at $0.10/$0.50 per million tokens.
  • Claude Sonnet 4.5 — anthropic: Claude Sonnet 4.5 offers strong coding and agentic capabilities at a mid-tier price point, with a 200k context window (1M in beta). It's positioned as Anthropic
  • Llama 3.3 70B — meta: Llama 3.3 70B delivers performance comparable to Llama 3.1 405B at a fraction of the size, focused purely on text tasks. It supports a 128k context window and m
  • Mistral Small 3.2 — mistral: Mistral Small 3.2 is a 24B parameter open-weight model with multimodal support, tuned for instruction following and reduced repetition errors. It offers a 128k
  • DeepSeek-V3.1 — deepseek: DeepSeek-V3.1 is a 671B parameter (37B active) mixture-of-experts model combining fast and thinking modes in one model. It offers strong coding and reasoning at
  • MiMo-V2.6-Pro — xiaomi: MIT-licensed 1.02T-parameter open MoE with 1M context.
  • Qwen2.5-VL-72B — alibaba: Qwen2.5-VL-72B provides advanced visual understanding including document parsing, video comprehension and object grounding. It supports a 128k context window an
  • Gemini 3.8 Flash — google: Google's September 2026 workhorse for coding and agents, with a defenders-only Cyber variant.
  • Gemini 3.8 Flash Cyber — google: Cyber-capability tier of Gemini 3.8 Flash, available to vetted security teams only.
  • Deepgram Nova-3 — deepgram: Deepgram's most accurate speech-to-text model, tuned for noisy, multi-speaker enterprise audio. It extends the Nova line's lead in real-time transcription laten
  • MiMo-V2.6-Flash — xiaomi: Efficient MIT-licensed 309B MoE sibling of MiMo-V2.6-Pro.
  • Step 5 Preview — stepfun: Preview of StepFun's fifth-generation multimodal model.
  • Qwen3.8-Omni-Flash — alibaba: Fast omni-modal Qwen3.8 model for real-time multimodal assistants.
  • Grok Voice Transcribe 2.0 — xai: xAI's second-generation low-latency speech-to-text model.
  • Kimi K2.8 Preview — moonshot ai: Preview of Moonshot's agentic long-context Kimi K2.8.
  • Qwen3-235B-A22B — alibaba: Qwen3-235B-A22B is a mixture-of-experts model with 235B total/22B active parameters, supporting seamless switching between thinking and non-thinking modes. It s
  • Qwen 3 235B — alibaba: Alibaba's Qwen 3 235B is an exceptionally powerful open-weight model optimized for elite multilingual understanding and advanced programming tasks. Built on a m
  • Gemini 3.7 Flash — google: Google's most intelligent workhorse model of August 2026 for coding and agents, with introductory pricing.
  • Gemini 3.8 Live — google: Google's most advanced live-dialogue voice model, with an Extended Thinking variant for multi-step reasoning mid-conversation.
  • Muse Spark 1.3 — meta: Meta's low-cost September 2026 model at a blended price near $0.10 per million tokens.
  • DeepSeek-V4.1-Flash — deepseek: DeepSeek's September 2026 release cutting agent memory costs fourfold over V4.
  • GPT-5.6 Terra — openai: GPT-5.6 Terra is a balanced powerhouse within OpenAI's latest model family, engineered to deliver a cost-effective blend of advanced intelligence and efficiency