Trouvez, comparez et évaluez des modèles d'IA

Reviuws suit 101 modèles d'IA répartis sur 29 fournisseurs avec leurs caractéristiques, prix par million de tokens, scores de benchmark et avis de la communauté.

  • GPT-5.5 — openai: GPT-5.5 represents OpenAI's next-generation frontier model, specifically optimized for highly complex reasoning, advanced coding tasks, and multi-step analytica
  • SDXL — stability: Stable Diffusion XL (SDXL) is the definitive open-source standard for text-to-image generation, renowned for its massive community-driven ecosystem. Utilizing a
  • Runway Gen-4 — runway: Runway Gen-4 represents the pinnacle of production-grade AI video generation, offering filmmakers and creators unprecedented control over physics, camera moveme
  • Whisper Large v3 — openai: Whisper Large v3 is OpenAI's state-of-the-art open-source speech recognition model, trained on millions of hours of diverse audio data to deliver industry-leadi
  • text-embedding-3-small — openai: Text-embedding-3-small is OpenAI's highly efficient and cost-effective embedding model designed to convert textual data into numerical vectors. It offers a sign
  • GPT Image 2 — openai: GPT Image 2 is OpenAI's state-of-the-art visual model designed for advanced image generation and editing directly through an API. It allows developers and creat
  • o4-mini — openai: o4-mini delivers strong reasoning performance at lower cost and latency than o3. It supports tool use and is optimized for math and coding. It's designed for hi
  • Gemini 2.5 Pro — google: Gemini 2.5 Pro is Google's flagship model with native multimodality and a 1M token context window, featuring built-in 'thinking' for complex reasoning. It leads
  • Gemini 2.5 Flash-Lite — google: Gemini 2.5 Flash-Lite is optimized for high-throughput, latency-sensitive tasks at the lowest cost in the Gemini 2.5 family. It retains a 1M token context windo
  • text-embedding-004 — google: text-embedding-004 generates vector representations for text used in search, clustering and classification. It's accessible via the Gemini API. It supports task
  • Claude Opus 4.1 — anthropic: Claude Opus 4.1 is Anthropic's top-tier model, excelling at agentic coding, complex reasoning and long-horizon tasks. It supports a 200k context window and visi
  • Claude Haiku 4.5 — anthropic: Claude Haiku 4.5 is optimized for speed and affordability while retaining solid reasoning ability inherited from the Claude 4 family. It supports a 200k context
  • Mistral Large 2 — mistral: Mistral Large 2 is a 123B parameter dense model with strong multilingual, coding and reasoning capabilities and a 128k context window. It's available under a re
  • Codestral 2508 — mistral: Codestral is fine-tuned specifically for code generation, completion and understanding across 80+ programming languages, with a 256k context window. It supports
  • Whisper-1 — openai: Whisper is a general-purpose speech recognition model trained on diverse multilingual audio. It's available open-source and via OpenAI's API. It handles transcr
  • Gemini 2.5 Flash — google: Gemini 2.5 Flash balances speed, cost and reasoning quality with a configurable thinking budget. It supports a 1M token context and full multimodal input. It's
  • Gemma 3 — google: Gemma 3 is a family of open-weight models (1B-27B) with multimodal and multilingual support, derived from Gemini research. It supports a 128k context window. It
  • Claude Sonnet 4.5 — anthropic: Claude Sonnet 4.5 offers strong coding and agentic capabilities at a mid-tier price point, with a 200k context window (1M in beta). It's positioned as Anthropic
  • Llama 3.3 70B — meta: Llama 3.3 70B delivers performance comparable to Llama 3.1 405B at a fraction of the size, focused purely on text tasks. It supports a 128k context window and m
  • Mistral Small 3.2 — mistral: Mistral Small 3.2 is a 24B parameter open-weight model with multimodal support, tuned for instruction following and reduced repetition errors. It offers a 128k
  • DeepSeek-V3.1 — deepseek: DeepSeek-V3.1 is a 671B parameter (37B active) mixture-of-experts model combining fast and thinking modes in one model. It offers strong coding and reasoning at
  • DeepSeek-R1 — deepseek: DeepSeek-R1 is a reasoning-focused model trained with large-scale reinforcement learning, achieving performance comparable to OpenAI's o1 on math and coding ben
  • Qwen2.5-VL-72B — alibaba: Qwen2.5-VL-72B provides advanced visual understanding including document parsing, video comprehension and object grounding. It supports a 128k context window an
  • Grok 4.3 — xai: Grok 4.3 by xAI is a highly efficient conversational model designed to handle massive datasets with its expansive 1-million-token context window. It offers a hi
  • Veo 3.1 — google: Google's Veo 3.1 is a cutting-edge generative AI model that produces high-quality video with native, synchronized audio. Supporting resolutions up to 1080p, it
  • Qwen3-235B-A22B — alibaba: Qwen3-235B-A22B is a mixture-of-experts model with 235B total/22B active parameters, supporting seamless switching between thinking and non-thinking modes. It s
  • Qwen 3 235B — alibaba: Alibaba's Qwen 3 235B is an exceptionally powerful open-weight model optimized for elite multilingual understanding and advanced programming tasks. Built on a m
  • Gemini Embedding 2 — google: Google's Gemini Embedding 2 is a state-of-the-art multimodal embedding model that maps text, images, video, audio, and PDF data into a single, unified vector sp
  • Claude Sonnet 5 — anthropic: Claude Sonnet 5 by Anthropic delivers an exceptional balance of speed and high-tier intelligence, specifically optimized for advanced chat and complex coding ta
  • GPT-5.6 Luna — openai: GPT-5.6 Luna by OpenAI is a highly efficient chat-based model designed specifically to handle high-volume, cost-sensitive workloads without compromising on mode
  • Claude Sonnet 4.5 — anthropic: Claude 4.5 Sonnet is Anthropic's premier mid-tier model designed to deliver elite-level reasoning, coding, and comprehension capabilities. It excels at deeply u
  • Grok 4.5 — xai: Grok 4.5 is xAI's frontier model engineered specifically for advanced coding, complex reasoning, and agentic workflows. Featuring a massive 500k token context w
  • GPT-5.6 Terra — openai: GPT-5.6 Terra is a balanced powerhouse within OpenAI's latest model family, engineered to deliver a cost-effective blend of advanced intelligence and efficiency
  • Claude Fable 5 — anthropic: Claude Fable 5 is Anthropic's premier model designed to power highly sophisticated, autonomous, long-running agents. It features a massive 1M-token context wind
  • Claude Opus 5 — anthropic: Claude Opus 5 is Anthropic's flagship model designed specifically for complex agentic coding and heavy enterprise workloads. Featuring an expansive 1-million-to
  • GPT-Realtime 2.1 — openai: GPT-Realtime 2.1 by OpenAI represents a major leap in conversational AI, offering native speech-to-speech interaction with extremely low latency. Combining adva
  • DeepSeek V4 Pro — deepseek: DeepSeek V4 Pro is an advanced mixture-of-experts model engineered for high-performance chat and coding tasks, leveraging 1.6 trillion total parameters with onl
  • Claude Haiku 4.5 — anthropic: Claude Haiku 4.5 is Anthropic's fastest and most cost-effective model, optimized for near-instantaneous response times. It delivers exceptional performance on h
  • GPT-5.6 Sol — openai: GPT-5.6 Sol represents OpenAI's absolute pinnacle of machine intelligence, specifically engineered to tackle ultra-complex reasoning challenges and autonomous a
  • GPT-5.4 mini — openai: GPT-5.4 mini is OpenAI's highly efficient and cost-effective model optimized for high-volume automated tasks. It excels in driving specialized coding sub-agents
  • GPT-5.4 nano — openai: GPT-5.4 nano is OpenAI's highly optimized, lightweight model designed for high-throughput, low-latency text processing tasks. Positioned as the most cost-effect
  • Gemini 3.5 Flash — google: Google's Gemini 3.5 Flash is a highly optimized, cost-effective model designed specifically for high-speed, high-volume conversational tasks and multi-turn agen
  • Gemini 3.1 Pro (Preview) — google: Gemini 3.1 Pro is a highly advanced multimodal reasoning model from Google, specifically optimized for complex chat interactions and deep analytical coding. Equ
  • Nano Banana 2 (Gemini 3.1 Flash Image) — google: Nano Banana 2 (Gemini 3.1 Flash Image) is Google's highly efficient model designed for rapid image generation and editing. It offers an exceptionally cost-effec
  • Mistral Large 3 — mistral: Mistral Large 3 is Mistral AI's premier European frontier model, engineered to deliver top-tier reasoning, advanced coding capabilities, and highly sophisticate
  • FLUX.2 [pro] — black-forest-labs: FLUX 2 Pro by Black Forest Labs is a top-tier text-to-image model engineered engineered specifically for elite-level photorealism and precise visual rendering.
  • FLUX.2 [dev] — black-forest-labs: FLUX 2 Dev is an advanced open-weights image generation model developed by Black Forest Labs, designed to deliver state-of-the-art visual quality and prompt adh
  • Qwen3.5-397B-A17B — alibaba: Alibaba's Qwen3.5-397B-A17B is the debut model of the Qwen3.5 family, leveraging a massive Mixture of Experts architecture with 397 billion total and 17 billion
  • text-embedding-3-large — openai: Text-embedding-3-large is OpenAI's flagship embedding model, engineered to provide highly accurate vector representations for complex semantic search and retrie
  • Sora 2 — openai: Sora 2 is OpenAI's cutting-edge text-to-video generator that seamlessly integrates high-fidelity synchronized audio directly into its video outputs. Building up
  • Llama 4 Maverick — meta: Llama 4 Maverick is a state-of-the-art natively multimodal Mixture-of-Experts (MoE) model developed by Meta, engineered to excel in high-performance chat and co
  • Gemini 3.6 Flash — google: Gemini 3.6 Flash is Google's premier speed-optimized model, engineered to deliver rapid responses without sacrificing deep intelligence or coding capabilities.
  • Llama 4 Scout — meta: Llama 4 Scout by Meta is an efficient Mixture of Experts (MoE) chat model featuring 109 billion total and 17 billion active parameters. It stands out in the lan
  • Grok 3 — xai: Grok 3 introduced 'Think' mode for extended reasoning and DeepSearch for web-integrated answers, with a 131k context window. It was trained on xAI's Colossus su
  • Command A — cohere: Command A is Cohere's most capable model, optimized for enterprise use cases like RAG, tool use and agents, with a 256k context window. It runs efficiently on j
  • Embed v4 — cohere: Embed v4 generates unified embeddings for text, images and mixed documents (like PDFs with charts), supporting a 128k token context. It's designed for enterpris
  • Jamba 1.6 — ai21: Jamba 1.6 combines Mamba and Transformer architectures in a mixture-of-experts design, offering a 256k context window with efficient long-context inference. It
  • Phi-4 — microsoft: Phi-4 is a 14B parameter dense model trained with a focus on data quality, achieving strong performance on reasoning and math benchmarks relative to its size. I
  • Phi-4-multimodal — microsoft: Phi-4-multimodal is a 5.6B parameter model that unifies text, image and audio understanding in a single small model with a 128k context window. It's released un
  • Nemotron-4 340B — nvidia: Nemotron-4 340B is a dense 340B parameter model optimized to generate high-quality synthetic training data for other LLMs. It's released under the NVIDIA Open M