Passer au contenu principal

Model Guides

43 guides

Explore Verdent guides for AI models, coding agents, benchmark interpretation, and practical model evaluation workflows.

Amazon Nova Pro: Features, Pricing, and Deployment

Model

Amazon Nova Pro: Features, Pricing, and Deployment

A practical guide to Amazon Nova Pro: multimodal inputs, Bedrock pricing, Nova Lite tradeoffs, deployment, coding evaluation, and Verdent access checks.

Claude Opus 4.6: Features, Pricing, and Coding Use

Model

Claude Opus 4.6: Features, Pricing, and Coding Use

Explore Claude Opus 4.6's 1M context, adaptive thinking, 128K output, agentic coding strengths, pricing, and successor options.

Claude Sonnet 4.6: Features, Pricing, and Agentic Coding

Model

Claude Sonnet 4.6: Features, Pricing, and Agentic Coding

Explore Claude Sonnet 4.6 for coding and agents, including its 1M context, published evaluation results, API pricing, and Verdent use.

DeepSeek R1-0528: Reasoning, Access, and Coding Use

Model

DeepSeek R1-0528: Reasoning, Access, and Coding Use

Explore DeepSeek R1-0528's reasoning update, benchmark conditions, open weights, coding evaluation, and current API and Verdent access limits.

DeepSeek R1: Reasoning, Pricing, and Coding Use

Model

DeepSeek R1: Reasoning, Pricing, and Coding Use

Explore DeepSeek R1's reasoning training, 128K open weights, MIT license, coding fit, historical API pricing, and current access limits.

DeepSeek V3.2: Features, Pricing, and API Use

Model

DeepSeek V3.2: Features, Pricing, and API Use

DeepSeek V3.2 explained: reasoning modes, MIT-licensed weights, historical API pricing, coding uses, and its current access status.

Gemini 2.5 Flash: Speed, Pricing, and Coding Use

Model

Gemini 2.5 Flash: Speed, Pricing, and Coding Use

Explore Gemini 2.5 Flash's 1M context, thinking controls, multimodal inputs, API pricing, coding fit, and Flash versus Pro tradeoffs.

Gemini 3.1 Pro: Features, Pricing, and Coding Use

Model

Gemini 3.1 Pro: Features, Pricing, and Coding Use

Explore Gemini 3.1 Pro Preview, including its 1M context, multimodal tools, tiered API pricing, coding uses, and Verdent access.

GLM-4.7-Flash: Features, Deployment, and Coding Use

Model

GLM-4.7-Flash: Features, Deployment, and Coding Use

Explore GLM-4.7-Flash's 30B-A3B MoE design, 200K context, MIT license, local deployment, API access, and coding use cases.

GLM-5: Features, Pricing, and Agentic Coding

Model

GLM-5: Features, Pricing, and Agentic Coding

Explore GLM-5's 200K context, MIT-licensed weights, API pricing, coding use cases, deployment choices, and Verdent access paths.

GPT-5: Features, Pricing, and Coding Use

Model

GPT-5: Features, Pricing, and Coding Use

Explore GPT-5's 400K context, reasoning controls, coding strengths, API pricing, family variants, and current migration considerations.

GPT-5.2: Features, Pricing, and Coding Use

Model

GPT-5.2: Features, Pricing, and Coding Use

Explore GPT-5.2's 400K context, reasoning controls, API pricing, coding use cases, and tradeoffs against GPT-5.4, GPT-5.5, and current models.

GPT-OSS 120B: Features, Deployment, and Use Cases

Model

GPT-OSS 120B: Features, Deployment, and Use Cases

Learn how GPT-OSS 120B works, how it differs from 20B, what an 80 GB deployment requires, and where this open-weight model fits.

Grok 3: Features, Pricing, and Use Cases

Model

Grok 3: Features, Pricing, and Use Cases

Explore Grok 3's 1M-token launch context, reasoning and coding use cases, API retirement, successor path, and access boundaries.

Kimi K2.5: Features, Pricing, and Use Cases

Model

Kimi K2.5: Features, Pricing, and Use Cases

Explore Kimi K2.5's 1T MoE design, 256K context, multimodal coding, Agent Swarm origins, open weights, and current access limits.

Llama 4 Scout: Context, Deployment, and Coding Use

Model

Llama 4 Scout: Context, Deployment, and Coding Use

Explore Llama 4 Scout's open-weight MoE design, 10M context boundary, deployment needs, coding evaluation, and choice against Maverick.

MiniMax M2.5: Features, Deployment, and Coding Use

Model

MiniMax M2.5: Features, Deployment, and Coding Use

Explore MiniMax M2.5's 204,800-token context, model license, API pricing, coding uses, deployment choices, and current Verdent status.

Token-Efficient Coding Models: Why API Prices Don’t Predict Real Project Cost

Model

Token-Efficient Coding Models: Why API Prices Don’t Predict Real Project Cost

Cheap API tokens can produce an expensive finished task. Benchmarks, agent runs, and user reports show when a premium coding model costs less by avoiding waste.

Claude Haiku 4.5

Model

Claude Haiku 4.5

A practical guide to Claude Haiku 4.5 — speed, pricing, best-fit coding tasks, and when to use it instead of Sonnet in Verdent.

Claude Sonnet 5

Model

Claude Sonnet 5

A complete guide to Claude Sonnet 5 — what's improved over Sonnet 4.6, coding benchmarks, agentic capabilities, and how to use it in Verdent for parallel coding tasks.

Devstral 2

Model

Devstral 2

A developer's guide to Devstral 2 — 72.2% SWE-bench Verified, multi-file editing, agentic coding, and how it compares to Devstral Small 2 and Claude Code.

Gemma 4

Model

Gemma 4

A complete guide to Gemma 4 — 26B MoE parameters, 85 tokens/second on consumer GPUs, Apache 2.0 license, and how it compares to Llama 4 Scout for local deployment.

Grok 4.1 Fast

Model

Grok 4.1 Fast

A migration guide to the retired Grok 4.1 Fast API slugs — historical pricing and tool use, redirect behavior, and how to validate a maintained replacement.

Grok Code Fast 1

Model

Grok Code Fast 1

A technical guide to Grok Code Fast 1 — xAI's dedicated coding model. Speed, SWE-bench scores, pricing, and how it compares to Devstral 2 and Claude Code.

Llama 4 Maverick

Model

Llama 4 Maverick

A complete guide to Llama 4 Maverick — Meta's open-source multimodal flagship with 1M context. Benchmarks, coding performance, and how it compares to Llama 4 Scout.

Mistral Large 3

Model

Mistral Large 3

A guide to Mistral Large 3 — 675B total and 41B active parameters, 256K context, multilingual support, pricing, and enterprise deployment tradeoffs.

Pixtral Large

Model

Pixtral Large

A migration guide to deprecated Pixtral Large — its multimodal capabilities, historical access, replacement options, and how to validate visual coding workflows.

Qwen3.5

Model

Qwen3.5

A complete guide to Qwen3.5 — Apache 2.0 licensed, frontier-class performance, and how it compares to Llama 4 and DeepSeek V3.2 for coding and agentic tasks.

Claude Opus 4.5

Model

Claude Opus 4.5

A complete guide to Claude Opus 4.5 — what's new, how it performs on coding tasks, and how it compares to Claude Opus 4.7 and GPT-5 for agentic workflows.

Codex CLI

Model

Codex CLI

A hands-on guide to OpenAI Codex CLI — how to install it, what it can do, and how it compares to Claude Code and Verdent for agentic software development.

Gemini 2.5 Pro

Model

Gemini 2.5 Pro

Everything you need to know about Gemini 2.5 Pro — coding benchmarks, 1M context window, pricing, and how to use it inside Verdent for agentic software development.

Gemini 3 Flash

Model

Gemini 3 Flash

A practical guide to Gemini 3 Flash — speed benchmarks, SWE-bench 78%, free tier access, and how to use it for fast agentic coding inside Verdent.

Gemini 3 Pro

Model

Gemini 3 Pro

Everything you need to know about Gemini 3 Pro — Deep Think mode, 1M context, SWE-bench 78%, and how to use it with Verdent for agentic coding workflows.

Gemini 3.5

Model

Gemini 3.5

A complete guide to the Gemini 3.5 series — starting with Gemini 3.5 Flash (I/O 2026). Outperforms Gemini 3.1 Pro on coding and agentic benchmarks, 4x faster, and Gemini 3.5 Pro confirmed coming next month.

Gemini Omni

Model

Gemini Omni

A first look at Gemini Omni — Google's new multimodal flagship announced at I/O 2026. Create anything from any input, with breakthrough video understanding, generation, and world modeling capabilities.

Gemma 3

Model

Gemma 3

A complete guide to Google's Gemma 3 — benchmarks, local deployment, and how it compares to Gemma 4 and Llama 4 for coding and agentic tasks.

Google Antigravity

Model

Google Antigravity

A complete breakdown of Google Antigravity — what it does, how it compares to Verdent and Cursor, and which AI coding tool is right for your workflow.

GPT-5.1 Codex

Model

GPT-5.1 Codex

A developer's guide to GPT-5.1 Codex — long-horizon coding tasks, API setup, real benchmarks, and how it compares to Claude Code and Verdent for agentic workflows.

GPT-OSS 20B

Model

GPT-OSS 20B

A practical guide to GPT-OSS 20B — OpenAI's first open-weight model. Benchmarks, deployment options, and how it compares to DeepSeek V3.2 and Llama 4.

Grok 4

Model

Grok 4

A complete guide to Grok 4 — benchmarks, coding capabilities, pricing, and how it compares to GPT-5 and Claude. See how Verdent uses Grok 4 for parallel agentic coding.

Grok 4.1

Model

Grok 4.1

A complete review of Grok 4.1 — LMArena #1 Elo rating, 65% lower hallucination rate, coding benchmarks, and how it compares to Claude Sonnet 4.6 and GPT-5.5.

Kimi K2 Thinking

Model

Kimi K2 Thinking

A deep dive into Kimi K2's Thinking mode — how it compares to DeepSeek R1, when to enable it, and how it powers agentic coding workflows inside Verdent.

Phi-4

Model

Phi-4

Everything about Microsoft Phi-4 — how a 14B model beats larger LLMs on reasoning, local deployment options, and when to use Phi-4 vs GPT-5 for your coding tasks.