MODEL REFERENCE · OCTOBER 4, 2026

Compare the models.

Choose up to three. Start with the task and access you need, then compare capability and cost.

These are dated reference profiles. Benchmark leaders refresh daily; provider links show the latest availability and pricing. An app subscription does not automatically include API credit.

13 models · 0/3 selected

DeepSeek

DeepSeek V4.1 Flash

General & coding

A multimodal model with reasoning and tool calling. Compare accuracy and total task cost rather than token prices alone.

Access
DeepSeek API as deepseek-flash
API / 1M tokens
$0.30 in / $1.20 out
AA index · max
39 Oct 4 reference
Context, pricing & sources

1M tokens. Peak uncached-input/output rates shown. Off-peak rates are $0.15 / $0.60; cached inputs cost less. See the provider schedule.

Checked 2026-10-04. Provider documentation ↗ · Benchmark source ↗

Platform overview →
Moonshot AI

Kimi K3

General & coding

A model for long-context work, coding and vision, with low, high and max reasoning settings.

Access
Kimi API
API / 1M tokens
$3.00 in / $15.00 out
AA index · max
44 Oct 4 reference
Context, pricing & sources

1M tokens. Uncached input/output rates; cached input is $0.30 per million tokens.

Checked 2026-10-04. Provider documentation ↗ · Benchmark source ↗

Platform overview →
Anthropic

Claude Opus 5.5

General & coding

A flagship option for demanding analysis and coding. Compare its additional cost with Sonnet on work you can verify.

Access
Claude and API; plan limits apply
API / 1M tokens
$4.00 in / $20.00 out
AA index · max with fallback
58 Oct 4 reference
Context, pricing & sources

1M tokens. Standard API rates; fast mode and caching differ.

Checked 2026-10-04. Provider documentation ↗ · Benchmark source ↗

Platform overview →
Anthropic

Claude Sonnet 5.5

General & coding

A lower-cost Claude option for regular coding and analysis. Its high-effort results make it worth comparing before choosing Opus.

Access
Claude and API; plan limits apply
API / 1M tokens
$2.00 in / $10.00 out
AA index · max with fallback
56 Oct 4 reference
Context, pricing & sources

See model documentation. Standard API rates; caching differs.

Checked 2026-10-04. Provider documentation ↗ · Benchmark source ↗

Platform overview →
OpenAI

GPT-6 Astra

General & coding

OpenAI’s highest-capability option for difficult reasoning, coding and tool-driven work.

Access
API; app access depends on plan
API / 1M tokens
$10.00 in / $50.00 out
AA index · max
53 Oct 4 reference
Context, pricing & sources

1.05M tokens. Standard text-token rates; inputs over 272K cost more. Tools, caching and faster service have separate rates.

Checked 2026-10-04. Provider documentation ↗ · Benchmark source ↗

Platform overview →
OpenAI

GPT-6.1 Sol

General & coding

A less expensive alternative to Astra for coding and professional work. Test whether the difference matters on your tasks.

Access
API; app access depends on plan
API / 1M tokens
$2.00 in / $10.00 out
AA index · max
52 Oct 4 reference
Context, pricing & sources

1.05M tokens. Standard text-token rates; inputs over 272K cost more. Tools, caching and faster service have separate rates.

Checked 2026-10-04. Provider documentation ↗ · Benchmark source ↗

Platform overview →
OpenAI

GPT-6 Luna

High volume

A low-cost option for focused tasks, extraction and larger request volumes.

Access
API
API / 1M tokens
$0.10 in / $0.50 out
AA index · max
38 Oct 4 reference
Context, pricing & sources

1.05M tokens. Standard text-token rates; inputs over 272K cost more. Tools, caching and faster service have separate rates.

Checked 2026-10-04. Provider documentation ↗ · Benchmark source ↗

Platform overview →
xAI

Grok 4.7

General & coding

Text and image understanding with adjustable reasoning effort. The Grok app adds search, voice and other tools beyond the model itself.

Access
Grok and API; plan limits apply
API / 1M tokens
$2.00 in / $6.00 out
AA index · xhigh
46 Oct 4 reference
Context, pricing & sources

500K tokens. Base API rates up to 200K input tokens; longer context and regional processing cost more.

Checked 2026-10-04. Provider documentation ↗ · Benchmark source ↗

Platform overview →
Google

Gemini 3.8 Flash

General & coding

Google’s current Flash model for agent workflows and software engineering. Useful to compare with smaller models when latency matters.

Access
Google AI Studio and Gemini API
API / 1M tokens
$0.75 in / $3.75 out
AA index · high
41 Oct 4 reference
Context, pricing & sources

See model documentation. Promotional standard rates through Dec 31, 2026; $1.50 input / $7.50 output from Jan 1, 2027. Grounding and caching cost extra.

Checked 2026-10-04. Provider documentation ↗ · Benchmark source ↗

Platform overview →
Google

Gemini 3.5 Flash-Lite

High volume

A lightweight model aimed at throughput and cost-sensitive tasks. Validate extraction quality before increasing volume.

Access
Google AI Studio and Gemini API
API / 1M tokens
$0.30 in / $2.50 out
Context, pricing & sources

See model documentation. Standard API rates; grounding and cache storage have separate charges.

Checked 2026-10-04. Provider documentation ↗

Platform overview →
Google

Gemini 4 Argon

Restricted research

Announced for a restricted rollout to trusted cyber defenders. Strong benchmark results are not an invitation to subscribe for access.

Access
Restricted rollout through Fairwind
API / 1M tokens
See provider pricing
AA index · high
53 Oct 4 reference
Context, pricing & sources

1M tokens. No broadly available price listed here.

Checked 2026-10-04. Provider documentation ↗ · Benchmark source ↗

Platform overview →
OpenAI

GPT Image 2.5 Sunburst

Image generation

An image model to consider for visual work. Compare text rendering, edits and composition using the image leaderboard and your own brief.

Access
Check current model availability
API / 1M tokens
See provider pricing
Context, pricing & sources

Image workflow. Image costs depend on dimensions and quality; text-token rates are not a useful comparison.

Checked 2026-10-04. Provider documentation ↗

Platform overview →
OpenAI

GPT Image 2.5 Flare

Image generation

Another current OpenAI image model. Use the same prompt and output size when comparing with Sunburst.

Access
Check current model availability
API / 1M tokens
See provider pricing
Context, pricing & sources

Image workflow. See current image pricing and output settings.

Checked 2026-10-04. Provider documentation ↗

Platform overview →

Model, app or agent?

Model

The engine that generates an answer. A model’s context window and reasoning settings affect what it can process.

App & tools

Search, files, connectors and editing features around that engine. Two apps using the same model can behave very differently.

Agent

A workflow that can take multiple steps and use tools. Finished-task quality depends on instructions, permissions and checks as well as the model.

Compare platform features →