AgMoDB
ModelsAgentsEvalsCompositesVisualizeIndustry
AgMoDB by @mistakeknot

Model picks

Current defaults by use case.

Reviewed Jul 24, 2026 · Benchmark, price, speed, and human preference.

Product

Production assistants and internal tools.

Default
Claude Sonnet 5 (Adaptive Reasoning, Max Effort)

Anthropic

AgMoBench 53.2$4.00/M79 tok/sEarly evidence

Current Sonnet balance of agency, quality, and cost.

View model
Value
Gemini 3.5 Flash-Lite

Google

AgMoBench 53.1$0.850/M351 tok/sEarly evidence

Fast, low-cost agentic throughput.

View model
Ceiling
Claude Opus 5 (Adaptive Reasoning, Max Effort)

Anthropic

AgMoBench 68.5$10.00/M56 tok/sEarly evidence

Highest-confidence production work.

View model
Browse all modelsCompare picks

Human frontier

See all
1Anthropic: Claude Opus 4.7AnthropicHuman Frontier 95.5$10.00/M—2Claude Opus 4.6 (Non-reasoning, High Effort)AnthropicHuman Frontier 95.2$10.00/M—3GLM-5.2 (max)Z AIHuman Frontier 94.5$2.15/M105 tok/s4Claude Sonnet 4.6 (Non-reasoning, High Effort)AnthropicHuman Frontier 93.4$6.00/M—5GLM-5.1 (Reasoning)Z AIHuman Frontier 92.9$2.13/M—6Claude Opus 5 (Adaptive Reasoning, High Effort)AnthropicHuman Frontier 92.8$10.00/M52 tok/s

Worth discovering

Frontier value

Kimi K3 (max)

Kimi

New open-frontier pressure.

Cheap reasoning

GLM-5.2 (max)

Z AI

Recent reasoning price/performance.

Fast batch work

Gemini 3.5 Flash-Lite

Google

Fast, cheap high-throughput lane.

Open pressure

Qwen3.7 Plus

Alibaba

Open-ish frontier compression.