ASCA06
MODEL DECISION PAGE

AI MODEL ATLAS · DECISION PAGE

Fastest LLM API models by OpenRouter order

OpenRouter models in the server-provided throughput-high-to-low order.

OBJECT IDasca:openrouter-model-page-fastest-llm-api@2026-08-27.v1
FETCHED
STATUSREPRODUCIBLE SNAPSHOT

Fastest LLM API models by OpenRouter order

Preserve server-provided order; do not interpret rank as a latency or tokens-per-second value.

OBSERVED MODELS · 10 SHOWN

Comparable fields from one dated catalog snapshot

RANK ≠ QUALITY
Rank / orderModelInput / 1MOutput / 1MRequestContextToolsStructured
1Morph: Morph V3 Fast$0.80$1.20$0.0081.92KNoNo
2Google: Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image)$0.25$1.50$0.0065.54KNoNo
3Meta: Llama 4 Scout$0.11$0.34$0.001.31MYesYes
4OpenAI: gpt-oss-safeguard-20b$0.075$0.30$0.00131.07KYesYes
5Qwen: Qwen3 Next 80B A3B Thinking$0.15$1.20$0.00262.14KYesYes
6Meta: Llama 3.2 1B Instruct$0.027$0.201$0.0060KNoNo
7Inception: Mercury 2$0.25$0.75$0.00128KYesYes
8Google: Nano Banana (Gemini 2.5 Flash Image)$0.30$2.50$0.0032.77KNoYes
9Google: Nano Banana 2 (Gemini 3.1 Flash Image Preview)$0.50$3.00$0.0065.54KNoYes
10Qwen: Qwen3 30B A3B$0.12$0.50$0.00131.07KYesNo

Prices are listed USD per token converted to USD per 1M tokens. Request fee is a separate listed field; it is not included in derived text-token cost. A free listing means zero listed token price in this snapshot, not unlimited access.

  • Observed range: 417 matching catalog records; 10 shown.
  • Bias: OpenRouter measurement and catalog bias.
  • Limitation: No numeric throughput, latency, hardware, prompt size, or reproducibility claim is made.
  • Confidence: medium for the recorded response and stated transformation.

DOWNLOAD · VERIFY · CITE

Keep the snapshot and object ID with the decision.

Parent model atlas

Preserve server-provided order; do not interpret rank as a latency or tokens-per-second value.asca:openrouter-model-page-fastest-llm-api@2026-08-27.v1