AI INFRASTRUCTURE METRICS / SIMPLIFIED

Choose the right
engine for the job.

Cost, capability, and real-world signals translated into decisions your business team can act on.

344OpenRouter models tracked
218Human-rated Arena entries
DailyAutomated refresh cadence
GLOBAL INDUSTRY AI INDEXChoose a sectorEach vertical opens into high-intent business tasks.
Healthcare & Life Sciences
Finance & Wealth
§Legal & Compliance
Agriculture & Supply Chain
Manufacturing & Industrial IoT
Education & EdTech
E-Commerce & Retail
Aerospace & Cybersecurity
01 / TASK-BASED DIRECTORY

Best for clinical documentation

Low-latency transcription and reliable context for regulated workflows.

OpenRouter pricing · LMSYS human preference
01tencent

Tencent: Hy3 (free)

Hy3 is a 295B-parameter Mixture-of-Experts model from Tencent (21B active, 192 experts with top-8 ro

BLENDED COST$0.00/ 1M tokens
CONTEXT262ktokens
ARENA ELOhuman score
Basic classification
tencent/hy3:free
// Cost-aware generation
const result = await ai.generateContent(prompt);
// Track usage and quality
02poolside

Poolside: Laguna XS 2.1 (free)

Laguna XS 2.1 is the latest coding agent model in the 33B-A3B category from [Poolside](https://pools

BLENDED COST$0.00/ 1M tokens
CONTEXT262ktokens
ARENA ELOhuman score
Basic classification
poolside/laguna-xs-2.1:free
// Cost-aware generation
const result = await ai.generateContent(prompt);
// Track usage and quality
03cohere

Cohere: North Mini Code (free)

North Mini Code is Cohere's first agentic coding model and the debut of its North family. A sparse m

BLENDED COST$0.00/ 1M tokens
CONTEXT256ktokens
ARENA ELOhuman score
Basic classification
cohere/north-mini-code:free
// Cost-aware generation
const result = await ai.generateContent(prompt);
// Track usage and quality
04nvidia

NVIDIA: Nemotron 3.5 Content Safety (free)

NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA,

BLENDED COST$0.00/ 1M tokens
CONTEXT128ktokens
ARENA ELOhuman score
Basic classification
nvidia/nemotron-3.5-content-safety:free
// Cost-aware generation
const result = await ai.generateContent(prompt);
// Track usage and quality
05nvidia

NVIDIA: Nemotron 3 Ultra (free)

NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B

BLENDED COST$0.00/ 1M tokens
CONTEXT1000ktokens
ARENA ELOhuman score
Basic classification
nvidia/nemotron-3-ultra-550b-a55b:free
// Cost-aware generation
const result = await ai.generateContent(prompt);
// Track usage and quality
06nvidia

NVIDIA: Nemotron 3 Nano Omni (free)

NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception

BLENDED COST$0.00/ 1M tokens
CONTEXT256ktokens
ARENA ELOhuman score
Basic classification
nvidia/nemotron-3-nano-omni-30b-a3b-reasoning:free
// Cost-aware generation
const result = await ai.generateContent(prompt);
// Track usage and quality
07poolside

Poolside: Laguna M.1 (free)

Laguna M.1 is the flagship coding agent model from [Poolside](https://poolside.ai/), optimized for c

BLENDED COST$0.00/ 1M tokens
CONTEXT262ktokens
ARENA ELOhuman score
Basic classification
poolside/laguna-m.1:free
// Cost-aware generation
const result = await ai.generateContent(prompt);
// Track usage and quality
08google

Google: Gemma 4 26B A4B (free)

Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Desp

BLENDED COST$0.00/ 1M tokens
CONTEXT262ktokens
ARENA ELOhuman score
Basic classification
google/gemma-4-26b-a4b-it:free
// Cost-aware generation
const result = await ai.generateContent(prompt);
// Track usage and quality
09google

Google: Gemma 4 31B (free)

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image inp

BLENDED COST$0.00/ 1M tokens
CONTEXT262ktokens
ARENA ELOhuman score
Basic classification
google/gemma-4-31b-it:free
// Cost-aware generation
const result = await ai.generateContent(prompt);
// Track usage and quality
10google

Google: Lyria 3 Pro Preview

Full-length songs are priced at $0.08 per song. Lyria 3 is Google's family of music generation model

BLENDED COST$0.00/ 1M tokens
CONTEXT1049ktokens
ARENA ELOhuman score
Basic classification
google/lyria-3-pro-preview
// Cost-aware generation
const result = await ai.generateContent(prompt);
// Track usage and quality
02 / HOW TO READ THIS

Business metrics,
without the noise.

Financial efficiency

Input and output prices from OpenRouter are normalized to a blended cost per million tokens.

Human preference

LMSYS Arena Elo adds a reality check from blind, human-voted model comparisons.

Operational fit

Context limits and provider availability help teams choose an engine that fits the actual workflow.