Skip to content

Compare flagship and efficient AI models for journalism — context windows, indicative token pricing, and newsroom workflow tags. Sources refresh from OpenRouter and Groq.

ROI calculator →Tools for journalism →Policy templates →

Showing 21 of 50 in this view

Live

Qwen3 Coder Next

API / route id

qwen3-coder-next

Open-weight coding specialist for CI, scrapers, CMS integrations, and internal bots.

Best forEngineering & automation · Structured extraction & fact packs · Tagging, routing & metadata
In / 1M tok$0.120
Out / 1M tok$0.800
Context262k tokens

GPT-5 Chat

API / route id

gpt-5-chat

Low-latency conversational UX — alerts, briefs, and interactive assistants.

Best forChat & newsroom assistants · Summaries & briefs
In / 1M tok$1.25
Out / 1M tok$10.00
Context128k tokens
Live

Meta: Muse Spark 1.1

API / route id

muse-spark-1.1

Muse Spark 1.1 is a multimodal reasoning model from Meta, built for agentic tasks. It accepts text, images, video, audio, and PDF documents and returns text, with a 1M-token context... Modality: text+image+file+audio+video->text. Inputs: text, image, video, file, audio. Outputs: text

Best forMultimodal & rich formats · Long documents & transcripts · Structured extraction & fact packs · Summaries & briefs
In / 1M tok$1.25
Out / 1M tok$4.25
Context1.0M tokens
Live

Claude Opus 5

API / route id

claude-opus-5

Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis... Modality: text+image+file->text. Inputs: text, image, file. Outputs: text

Best forEngineering & automation · Multimodal & rich formats · Long documents & transcripts
In / 1M tok$5.00
Out / 1M tok$25.00
Context1.0M tokens
Live

Whisper V3 Large

API / route id

whisper-large-v3

Whisper Large v3 is OpenAI's most advanced and capable speech recognition model, delivering state-of-the-art accuracy across a wide range of audio conditions and languages. This flagship model excels at handling challenging audio scenarios including background noise, accents, and technical terminology.

Best forVoice & audio workflows · Structured extraction & fact packs
In / 1M tok
Out / 1M tok
STT / audio hr$0.111
Context
Live

GPT-OSS 120B (high)

API / route id

openai/gpt-oss-120b

Highest intelligence on Groq — complex analysis with very fast tokens/sec.

Best forInvestigations & deep research · Drafting & rewriting · Structured extraction & fact packs
In / 1M tok$0.150
Out / 1M tok$0.600
Speed (tok/s)500
Context131k tokens
Live

GPT-OSS 20B (high)

API / route id

openai/gpt-oss-20b

Fastest throughput for lightweight drafting and realtime tools.

Best forDrafting & rewriting · Chat & newsroom assistants
In / 1M tok$0.075
Out / 1M tok$0.300
Speed (tok/s)1,000
Context131k tokens
Live

😱 DEPRICATED July 17. Llama 3.1 8B

API / route id

llama-3.1-8b-instant

Cheapest Groq tier — triage, language detection, simple formatting.

Best forTagging, routing & metadata · Summaries & briefs
In / 1M tok$0.050
Out / 1M tok$0.080
Speed (tok/s)840
Context131k tokens
Live

😱 DEPRICATED July 17. Llama 3.3 70B

API / route id

llama-3.3-70b-versatile

General-purpose Groq default when you want open-model economics.

Best forDrafting & rewriting · Summaries & briefs · Structured extraction & fact packs
In / 1M tok$0.590
Out / 1M tok$0.790
Speed (tok/s)394
Context131k tokens
Live

Whisper Large V3 Turbo

API / route id

whisper-large-v3-turbo

Fast multilingual speech-to-text on Groq — interviews, briefings, and rough-cut transcripts at low cost per audio hour.

Best forVoice & audio workflows · Long documents & transcripts · Summaries & briefs
In / 1M tok
Out / 1M tok
STT / audio hr$0.040
Context
List price

LightOnOCR-2-1B

API / route id

lightonai/LightOnOCR-2-1B

LightOnOCR-2 is an efficient end-to-end 1B-parameter vision-language model for converting documents (PDFs, scans, images) into clean, naturally ordered text without relying on brittle pipelines.

Best forStructured extraction & fact packs · Long documents & transcripts
In / 1M tok
Out / 1M tok
Per page (1000 pages)$0.010
Context
Live

Gemma 3 27B Instruct

API / route id

gemma-3-27b-it

Open-weight Google model for self-hosted or budget-conscious pipelines and tooling.

Best forEngineering & automation · Tagging, routing & metadata · Summaries & briefs · EU / compliance-first workflows
In / 1M tok$0.080
Out / 1M tok$0.450
Context131k tokens
List price

FLUX.2 Klein 4B

API / route id

flux.2-klein-4b

Fastest, most cost-effective FLUX.2 tier — high-throughput thumbnails, social crops, and batch image generation.

Best forStills, graphics & layouts · Tagging, routing & metadata · Drafting & rewriting
In / 1M tok
Out / 1M tok
Image (per MP)$0.014
Image tiersFirst MP $0.014; each additional MP $0.001
Context41k tokens
Live

Kimi K2.5

API / route id

kimi-k2.5

Agentic coding and large-repo work — strong for newsroom engineering and data tooling.

Best forEngineering & automation · Structured extraction & fact packs · Drafting & rewriting
In / 1M tok$0.570
Out / 1M tok$2.85
Context262k tokens
Live

GPT-OSS 120B

API / route id

gpt-oss-120b

Open-weight workhorse when you self-host or run on fast inference (e.g. Groq).

Best forDrafting & rewriting · Structured extraction & fact packs · Engineering & automation
In / 1M tok$0.037
Out / 1M tok$0.170
Context131k tokens

GPT-5.6 Luna Pro

API / route id

gpt-5.6

Reasoning model for higher-quality responses on complex tasks.

Best forInvestigations & deep research · Structured extraction & fact packs · Engineering & automation
In / 1M tok$1.00
Out / 1M tok$6.00
Context1.0M tokens
Live

Claude Fable 5

API / route id

claude-fable-5

Claude Fable 5 is state-of-the-art on nearly all tested benchmarks of AI capability, showing exceptional performance in software engineering, knowledge work, vision, scientific research, and many other areas. The longer and more complex the task, the larger Fable 5’s lead over our other models. Modality: text+image+file->text. Inputs: text, image, file. Outputs: text

Best forStructured extraction & fact packs · Long documents & transcripts · Multimodal & rich formats · Investigations & deep research · Engineering & automation
In / 1M tok$10.00
Out / 1M tok$50.00
Context1.0M tokens
Live

Anthropic: Claude Sonnet 5

API / route id

claude-sonnet-5

Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels. Modality: text+image+file->text. Inputs: text, image, file. Outputs: text

Best forDrafting & rewriting · Summaries & briefs · Structured extraction & fact packs · Tagging, routing & metadata · Copy editing & style
In / 1M tok$2.00
Out / 1M tok$10.00
Context1.0M tokens
List price

Mistral OCR 4

API / route id

mistral-ocr-latest

Mistral OCR 4 is a document-intelligence model that turns PDFs and other files into structured, richly annotated content instead of plain text. It extracts text with bounding boxes, block types, and confidence scores, enabling better RAG chunking, agentic workflows (like form filling or invoice processing), and robust ingestion/search pipelines across 170 languages.

Best forStructured extraction & fact packs · Tagging, routing & metadata · Long documents & transcripts · Investigations & deep research
In / 1M tok
Out / 1M tok
Context
List price

😱 DEPRICATED July 17. Llama 4 Scout 17B Instruct

API / route id

meta-llama/llama-4-scout-17b-16e-instruct

Groq-hosted Llama 4 Scout instruct: OCR, charts, and image understanding with 10M-token context for docs and AV material.

Best forMultimodal & rich formats · Structured extraction & fact packs · Long documents & transcripts · Investigations & deep research
In / 1M tok$0.110
Out / 1M tok$0.340
Speed (tok/s)594
Context10.0M tokens
List price

😱 DEPRICATED July 17. Qwen3 32B

API / route id

qwen3-32b

Reasoning + tool use at mid cost — good for structured extraction.

Best forStructured extraction & fact packs · Engineering & automation
In / 1M tok$0.290
Out / 1M tok$0.590
Speed (tok/s)662
Context131k tokens

Showing 21 of 50 in this snapshot · 50 total in the live catalog.