Skip to content

Compare flagship and efficient AI models for journalism — context windows, indicative token pricing, and newsroom workflow tags. Sources refresh from OpenRouter and Groq.

ROI calculator →Tools for journalism →Policy templates →

Could not refresh Groq prices from groq.com/pricing (Could not parse any pricing rows from Groq pricing page (markup may have changed).). Groq rows use bundled catalog fallback until the next sync.

Showing 21 of 48 in this view

Live

GPT-OSS 120B

API / route id

gpt-oss-120b

Open-weight workhorse when you self-host or run on fast inference (e.g. Groq).

Best forDrafting & rewriting · Structured extraction & fact packs · Engineering & automation
In / 1M tok$0.030
Out / 1M tok$0.170
Context131k tokens
Live

Meta: Muse Glimmer 30B

API / route id

muse-glimmer-30b

Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for autonomous agents on consumer hardware.

Best forDrafting & rewriting · Summaries & briefs · Engineering & automation
In / 1M tok$0.350
Out / 1M tok$1.50
Context131k tokens
Live

Meta: Muse Spark 1.2

API / route id

muse-spark-1.2

Muse Spark 1.2 is a reasoning model from Meta, designed for complex agentic tasks. It accepts text, images, video, audio, and PDF documents, returns text, and offers a 1M-token context window. Modality: text+image+file+audio+video->text. Inputs: text, image, video, file, audio. Outputs: text

Best forMultimodal & rich formats · Long documents & transcripts · Structured extraction & fact packs · Summaries & briefs
In / 1M tok$1.25
Out / 1M tok$4.25
Context1.0M tokens
Live

Qwen3 Coder Next

API / route id

qwen3-coder-next

Open-weight coding specialist for CI, scrapers, CMS integrations, and internal bots.

Best forEngineering & automation · Structured extraction & fact packs · Tagging, routing & metadata
In / 1M tok$0.120
Out / 1M tok$0.800
Context262k tokens

GPT-5 Chat

API / route id

gpt-5-chat

Low-latency conversational UX — alerts, briefs, and interactive assistants.

Best forChat & newsroom assistants · Summaries & briefs
In / 1M tok$1.25
Out / 1M tok$10.00
Context128k tokens
Live

Meta: Muse Spark 1.1

API / route id

muse-spark-1.1

Muse Spark 1.1 is a multimodal reasoning model from Meta, built for agentic tasks. It accepts text, images, video, audio, and PDF documents and returns text, with a 1M-token context... Modality: text+image+file+audio+video->text. Inputs: text, image, video, file, audio. Outputs: text

Best forMultimodal & rich formats · Long documents & transcripts · Structured extraction & fact packs · Summaries & briefs
In / 1M tok$1.25
Out / 1M tok$4.25
Context1.0M tokens
Live

Claude Opus 5

API / route id

claude-opus-5

Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis... Modality: text+image+file->text. Inputs: text, image, file. Outputs: text

Best forEngineering & automation · Multimodal & rich formats · Long documents & transcripts
In / 1M tok$5.00
Out / 1M tok$25.00
Context1.0M tokens
List price

Whisper V3 Large

API / route id

whisper-large-v3

Whisper Large v3 is OpenAI's most advanced and capable speech recognition model, delivering state-of-the-art accuracy across a wide range of audio conditions and languages. This flagship model excels at handling challenging audio scenarios including background noise, accents, and technical terminology.

Best forVoice & audio workflows · Structured extraction & fact packs
In / 1M tok
Out / 1M tok
STT / audio hr$0.111
Context
List price

GPT-OSS 120B (high)

API / route id

openai/gpt-oss-120b

Highest intelligence on Groq — complex analysis with very fast tokens/sec.

Best forInvestigations & deep research · Drafting & rewriting · Structured extraction & fact packs
In / 1M tok$0.150
Out / 1M tok$0.600
Speed (tok/s)500
Context131k tokens
List price

GPT-OSS 20B (high)

API / route id

openai/gpt-oss-20b

Fastest throughput for lightweight drafting and realtime tools.

Best forDrafting & rewriting · Chat & newsroom assistants
In / 1M tok$0.075
Out / 1M tok$0.300
Speed (tok/s)1,000
Context131k tokens
List price

Whisper Large V3 Turbo

API / route id

whisper-large-v3-turbo

Fast multilingual speech-to-text on Groq — interviews, briefings, and rough-cut transcripts at low cost per audio hour.

Best forVoice & audio workflows · Long documents & transcripts · Summaries & briefs
In / 1M tok
Out / 1M tok
STT / audio hr$0.040
Context
List price

LightOnOCR-2-1B

API / route id

lightonai/LightOnOCR-2-1B

LightOnOCR-2 is an efficient end-to-end 1B-parameter vision-language model for converting documents (PDFs, scans, images) into clean, naturally ordered text without relying on brittle pipelines.

Best forStructured extraction & fact packs · Long documents & transcripts
In / 1M tok
Out / 1M tok
Per page (1000 pages)$0.010
Context
Live

Gemma 3 27B Instruct

API / route id

gemma-3-27b-it

Open-weight Google model for self-hosted or budget-conscious pipelines and tooling.

Best forEngineering & automation · Tagging, routing & metadata · Summaries & briefs · EU / compliance-first workflows
In / 1M tok$0.080
Out / 1M tok$0.450
Context131k tokens
List price

FLUX.2 Klein 4B

API / route id

flux.2-klein-4b

Fastest, most cost-effective FLUX.2 tier — high-throughput thumbnails, social crops, and batch image generation.

Best forStills, graphics & layouts · Tagging, routing & metadata · Drafting & rewriting
In / 1M tok
Out / 1M tok
Image (per MP)$0.014
Image tiersFirst MP $0.014; each additional MP $0.001
Context41k tokens
Live

Kimi K2.5

API / route id

kimi-k2.5

Agentic coding and large-repo work — strong for newsroom engineering and data tooling.

Best forEngineering & automation · Structured extraction & fact packs · Drafting & rewriting
In / 1M tok$0.570
Out / 1M tok$2.85
Context262k tokens

GPT-5.6 Luna Pro

API / route id

gpt-5.6

Reasoning model for higher-quality responses on complex tasks.

Best forInvestigations & deep research · Structured extraction & fact packs · Engineering & automation
In / 1M tok$1.00
Out / 1M tok$6.00
Context1.0M tokens
Live

Claude Fable 5

API / route id

claude-fable-5

Claude Fable 5 is state-of-the-art on nearly all tested benchmarks of AI capability, showing exceptional performance in software engineering, knowledge work, vision, scientific research, and many other areas. The longer and more complex the task, the larger Fable 5’s lead over our other models. Modality: text+image+file->text. Inputs: text, image, file. Outputs: text

Best forStructured extraction & fact packs · Long documents & transcripts · Multimodal & rich formats · Investigations & deep research · Engineering & automation
In / 1M tok$10.00
Out / 1M tok$50.00
Context1.0M tokens
Live

Anthropic: Claude Sonnet 5

API / route id

claude-sonnet-5

Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels. Modality: text+image+file->text. Inputs: text, image, file. Outputs: text

Best forDrafting & rewriting · Summaries & briefs · Structured extraction & fact packs · Tagging, routing & metadata · Copy editing & style
In / 1M tok$2.00
Out / 1M tok$10.00
Context1.0M tokens
List price

Mistral OCR 4

API / route id

mistral-ocr-latest

Mistral OCR 4 is a document-intelligence model that turns PDFs and other files into structured, richly annotated content instead of plain text. It extracts text with bounding boxes, block types, and confidence scores, enabling better RAG chunking, agentic workflows (like form filling or invoice processing), and robust ingestion/search pipelines across 170 languages.

Best forStructured extraction & fact packs · Tagging, routing & metadata · Long documents & transcripts · Investigations & deep research
In / 1M tok
Out / 1M tok
Context

GPT-4o Audio

API / route id

gpt-4o-audio-preview

GPT-4–class audio conversations when quality matters more than Mini pricing.

Best forVoice & audio workflows · Chat & newsroom assistants · Drafting & rewriting
In / 1M tok$2.50
Out / 1M tok$10.00
Audio / 1M tok$40.00
Context128k tokens
Live

Claude Opus 4.8

API / route id

claude-opus-4.8

Model is designed for complex reasoning, advanced coding, and long-running AI workflows. With support for text, images, files, and a 1M-token context window, it excels at autonomous agents, large codebases, building presentations, multi-step problem solving, document creation, data analysis, and other knowledge-intensive tasks that require maintaining coherence across long sessions.

Best forLong documents & transcripts · Investigations & deep research · Drafting & rewriting
In / 1M tok$5.00
Out / 1M tok$25.00
Context1.0M tokens

Showing 21 of 48 in this snapshot · 48 total in the live catalog.