Marketplace

AI Model Marketplace

Search, filter, compare, and audit live model pricing with TVP buyer pricing.

443Models72ProvidersLivePricing
Showing 12 of 443 models
Select up to 3 live models to compare pricing, context, and endpoint coverage.
Anthropic

Claude Opus 5 (Fast)

Fast-mode variant of [Opus 5](/anthropic/claude-opus-5) - identical capabilities with higher output speed at 2x pricing relative to regular Opus 5. Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-mode

text+image+file->texttextimagefile
Input$10.50per 1M
Output$52.50per 1M
Context1M
Max output128,000
Details
Anthropic

Claude Opus 5

Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...

text+image+file->texttextimagefile
Input$5.25per 1M
Output$26.25per 1M
Context1M
Max output128,000
Details
Microsoft

Microsoft: MAI-Image-2.5 Pro

Microsoft's MAI-Image-2.5 is a high-quality image generation model available via Azure AI Foundry. It produces photorealistic and artistic images from text prompts with support for various aspect ratios.

text+image->imagetextimage
Input$5.25per 1M
Output$0.00per 1M
Context4.1K
Max output1,024
Details
Microsoft

Microsoft: MAI-Voice-2-Flash

MAI-Voice-2-Flash is a low-latency text-to-speech model from Microsoft for voice agents, assistants, call centers, accessibility, narration, and other interactive applications. It generates expressive 24 kHz mono speech across 15 languages...

text->speechtextspeech
Input$15.75per 1M
Output$0.00per 1M
Context0
Max outputn/a
Details
inclusionAI

Ling-3.0-flash (free)

*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...

text->texttexttools
Input$0.00per 1M
Output$0.00per 1M
Context262.1K
Max output32,768
Details
Qwen

Qwen: Qwen-Audio-3.0-TTS Flash

Qwen-Audio-3.0-TTS Flash is Alibaba's fast, cost-efficient text-to-speech model, generating spoken audio from text via the DashScope Speech Synthesizer API.

text->speechtextspeechstructured
Input$15.75per 1M
Output$0.00per 1M
Context0
Max outputn/a
Details
Qwen

Qwen: Qwen-Audio-3.0-TTS Plus

Qwen-Audio-3.0-TTS Plus is Alibaba's higher-quality text-to-speech model, generating spoken audio from text via the DashScope Speech Synthesizer API.

text->speechtextspeechstructured
Input$21.00per 1M
Output$0.00per 1M
Context0
Max outputn/a
Details
xAI

xAI: Grok STT 1.0

Grok STT is xAI's speech-to-text model, available via the REST /v1/stt endpoint. It supports transcription with word-level timestamps, optional speaker diarization, and multichannel audio.

audio->transcriptionaudiotranscriptionstructured
Input$105000.00per 1M
Output$0.00per 1M
Context0
Max outputn/a
Details
Poolside

Poolside: Laguna S 2.1

Laguna S 2.1 is the latest coding agent model from [Poolside](<https://poolside.ai/>). Laguna S 2.1 is a 118B total parameter model with 8B active parameters, scoring 70.2% on Terminal-Bench 2.1 and...

text->texttexttools
Input$0.105per 1M
Output$0.21per 1M
Context1M
Max output131,072
Details
Poolside

Poolside: Laguna S 2.1 (free)

Laguna S 2.1 is the latest coding agent model from [Poolside](<https://poolside.ai/>). Laguna S 2.1 is a 118B total parameter model with 8B active parameters, scoring 70.2% on Terminal-Bench 2.1 and...

text->texttexttools
Input$0.00per 1M
Output$0.00per 1M
Context262.1K
Max output32,768
Details
Google

Google: Gemini 3.6 Flash

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

text+image+file+audio+video->texttextimagevideo
Input$1.58per 1M
Output$7.88per 1M
Context1M
Max output65,536
Details
Google

Google: Gemini 3.5 Flash Lite

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

text+image+file+audio+video->texttextimagevideo
Input$0.315per 1M
Output$2.63per 1M
Context1M
Max output65,536
Details
Showing 12 of 443 models
Internal distribution

Use the marketplace with guide and comparison pages

These internal routes help buyers move from generic catalog browsing into comparison and buying-intent pages.

How TVP pricing works

Transparent pricing buyers can audit

Simple pricing rules, live source data, and exact decimals remain visible for review.

1

Provider pricing syncs live

TVP reads current provider catalog data and refreshes on demand.

2

TVP price is shown for planning

Buyers can compare options using one consistent TVP pricing view.

3

Prices are shown per 1M tokens

Tiny token decimals are normalized into buyer-readable rates.

4

Exact decimals stay available

Detailed pricing fields remain visible for audit and billing checks.

Last synced: 7/24/2026, 8:48:30 PM