Marketplace

AI Model Marketplace

Search, filter, compare, and audit live model pricing with TVP buyer pricing.

653Models87ProvidersLivePricing
Showing 12 of 653 models
Select up to 3 live models to compare pricing, context, and endpoint coverage.
Microsoft

Microsoft: Microsoft-Decision-1

Microsoft-Decision-1 is a small model built for fast decision-making. Instead of generating text, it reads the provided content and returns a calibrated probability for each fixed answer option, so the...

text->decisionstextdecisions
Input$0.0441per 1M
Output$0.00per 1M
Context32.8K
Max output29,491
Details
nace-ai

Nace.AI: Drex v1.5

Drex 1.5 is a small decision model from Nace AI, served over the same /v1/systemone contract as TypeSafe's Jev. Send a state and typed questions (yes/no, multiple choice, or score)...

text->decisionstextdecisions
Input$0.042per 1M
Output$0.00per 1M
Context131.1K
Max output117,964
Details
cloudflare

Cloudflare: Clef Omni

Clef Omni is the mixture-of-experts member of Cloudflare's open-source Clef decision model family, a fine-tune of Qwen3-Omni-30B-A3B (30B total, 3B active parameters) served on Workers AI. It turns a state...

text+image->decisionstextimagedecisions
Input$0.1575per 1M
Output$0.00per 1M
Context65.5K
Max output58,982
Details
StepFun

StepFun: Step 5 Preview

Step 5 Preview is StepFun's flagship model for agentic work, built on a sparse Mixture-of-Experts architecture (27B active / 600B total parameters). It performs strongly in software engineering and professional...

text+image+video->texttextimagevideo
Input$1.05per 1M
Output$2.84per 1M
Context1M
Max output64,000
Details
Upstage

Upstage: Solar Decide Flash

Solar Decide Flash is Upstage's low-latency structured decision model, a faster variant of [Solar Decide](/upstage/solar-decide) built on Solar Mini 4 and served through the System One (`/v1/systemone`) API. Instead of...

text->decisionstextdecisions
Input$0.0525per 1M
Output$0.00per 1M
Context524.3K
Max output471,859
Details
Anthropic

Anthropic: Claude Haiku 5.5

Claude Haiku 5.5 is Anthropic's small, fast model for high-volume, cost-sensitive work such as summarization, subagents, and browser use. It succeeds Claude Haiku 4.5 with stronger coding, computer use, and...

text+image+file->texttextimagefile
Input$0.105per 1M
Output$0.525per 1M
Context1M
Max output128,000
Details
Anthropic

Anthropic: Claude Haiku 5.5 (batch)

Claude Haiku 5.5 is Anthropic's small, fast model for high-volume, cost-sensitive work such as summarization, subagents, and browser use. It succeeds Claude Haiku 4.5 with stronger coding, computer use, and...

text+image+file->texttextimagefile
Input$0.0525per 1M
Output$0.2625per 1M
Context1M
Max output128,000
Details
Perplexity

Perplexity: Decider V1.1 27B

Decider V1.1 27B is a new checkpoint of Perplexity's decision model, succeeding [Decider V1 27B](/perplexity/pplx-decider-v1-27b) with the same API contract. Instead of generating text, it reads content passed as `state`...

text+image->decisionstextimagedecisions
Input$0.021per 1M
Output$0.00per 1M
Context262.1K
Max output235,929
Details
elevenlabs

ElevenLabs: Eleven v4

Eleven v4 is a text-to-speech model from ElevenLabs. It is ElevenLabs' most expressive model, with inline audio tags for emotional and delivery control, support for 90+ languages, and a 10,000-character...

text->speechtextspeech
Input$42.00per 1M
Output$0.00per 1M
Context0
Max output0
Details
elevenlabs

ElevenLabs: Eleven v4 Turbo

Eleven v4 Turbo is a low-latency text-to-speech model from ElevenLabs. It keeps Eleven v4's expressive delivery and audio tags while being tuned for faster generation, with support for 90+ languages...

text->speechtextspeech
Input$21.00per 1M
Output$0.00per 1M
Context0
Max output0
Details
elevenlabs

ElevenLabs: Eleven v3

Eleven v3 is a text-to-speech model from ElevenLabs. It produces emotionally rich, highly expressive speech with inline audio tags, supports 70+ languages, and has a 5,000-character request limit.

text->speechtextspeech
Input$42.00per 1M
Output$0.00per 1M
Context0
Max output0
Details
elevenlabs

ElevenLabs: Eleven v3 Conversational

Eleven v3 Conversational is a text-to-speech model from ElevenLabs, a variant of Eleven v3 optimized for natural dialogue in conversational agents. It supports 70+ languages and has a 5,000-character request...

text->speechtextspeech
Input$21.00per 1M
Output$0.00per 1M
Context0
Max output0
Details
Showing 12 of 653 models
Internal distribution

Use the marketplace with guide and comparison pages

These internal routes help buyers move from generic catalog browsing into comparison and buying-intent pages.

How TVP pricing works

Transparent pricing buyers can audit

Simple pricing rules, live source data, and exact decimals remain visible for review.

1

Provider pricing syncs live

TVP reads current provider catalog data and refreshes on demand.

2

TVP price is shown for planning

Buyers can compare options using one consistent TVP pricing view.

3

Prices are shown per 1M tokens

Tiny token decimals are normalized into buyer-readable rates.

4

Exact decimals stay available

Detailed pricing fields remain visible for audit and billing checks.

Last synced: 10/9/2026, 11:51:06 PM