Marketplace

AI Model Marketplace

Search, filter, compare, and audit live model pricing with TVP buyer pricing.

662Models86ProvidersLivePricing
Showing 12 of 662 models
Select up to 3 live models to compare pricing, context, and endpoint coverage.
StepFun

StepFun: Step 5 Preview

Step 5 Preview is StepFun's flagship model for agentic work, built on a sparse Mixture-of-Experts architecture (27B active / 600B total parameters). It performs strongly in software engineering and professional...

text+image+video->texttextimagevideo
Input$1.05per 1M
Output$2.84per 1M
Context1M
Max output64,000
Details
Upstage

Upstage: Solar Decide Flash

Solar Decide Flash is Upstage's low-latency structured decision model, a faster variant of [Solar Decide](/upstage/solar-decide) built on Solar Mini 4 and served through the System One (`/v1/systemone`) API. Instead of...

text->decisionstextdecisions
Input$0.0525per 1M
Output$0.00per 1M
Context524.3K
Max output471,859
Details
Anthropic

Anthropic: Claude Haiku 5.5

Claude Haiku 5.5 is Anthropic's small, fast model for high-volume, cost-sensitive work such as summarization, subagents, and browser use. It succeeds Claude Haiku 4.5 with stronger coding, computer use, and...

text+image+file->texttextimagefile
Input$0.105per 1M
Output$0.525per 1M
Context1M
Max output128,000
Details
Anthropic

Anthropic: Claude Haiku 5.5 (batch)

Claude Haiku 5.5 is Anthropic's small, fast model for high-volume, cost-sensitive work such as summarization, subagents, and browser use. It succeeds Claude Haiku 4.5 with stronger coding, computer use, and...

text+image+file->texttextimagefile
Input$0.0525per 1M
Output$0.2625per 1M
Context1M
Max output128,000
Details
Perplexity

Perplexity: Decider V1.1 27B

Decider V1.1 27B is a new checkpoint of Perplexity's decision model, succeeding [Decider V1 27B](/perplexity/pplx-decider-v1-27b) with the same API contract. Instead of generating text, it reads content passed as `state`...

text+image->decisionstextimagedecisions
Input$0.021per 1M
Output$0.00per 1M
Context262.1K
Max output235,929
Details
elevenlabs

ElevenLabs: Eleven v4

Eleven v4 is a text-to-speech model from ElevenLabs. It is ElevenLabs' most expressive model, with inline audio tags for emotional and delivery control, support for 90+ languages, and a 10,000-character...

text->speechtextspeech
Input$42.00per 1M
Output$0.00per 1M
Context0
Max output0
Details
elevenlabs

ElevenLabs: Eleven v4 Turbo

Eleven v4 Turbo is a low-latency text-to-speech model from ElevenLabs. It keeps Eleven v4's expressive delivery and audio tags while being tuned for faster generation, with support for 90+ languages...

text->speechtextspeech
Input$21.00per 1M
Output$0.00per 1M
Context0
Max output0
Details
elevenlabs

ElevenLabs: Eleven v3

Eleven v3 is a text-to-speech model from ElevenLabs. It produces emotionally rich, highly expressive speech with inline audio tags, supports 70+ languages, and has a 5,000-character request limit.

text->speechtextspeech
Input$42.00per 1M
Output$0.00per 1M
Context0
Max output0
Details
elevenlabs

ElevenLabs: Eleven v3 Conversational

Eleven v3 Conversational is a text-to-speech model from ElevenLabs, a variant of Eleven v3 optimized for natural dialogue in conversational agents. It supports 70+ languages and has a 5,000-character request...

text->speechtextspeech
Input$21.00per 1M
Output$0.00per 1M
Context0
Max output0
Details
elevenlabs

ElevenLabs: Eleven Flash v2.5

Eleven Flash v2.5 is an ultra-low-latency text-to-speech model from ElevenLabs. It is suited for conversational and real-time use cases, supports 32 languages, and has a 40,000-character request limit.

text->speechtextspeech
Input$21.00per 1M
Output$0.00per 1M
Context0
Max output0
Details
elevenlabs

ElevenLabs: Eleven Turbo v2

Eleven Turbo v2 is an English-only, low-latency text-to-speech model from ElevenLabs. It is suited for developer use cases where speed matters and only English is needed, and has a 30,000-character...

text->speechtextspeech
Input$21.00per 1M
Output$0.00per 1M
Context0
Max output0
Details
elevenlabs

ElevenLabs: Eleven Multilingual v2

Eleven Multilingual v2 is a text-to-speech model from ElevenLabs. It is suited for lifelike, consistent long-form narration such as voice-overs and audiobooks, supports 29 languages, and has a 10,000-character request...

text->speechtextspeech
Input$42.00per 1M
Output$0.00per 1M
Context0
Max output0
Details
Showing 12 of 662 models
Internal distribution

Use the marketplace with guide and comparison pages

These internal routes help buyers move from generic catalog browsing into comparison and buying-intent pages.

How TVP pricing works

Transparent pricing buyers can audit

Simple pricing rules, live source data, and exact decimals remain visible for review.

1

Provider pricing syncs live

TVP reads current provider catalog data and refreshes on demand.

2

TVP price is shown for planning

Buyers can compare options using one consistent TVP pricing view.

3

Prices are shown per 1M tokens

Tiny token decimals are normalized into buyer-readable rates.

4

Exact decimals stay available

Detailed pricing fields remain visible for audit and billing checks.

Last synced: 10/8/2026, 11:05:39 PM