Marketplace

AI Model Marketplace

Search, filter, compare, and audit live model pricing with TVP buyer pricing.

450Models75ProvidersLivePricing
Showing 12 of 450 models
Select up to 3 live models to compare pricing, context, and endpoint coverage.
Qwen

Qwen: Qwen3.8 Max

Qwen3.8 Max is the flagship model in Alibaba's Qwen3.8 series, the general-availability successor to the Qwen3.8 Max Preview. It is a multimodal reasoning model intended for complex reasoning, visual understanding,...

text+image+video->texttextimagevideo
Input$2.10per 1M
Output$6.30per 1M
Context1M
Max output131,072
Details
DeepSeek

DeepSeek V4 Flash Latest

This model always redirects to the latest model in the DeepSeek V4 Flash family.

text->texttexttoolsstructured
Input$0.0945per 1M
Output$0.189per 1M
Context1M
Max output65,536
Details
DeepSeek

DeepSeek: DeepSeek V4 Flash 0731

DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows.

text->texttexttoolsstructured
Input$0.0945per 1M
Output$0.189per 1M
Context1M
Max output65,536
Details
thinkingmachines

Thinking Machines: Inkling Small

Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...

text+image+audio->texttextimageaudio
Input$0.525per 1M
Output$1.26per 1M
Context524.3K
Max outputn/a
Details
MiniMax

MiniMax: H3

MiniMax H3 is a lightweight, open-weights video generation model from MiniMax. It is designed for precise multimodal editing and controlled content generation, including instruction-guided edits, text and brand rendering, and...

text+image+audio+video->videotextimagevideo
Input$0.00per 1M
Output$0.00per 1M
Context0
Max outputn/a
Details
fish-audio

Fish Audio: Transcribe 1

Transcribe 1 is a speech-to-text model from Fish Audio. It is suited for audio transcription with automatic language detection and can return timestamped word-level segments when alignment details are requested.

audio->transcriptionaudiotranscription
Input$105.00per 1M
Output$0.00per 1M
Context0
Max outputn/a
Details
fish-audio

Fish Audio: S1

S1 is a multilingual text-to-speech model from Fish Audio. It is suited for voice applications that need broad emotional expression, using parenthetical controls to guide speaking style across its supported...

text->speechtextspeech
Input$15.75per 1M
Output$0.00per 1M
Context0
Max outputn/a
Details
fish-audio

Fish Audio: S2 Pro

S2 Pro is a multilingual text-to-speech model from Fish Audio. It is suited for expressive narration and multi-speaker dialogue, with natural-language controls for speaking style and emotion.

text->speechtextspeech
Input$15.75per 1M
Output$0.00per 1M
Context0
Max outputn/a
Details
fish-audio

Fish Audio: S2.1 Pro Free (free)

S2.1 Pro Free is the no-cost variant of Fish Audio S2.1 Pro, intended for testing, prototyping, and low-volume applications. It provides the same synthesis capabilities without production latency or availability...

text->speechtextspeech
Input$0.00per 1M
Output$0.00per 1M
Context0
Max outputn/a
Details
fish-audio

Fish Audio: S2.1 Pro

S2.1 Pro is a production-oriented text-to-speech model from Fish Audio. It is suited for multilingual voice applications, expressive narration, and dialogue synthesis, with open-ended natural-language controls for speaking style and...

text->speechtextspeech
Input$15.75per 1M
Output$0.00per 1M
Context0
Max outputn/a
Details
runway

Runway: Aleph 2.0

Runway Aleph 2.0 is an in-context video editing model from Runway. It applies text instructions and keyframe-guided edits across existing footage while preserving details that are not meant to change....

text+image+video->videotextimagevideo
Input$0.00per 1M
Output$0.00per 1M
Context0
Max outputn/a
Details
runway

Runway: Gen-4.5

Runway Gen-4.5 is a video generation model from Runway for text-to-video and image-to-video workflows. It is designed for cinematic scene creation with strong motion quality, visual fidelity, and prompt adherence....

text+image->videotextimagevideo
Input$0.00per 1M
Output$0.00per 1M
Context0
Max outputn/a
Details
Showing 12 of 450 models
Internal distribution

Use the marketplace with guide and comparison pages

These internal routes help buyers move from generic catalog browsing into comparison and buying-intent pages.

How TVP pricing works

Transparent pricing buyers can audit

Simple pricing rules, live source data, and exact decimals remain visible for review.

1

Provider pricing syncs live

TVP reads current provider catalog data and refreshes on demand.

2

TVP price is shown for planning

Buyers can compare options using one consistent TVP pricing view.

3

Prices are shown per 1M tokens

Tiny token decimals are normalized into buyer-readable rates.

4

Exact decimals stay available

Detailed pricing fields remain visible for audit and billing checks.

Last synced: 8/3/2026, 7:23:32 PM