Marketplace

AI Model Marketplace

Search, filter, compare, and audit live model pricing with TVP buyer pricing.

449Models75ProvidersLivePricing
Showing 12 of 449 models
Select up to 3 live models to compare pricing, context, and endpoint coverage.
DeepSeek

DeepSeek V4 Flash Latest

This model always redirects to the latest model in the DeepSeek V4 Flash family.

text->texttexttoolsstructured
Input$0.0945per 1M
Output$0.189per 1M
Context1M
Max output65,536
Details
DeepSeek

DeepSeek: DeepSeek V4 Flash 0731

DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows.

text->texttexttoolsstructured
Input$0.0945per 1M
Output$0.189per 1M
Context1M
Max output65,536
Details
thinkingmachines

Thinking Machines: Inkling Small

Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...

text+image+audio->texttextimageaudio
Input$0.525per 1M
Output$1.26per 1M
Context524.3K
Max outputn/a
Details
MiniMax

MiniMax: H3

MiniMax H3 is a lightweight, open-weights video generation model from MiniMax. It is designed for precise multimodal editing and controlled content generation, including instruction-guided edits, text and brand rendering, and...

text+image+audio+video->videotextimagevideo
Input$0.00per 1M
Output$0.00per 1M
Context0
Max outputn/a
Details
fish-audio

Fish Audio: Transcribe 1

Transcribe 1 is a speech-to-text model from Fish Audio. It is suited for audio transcription with automatic language detection and can return timestamped word-level segments when alignment details are requested.

audio->transcriptionaudiotranscription
Input$105.00per 1M
Output$0.00per 1M
Context0
Max outputn/a
Details
fish-audio

Fish Audio: S1

S1 is a multilingual text-to-speech model from Fish Audio. It is suited for voice applications that need broad emotional expression, using parenthetical controls to guide speaking style across its supported...

text->speechtextspeech
Input$15.75per 1M
Output$0.00per 1M
Context0
Max outputn/a
Details
fish-audio

Fish Audio: S2 Pro

S2 Pro is a multilingual text-to-speech model from Fish Audio. It is suited for expressive narration and multi-speaker dialogue, with natural-language controls for speaking style and emotion.

text->speechtextspeech
Input$15.75per 1M
Output$0.00per 1M
Context0
Max outputn/a
Details
fish-audio

Fish Audio: S2.1 Pro Free (free)

S2.1 Pro Free is the no-cost variant of Fish Audio S2.1 Pro, intended for testing, prototyping, and low-volume applications. It provides the same synthesis capabilities without production latency or availability...

text->speechtextspeech
Input$0.00per 1M
Output$0.00per 1M
Context0
Max outputn/a
Details
fish-audio

Fish Audio: S2.1 Pro

S2.1 Pro is a production-oriented text-to-speech model from Fish Audio. It is suited for multilingual voice applications, expressive narration, and dialogue synthesis, with open-ended natural-language controls for speaking style and...

text->speechtextspeech
Input$15.75per 1M
Output$0.00per 1M
Context0
Max outputn/a
Details
runway

Runway: Aleph 2.0

Runway Aleph 2.0 is an in-context video editing model from Runway. It applies text instructions and keyframe-guided edits across existing footage while preserving details that are not meant to change....

text+image+video->videotextimagevideo
Input$0.00per 1M
Output$0.00per 1M
Context0
Max outputn/a
Details
runway

Runway: Gen-4.5

Runway Gen-4.5 is a video generation model from Runway for text-to-video and image-to-video workflows. It is designed for cinematic scene creation with strong motion quality, visual fidelity, and prompt adherence....

text+image->videotextimagevideo
Input$0.00per 1M
Output$0.00per 1M
Context0
Max outputn/a
Details
Qwen

Qwen: Qwen3.7 Flash

Qwen3.7 Flash is a vision-language reasoning model from Alibaba. It is suited for multimodal agents, visual coding, search, and computer interaction, with strengths in object recognition, spatial understanding, and real-world...

text+image+video->texttextimagevideo
Input$0.0315per 1M
Output$0.1365per 1M
Context1M
Max output65,536
Details
Showing 12 of 449 models
Internal distribution

Use the marketplace with guide and comparison pages

These internal routes help buyers move from generic catalog browsing into comparison and buying-intent pages.

How TVP pricing works

Transparent pricing buyers can audit

Simple pricing rules, live source data, and exact decimals remain visible for review.

1

Provider pricing syncs live

TVP reads current provider catalog data and refreshes on demand.

2

TVP price is shown for planning

Buyers can compare options using one consistent TVP pricing view.

3

Prices are shown per 1M tokens

Tiny token decimals are normalized into buyer-readable rates.

4

Exact decimals stay available

Detailed pricing fields remain visible for audit and billing checks.

Last synced: 8/1/2026, 10:32:49 PM