Marketplace

AI Model Marketplace

Search, filter, compare, and audit live model pricing with TVP buyer pricing.

620Models83ProvidersLivePricing
Showing 12 of 620 models
Select up to 3 live models to compare pricing, context, and endpoint coverage.
ByteDance Seed

ByteDance Seed: Seed Audio 1.0

Seed Audio 1.0 is ByteDance Seed's non-streaming audio generation model. It produces speech and other audio from a natural-language text prompt that can describe the desired voice, tone, and sound...

text->speechtextspeech
Input$0.00per 1M
Output$2625.00per 1M
Context0
Max output0
Details
typesafe

TypeSafe: Jev Router

Jev Router picks the best model and reasoning effort for each request, balancing quality, speed, and cost. It runs on [Jev](TVP catalog), TypeSafe's first System One model, and adapts as your...

text+image+file+audio+video->textaudiofileimage
Inputn/aper 1M
Outputn/aper 1M
Context1M
Max outputn/a
Details
jaredpalmer

Jared Palmer: Kev 4B

Kev 4B is a small open-weight decision model from Jared Palmer, built as a LoRA adapter and pointer head on Qwen3.5-4B-Base and served over the same /v1/systemone contract as TypeSafe's...

text->decisionstextdecisions
Input$0.0441per 1M
Output$0.00per 1M
Context8.2K
Max output7,372
Details
Perceptron

Perceptron: Perceptron Mk1.5

Perceptron Mk1.5 is Perceptron's embodied reasoning model for physical agents. It accepts text, image, video, and audio input, and answers with text plus optional structured annotations: points, boxes, polygons, tracks,...

text+image+audio+video->texttextimagevideo
Input$0.1575per 1M
Output$1.58per 1M
Context36.9K
Max output8,192
Details
Google

Google: Gemini 3.5 Transcribe

Gemini 3.5 Transcribe is a speech-to-text model from Google. It is suited for synchronous transcription that needs word-level timestamps or speaker diarization, with support for up to eight speakers. Audio...

audio->transcriptionaudiotranscription
Input$2.10per 1M
Output$12.60per 1M
Context98.3K
Max output32,768
Details
fish-audio

Fish Audio: Transcribe 1 Pro

Transcribe 1 Pro is a speech-to-text model from Fish Audio tuned for interviews, meetings, and podcasts. It labels speakers with inline `speaker` markers, preserves emotion and vocal-event cues such as...

audio->transcriptionaudiotranscription
Input$105.00per 1M
Output$0.00per 1M
Context0
Max output0
Details
fireworks

Fireworks: Ember-1

Ember-1 is a specialized reasoning model from Fireworks Research, built on [Kimi K3](TVP catalog). It is designed to make every token go further: it produces shorter reasoning traces, using roughly 40%...

text+image->texttextimagetools
Input$3.15per 1M
Output$15.75per 1M
Context1M
Max output943,718
Details
inclusionAI

inclusionAI: Ming Image 0.1 Design Layer

Ming Image 0.1 Design Layer is an image-to-image model from inclusionAI that decomposes a flattened design image into separate RGBA layers, such as a background layer and foreground elements, and...

text+image->imagetextimage
Input$0.00per 1M
Output$0.00per 1M
Context0
Max output0
Details
Google

Google: Gemini 3.8 Flash Lite TTS

Gemini 3.8 Flash Lite TTS is a text-to-speech model from Google and the fast, high-throughput member of the 3.8 TTS family alongside [Gemini 3.8 Flash TTS](TVP catalog). It is suited for...

text->speechtextspeech
Input$0.525per 1M
Output$6.30per 1M
Context8.2K
Max output7,372
Details
Google

Google: Gemini 3.8 Flash TTS

Gemini 3.8 Flash TTS is a text-to-speech model from Google and the successor to [Gemini 3.1 Flash TTS Preview](TVP catalog). It is the creative tier of the 3.8 TTS family, suited...

text->speechtextspeech
Input$0.525per 1M
Output$9.45per 1M
Context8.2K
Max output7,372
Details
Z.ai

Z.ai: GLM 5.3 Prime

GLM-5.3-Prime is the high-speed variant of Z.ai's GLM-5.3, inheriting its full capabilities while delivering 1.5–2× the output throughput through inference acceleration. It supports text input and output with a 1M-token...

text->texttexttoolsstructured
Input$2.94per 1M
Output$9.24per 1M
Context1M
Max output131,072
Details
Qwen

Qwen: Qwen3.8 Max Prime

Qwen3.8 Max Prime is a higher-throughput variant of Qwen3.8 Max from Alibaba's Qwen team, served as a separate SKU at a higher price point. It accepts text, image, and video...

text+image+video->texttextimagevideo
Input$4.20per 1M
Output$12.60per 1M
Context1M
Max output131,072
Details
Showing 12 of 620 models
Internal distribution

Use the marketplace with guide and comparison pages

These internal routes help buyers move from generic catalog browsing into comparison and buying-intent pages.

How TVP pricing works

Transparent pricing buyers can audit

Simple pricing rules, live source data, and exact decimals remain visible for review.

1

Provider pricing syncs live

TVP reads current provider catalog data and refreshes on demand.

2

TVP price is shown for planning

Buyers can compare options using one consistent TVP pricing view.

3

Prices are shown per 1M tokens

Tiny token decimals are normalized into buyer-readable rates.

4

Exact decimals stay available

Detailed pricing fields remain visible for audit and billing checks.

Last synced: 9/25/2026, 10:53:15 PM