Marketplace

AI Model Marketplace

Search, filter, compare, and audit live model pricing with TVP buyer pricing.

474Models72ProvidersLivePricing
Showing 12 of 474 models
Select up to 3 live models to compare pricing, context, and endpoint coverage.
Qwen

Qwen: Qwen3.7 Flash

Qwen3.7 Flash is a vision-language reasoning model from Alibaba. It is suited for multimodal agents, visual coding, search, and computer interaction, with strengths in object recognition, spatial understanding, and real-world...

text+image+video->texttextimagevideo
Input$0.0315per 1M
Output$0.1365per 1M
Context1M
Max output65,536
Details
voyageai

VoyageAI by MongoDB: rerank-2.5-lite

Rerank 2.5 Lite is an instruction-following text reranker from Voyage AI by MongoDB, balancing ranking quality with lower latency. It is suited for high-throughput retrieval pipelines and accepts natural-language guidance...

text->reranktextrerank
Input$0.00per 1M
Output$0.00per 1M
Context32K
Max outputn/a
Details
voyageai

VoyageAI by MongoDB: rerank-2.5

Rerank 2.5 is an instruction-following text reranker from Voyage AI by MongoDB, optimized for retrieval quality. It is suited for rescoring candidates in general-purpose and domain-specific search, and accepts natural-language...

text->reranktextrerank
Input$0.00per 1M
Output$0.00per 1M
Context32K
Max outputn/a
Details
voyageai

VoyageAI by MongoDB: voyage-multimodal-3.5

Voyage Multimodal 3.5 is a multimodal embedding model from Voyage AI by MongoDB. It maps interleaved text and visual content into a shared embedding space for cross-modal retrieval over documents,...

text+image->embeddingsembeddingstextimage
Input$0.126per 1M
Output$0.00per 1M
Context32K
Max outputn/a
Details
voyageai

VoyageAI by MongoDB: voyage-4-lite

Voyage 4 Lite is an efficiency-focused general-purpose embedding model from Voyage AI by MongoDB, optimized for low-latency and cost-sensitive retrieval. It supports Matryoshka embeddings at 2048, 1024, 512, and 256...

text->embeddingsembeddingstext
Input$0.021per 1M
Output$0.00per 1M
Context32K
Max outputn/a
Details
voyageai

VoyageAI by MongoDB: voyage-4

Voyage 4 is a general-purpose multilingual embedding model from Voyage AI by MongoDB. It is suited for retrieval, semantic search, and RAG applications, with Matryoshka embeddings at 2048, 1024, 512,...

text->embeddingsembeddingstext
Input$0.063per 1M
Output$0.00per 1M
Context32K
Max outputn/a
Details
voyageai

VoyageAI by MongoDB: voyage-4-large

Voyage 4 Large is a general-purpose multilingual embedding model from Voyage AI by MongoDB, optimized for retrieval quality. It supports Matryoshka embeddings at 2048, 1024, 512, and 256 dimensions, with...

text->embeddingsembeddingstext
Input$0.126per 1M
Output$0.00per 1M
Context32K
Max outputn/a
Details
Anthropic

Claude Opus 5 (Fast)

Fast-mode variant of [Opus 5](/anthropic/claude-opus-5) - identical capabilities with higher output speed at 2x pricing relative to regular Opus 5. Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-mode

text+image+file->texttextimagefile
Input$10.50per 1M
Output$52.50per 1M
Context1M
Max output128,000
Details
Anthropic

Claude Opus 5

Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...

text+image+file->texttextimagefile
Input$5.25per 1M
Output$26.25per 1M
Context1M
Max output128,000
Details
Microsoft

Microsoft: MAI-Image-2.5 Pro

Microsoft's MAI-Image-2.5 is a high-quality image generation model available via Azure AI Foundry. It produces photorealistic and artistic images from text prompts with support for various aspect ratios.

text+image->imagetextimage
Input$5.25per 1M
Output$0.00per 1M
Context4.1K
Max output1,024
Details
Microsoft

Microsoft: MAI-Voice-2-Flash

MAI-Voice-2-Flash is a low-latency text-to-speech model from Microsoft for voice agents, assistants, call centers, accessibility, narration, and other interactive applications. It generates expressive 24 kHz mono speech across 15 languages...

text->speechtextspeech
Input$15.75per 1M
Output$0.00per 1M
Context0
Max outputn/a
Details
inclusionAI

Ling-3.0-flash (free)

*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...

text->texttexttools
Input$0.00per 1M
Output$0.00per 1M
Context262.1K
Max output32,768
Details
Showing 12 of 474 models
Internal distribution

Use the marketplace with guide and comparison pages

These internal routes help buyers move from generic catalog browsing into comparison and buying-intent pages.

How TVP pricing works

Transparent pricing buyers can audit

Simple pricing rules, live source data, and exact decimals remain visible for review.

1

Provider pricing syncs live

TVP reads current provider catalog data and refreshes on demand.

2

TVP price is shown for planning

Buyers can compare options using one consistent TVP pricing view.

3

Prices are shown per 1M tokens

Tiny token decimals are normalized into buyer-readable rates.

4

Exact decimals stay available

Detailed pricing fields remain visible for audit and billing checks.

Last synced: 7/29/2026, 1:17:09 AM