Qwen: Qwen3.7 Flash
Qwen3.7 Flash is a vision-language reasoning model from Alibaba. It is suited for multimodal agents, visual coding, search, and computer interaction, with strengths in object recognition, spatial understanding, and real-world...
Qwen3.7 Flash is a vision-language reasoning model from Alibaba. It is suited for multimodal agents, visual coding, search, and computer interaction, with strengths in object recognition, spatial understanding, and real-world...
Rerank 2.5 Lite is an instruction-following text reranker from Voyage AI by MongoDB, balancing ranking quality with lower latency. It is suited for high-throughput retrieval pipelines and accepts natural-language guidance...
Rerank 2.5 is an instruction-following text reranker from Voyage AI by MongoDB, optimized for retrieval quality. It is suited for rescoring candidates in general-purpose and domain-specific search, and accepts natural-language...
Voyage Multimodal 3.5 is a multimodal embedding model from Voyage AI by MongoDB. It maps interleaved text and visual content into a shared embedding space for cross-modal retrieval over documents,...
Voyage 4 Lite is an efficiency-focused general-purpose embedding model from Voyage AI by MongoDB, optimized for low-latency and cost-sensitive retrieval. It supports Matryoshka embeddings at 2048, 1024, 512, and 256...
Voyage 4 is a general-purpose multilingual embedding model from Voyage AI by MongoDB. It is suited for retrieval, semantic search, and RAG applications, with Matryoshka embeddings at 2048, 1024, 512,...
Voyage 4 Large is a general-purpose multilingual embedding model from Voyage AI by MongoDB, optimized for retrieval quality. It supports Matryoshka embeddings at 2048, 1024, 512, and 256 dimensions, with...
Fast-mode variant of [Opus 5](/anthropic/claude-opus-5) - identical capabilities with higher output speed at 2x pricing relative to regular Opus 5. Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-mode
Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...
Microsoft's MAI-Image-2.5 is a high-quality image generation model available via Azure AI Foundry. It produces photorealistic and artistic images from text prompts with support for various aspect ratios.
MAI-Voice-2-Flash is a low-latency text-to-speech model from Microsoft for voice agents, assistants, call centers, accessibility, narration, and other interactive applications. It generates expressive 24 kHz mono speech across 15 languages...
*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...
These internal routes help buyers move from generic catalog browsing into comparison and buying-intent pages.
Shortlist useful models before drilling into the full marketplace.
Strong fit for complex work with reasoning support and large context.
Lowest combined TVP input and output price among paid text models.
Image-capable model selected from live pricing and capability data.
Largest context window with numeric input and output pricing.
Simple pricing rules, live source data, and exact decimals remain visible for review.
TVP reads current provider catalog data and refreshes on demand.
Buyers can compare options using one consistent TVP pricing view.
Tiny token decimals are normalized into buyer-readable rates.
Detailed pricing fields remain visible for audit and billing checks.
Last synced: 7/29/2026, 1:11:24 AM