DeepSeek: DeepSeek Pro Latest
This model always redirects to the latest model in the DeepSeek Pro family.
This model always redirects to the latest model in the DeepSeek Pro family.
This model always redirects to the latest model in the DeepSeek Flash family.
Schematron V2 Turbo is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes throughput for high-volume extraction workloads. Extraction instructions must be supplied through a JSON schema in response_format rather...
Schematron V2 Small is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes extraction quality for complex schemas and long pages. Extraction instructions must be supplied through a JSON schema...
Muse Voice Transcribe 1.0 is a synchronous speech-to-text model from Meta. It is suited for push-to-talk, endpointing, and speaker-aware transcription, with keyword biasing for domain terms and language biasing through...
This model always redirects to the latest model in the GPT Astra family.
This model always redirects to the latest model in the GPT Sol family.
This model always redirects to the latest model in the GPT Terra family.
This model always redirects to the latest model in the GPT Luna family.
Fugu Ultra v2 is the higher-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to...
Fugu Max is the cost-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to route...
FLUX Video Edit [fast] takes a source video and an edit prompt and returns a precisely edited video. Add, remove, or replace objects and characters, rebuild the setting, edit on-screen...
These internal routes help buyers move from generic catalog browsing into comparison and buying-intent pages.
Shortlist useful models before drilling into the full marketplace.
Strong fit for complex work with reasoning support and large context.
Lowest combined TVP input and output price among paid text models.
Image-capable model selected from live pricing and capability data.
Largest context window with numeric input and output pricing.
Simple pricing rules, live source data, and exact decimals remain visible for review.
TVP reads current provider catalog data and refreshes on demand.
Buyers can compare options using one consistent TVP pricing view.
Tiny token decimals are normalized into buyer-readable rates.
Detailed pricing fields remain visible for audit and billing checks.
Last synced: 9/14/2026, 11:10:18 PM