Jared Palmer: Kev 4B
Kev 4B is a small open-weight decision model from Jared Palmer, built as a LoRA adapter and pointer head on Qwen3.5-4B-Base and served over the same /v1/systemone contract as TypeSafe's...
Kev 4B is a small open-weight decision model from Jared Palmer, built as a LoRA adapter and pointer head on Qwen3.5-4B-Base and served over the same /v1/systemone contract as TypeSafe's...
Perceptron Mk1.5 is Perceptron's embodied reasoning model for physical agents. It accepts text, image, video, and audio input, and answers with text plus optional structured annotations: points, boxes, polygons, tracks,...
Gemini 3.5 Transcribe is a speech-to-text model from Google. It is suited for synchronous transcription that needs word-level timestamps or speaker diarization, with support for up to eight speakers. Audio...
Transcribe 1 Pro is a speech-to-text model from Fish Audio tuned for interviews, meetings, and podcasts. It labels speakers with inline `speaker` markers, preserves emotion and vocal-event cues such as...
Ember-1 is a specialized reasoning model from Fireworks Research, built on [Kimi K3](TVP catalog). It is designed to make every token go further: it produces shorter reasoning traces, using roughly 40%...
Ming Image 0.1 Design Layer is an image-to-image model from inclusionAI that decomposes a flattened design image into separate RGBA layers, such as a background layer and foreground elements, and...
Gemini 3.8 Flash Lite TTS is a text-to-speech model from Google and the fast, high-throughput member of the 3.8 TTS family alongside [Gemini 3.8 Flash TTS](TVP catalog). It is suited for...
Gemini 3.8 Flash TTS is a text-to-speech model from Google and the successor to [Gemini 3.1 Flash TTS Preview](TVP catalog). It is the creative tier of the 3.8 TTS family, suited...
GLM-5.3-Prime is the high-speed variant of Z.ai's GLM-5.3, inheriting its full capabilities while delivering 1.5–2× the output throughput through inference acceleration. It supports text input and output with a 1M-token...
Qwen3.8 Max Prime is a higher-throughput variant of Qwen3.8 Max from Alibaba's Qwen team, served as a separate SKU at a higher price point. It accepts text, image, and video...
Recraft V4.1 Flash is a text-to-image model from Recraft, the speed and cost tier of the V4.1 family. It generates ~1K raster images in about 1.5 seconds end to end,...
Space Bunny Alpha is an anonymous large model with blazing-fast inference, strong coding capabilities and native multimodal input support. It delivers adjustable reasoning effort, and a 1M-token context window. Space...
These internal routes help buyers move from generic catalog browsing into comparison and buying-intent pages.
Shortlist useful models before drilling into the full marketplace.
Strong fit for complex work with reasoning support and large context.
Lowest combined TVP input and output price among paid text models.
Image-capable model selected from live pricing and capability data.
Largest context window with numeric input and output pricing.
Simple pricing rules, live source data, and exact decimals remain visible for review.
TVP reads current provider catalog data and refreshes on demand.
Buyers can compare options using one consistent TVP pricing view.
Tiny token decimals are normalized into buyer-readable rates.
Detailed pricing fields remain visible for audit and billing checks.
Last synced: 9/25/2026, 8:56:29 PM