Skip to content

Embed V1 4B

Perplexity · 32K context · released Mar 16, 2026 · checked 2h ago

Ways to buy 2

Way to buyIn / out per 1M
Perplexity API
The lab's own API
$0.03 / FreeList price
Perplexity
int8
$0.03 / FreeSame as list

USD per 1M tokens. Regions are separate routes.

Through a gateway

Your own numbers

100M input and 20M output tokens a month: $3.00 direct.

A dash: not on that gateway's own model list.

Related models

In / out per 1M
Embed V1 0.6B
Perplexity · 32K
$0.004 / Free
Sonar Pro Search
Perplexity · 200K
$3 / $15
Sonar Deep Research
Perplexity · 128K
$2 / $8
Sonar Pro
Perplexity · 200K
$3 / $15
Text Embedding 3 Small
OpenAI · 8K
$0.02 / Free
Qwen3 Embedding 4B
Qwen (Alibaba) · 33K
$0.02 / Free
gpt-oss-20b
OpenAI · 131K
$0.018 / $0.09
Llama 3.1 8B Instruct
Meta · 131K
$0.02 / $0.04

About Embed V1 4B

pplx-embed-v1 -4B is one of Perplexity's state-of-the-art text embedding models built for real-world, web-scale retrieval.

Takes text in and returns embeddings. First listed Sep 26, 2026. Blended price $0.022 per 1M tokens at 3 input to 1 output. Model id pplx-embed-v1-4b.

Price history

List price since Sep 26, 2026, from daily reads of the OpenRouter catalog.

Questions

How much does Embed V1 4B cost?

$0.03 per million input tokens and Free per million output tokens.

What is the context window of Embed V1 4B?

32K tokens.

Where is Embed V1 4B cheapest?

Of 2 routes tracked, Perplexity is cheapest at $0.03 input and Free output per million tokens.

Price changes by email

Thursdays, only in weeks when an LLM API list price changed.

Double opt-in. Unsubscribe any time.