Skip to content

Compare model prices

Input, output, cached and batch prices, every route and a month of tokens, side by side.

Embed V1 0.6BEmbed V1 4B

A month of tokens

  • Embed V1 0.6B$0.40
  • Embed V1 4B$3.00

100M input and 20M output tokens, list price. Use your own numbers.

Side by side

Embed V1 0.6BEmbed V1 4B
Input per 1M$0.004$0.03
Output per 1MFreeFree
Cached inputn/an/a
Batch in / outn/an/a
Blended (3:1)$0.003$0.022
Context32K32K
Routes22
ReleasedMar 16, 2026Mar 16, 2026
Leadersboard scoren/an/a

On the price map

Size: context window, 8K to 2MEqual blended cost (3:1) 254 models, lower left is cheaper

Every route

Embed V1 0.6B

Way to buyIn / out per 1M
Perplexity API
The lab's own API
$0.004 / FreeList price
Perplexity
int8
$0.004 / FreeSame as list

Embed V1 4B

Way to buyIn / out per 1M
Perplexity API
The lab's own API
$0.03 / FreeList price
Perplexity
int8
$0.03 / FreeSame as list

Price changes by email

Thursdays, only in weeks when an LLM API list price changed.

Double opt-in. Unsubscribe any time.