Get an email when this model's blended price drops. No account needed.
One email per drop. Unsubscribe anytime.
Daily blended price ($/1M) — recorded each day, builds into a trend over time.
Typical 3:1 output-to-input mix, per 1M tokens
Source: litellm
General-purpose text generation, chat, summarization and content workloads where broad capability and low cost matter most.
Embed v4.0 creates unified text and image embeddings for enterprise retrieval and semantic search.
embed-v4.0 is a Text model from Cohere (US). HotON.ai tracks it at $0.12 per 1M input tokens and $0.00 per 1M output tokens, with a 128K-token context window. Its composite efficiency score is 89/100 at an estimated $0.000 per successful task.
embed-v4.0 is tracked at $0.12 per 1M input tokens and $0.00 per 1M output tokens. A typical 3:1 output-to-input workload blends to roughly $0.03 per 1M tokens. Figures are illustrative demo data.
General-purpose text generation, chat, summarization and content workloads where broad capability and low cost matter most.
embed-v4.0 supports up to a 128K-token context window — large enough for long documents and extended conversations in a single request.
Within the HotON.ai tracked set, embed-v4.0 is cheaper than 77% of models on input price and ranks #181 of 535 by overall efficiency.
Yes — morph-rerank-v3 is a lower-cost option at $0.00 per 1M output tokens, while still covering similar Text use cases. Compare them side by side on HotON.ai.
Ready to paste into articles, papers or AI prompts — prices and date refresh with the live data.
HotON.ai — embed-v4.0 (Cohere): $0.12/1M input, $0.00/1M output. https://hoton.ai/en/models/cohere-embed-v4-0Pricing is real (via the TestKey catalog, updated daily). Quality (Arena Elo) is real where the model is ranked on LMArena. Efficiency is a modeled composite of real price and context.