Get an email when this model's blended price drops. No account needed.
One email per drop. Unsubscribe anytime.
Daily blended price ($/1M) — recorded each day, builds into a trend over time.
Typical 3:1 output-to-input mix, per 1M tokens
Price as of 2026-05-11 · Source: alibaba_reference_catalog
General-purpose text generation, chat, summarization and content workloads where broad capability and low cost matter most.
Tongyi DeepResearch is an agentic large language model developed by Tongyi Lab, with 30 billion total parameters activating only 3 billion per token. It's optimized for long-horizon, deep information-seeking tasks...
tongyi-deepresearch-30b-a3b is a Text model from Alibaba Group (CN). HotON.ai tracks it at $0.09 per 1M input tokens and $0.45 per 1M output tokens, with a 131K-token context window. Its composite efficiency score is 89/100 at an estimated $0.000 per successful task.
tongyi-deepresearch-30b-a3b is tracked at $0.09 per 1M input tokens and $0.45 per 1M output tokens. A typical 3:1 output-to-input workload blends to roughly $0.36 per 1M tokens. Figures are illustrative demo data.
General-purpose text generation, chat, summarization and content workloads where broad capability and low cost matter most.
tongyi-deepresearch-30b-a3b supports up to a 131K-token context window — large enough for long documents and extended conversations in a single request.
Within the HotON.ai tracked set, tongyi-deepresearch-30b-a3b is cheaper than 84% of models on input price and ranks #186 of 535 by overall efficiency.
Yes — deepseek-v4-flash is a lower-cost option at $0.28 per 1M output tokens, while still covering similar Text use cases. Compare them side by side on HotON.ai.
Ready to paste into articles, papers or AI prompts — prices and date refresh with the live data.
HotON.ai — tongyi-deepresearch-30b-a3b (Alibaba Group): $0.09/1M input, $0.45/1M output, as of 2026-05-11. https://hoton.ai/en/models/alibaba-tongyi-deepresearch-30b-a3bPricing is real (via the TestKey catalog, updated daily). Quality (Arena Elo) is real where the model is ranked on LMArena. Efficiency is a modeled composite of real price and context.