
TokenDance
Officially listedAn LLM API aggregation gateway from the Watcha team that unifies 97 Chinese models behind one key.
TokenDance
TokenDance is an LLM API aggregation gateway from the team behind the AI product community Watcha, self-described as "a Chinese OpenRouter", unifying access to 97 Chinese models from 16+ providers behind a single key. It packages multi-model access, intelligent routing, unified billing, and failover into one developer-friendly layer, mainly serving independent developers, agent teams, and early-stage AI startups in mainland China. Note: the same team also runs the Watcha AI product community — two products from the same people, not duplicate entries.
Core Capabilities
- Native three protocols: Natively supports OpenAI / Anthropic (Claude) / Google GenAI (Gemini) — official SDKs connect by changing only the base URL and key, unlike relays that only emulate OpenAI
- Smart routing and failover: Routes automatically by model name to the right provider, switches automatically among multiple endpoints of the same model, and supports candidate lists via the extended models:[...] field for stepped fallback
- Unified billing: Cross-provider token accounting and balance deduction in one place — one dashboard for all usage, no topping up vendor by vendor
- Rich model catalog: 97 models verified live via the API (the homepage's "8+" is stale), including MiniMax, Qwen, Kimi, Zhipu, DeepSeek, and Doubao
- Developer infrastructure: Playground, API key management, call logs, async webhooks, OAuth-style key authorization, and a balance-query Open API
- AI-facing docs: Provides llms.txt / llms-full.txt so Claude Code, Cursor, and similar tools can integrate on their own
Use Cases Independent developers, one-person companies, and early AI teams in mainland China, especially Claude Code / Codex / Cline users blocked by "OpenAI-compatible-only" relays, and agent developers needing low-cost multi-model switching with failover. Not for non-technical users who never write code (user reviews agree it is "of no use to ordinary users").
Unique Advantages Against the many OpenAI-compatible-only relays in China, native three-protocol support is a hard differentiator — Anthropic-format tools connect directly, rare among Chinese gateways. The background is also more verifiable than a typical "individual webmaster relay": the operator Watcha (Hangzhou Siyi Technology Co., Ltd.) has Hangzhou city government media coverage and a Sequoia China seed round, with underlying compute supplied by Tsinghua-affiliated Infinigence (over ¥2.2 billion raised to date) under a publicly declared strategic partnership; per-model prices are public and subsidies generous. Shortcomings are equally clear: almost no overseas models, and fewer models in absolute terms than OpenRouter or SiliconFlow.
Editor's Verdict The product is real — 97 working models, high-quality documentation engineering, and native three-protocol support beyond its peers, making it a legitimate choice as a domestic multi-model gateway. But the risks must be stated honestly: no terms of service or privacy policy anywhere on the site (the login page says "by logging in you agree to the terms of service" yet that link 404s, with no public terms on how data is handled or how keys are stored); the operating entity is not disclosed on the site (the Zhejiang ICP filing number 2024107375 points to Hangzhou Siyi Technology Co., Ltd.); the domain was registered only on 2026-03-25, roughly six months ago; official PR and changelog contradict each other on whether online top-ups are supported; some vendor logos are hot-linked directly from their CDNs; and independent third-party reviews are essentially zero — praise on Watcha comes from the same source as the vendor, a conflict of interest. Disambiguation: searching "TokenDance" also surfaces a UCLA arXiv paper and the GitHub org TokenDanceLab, both unrelated to this product. Use caution for sensitive data or strict compliance scenarios.
Pricing
### 💰 定价模式:纯按量付费(预充值额度制,无订阅) **起步价**:免费注册赠额度/按量扣费,最低约 ¥0.14 / M tokens #### 主要单价(人民币,均以 /models 公开价格表为准) - **DeepSeek V4.1 Flash**:输入 ¥0.6 / M tokens,输出 ¥2.4 / M tokens(限时 6 折) - **Z.ai GLM 5.3**:输入 ¥7.2 / M,输出 ¥25.2 / M(Flash 档为 ¥0.72 / ¥2.52) - **Kimi K2.6**:输入 ¥5.2 / M,输出 ¥21.6 / M - **视频生成(MiniMax H3 Max / 可灵类)**:¥0.28–2.4 / 秒 - **Seedream 图像**:输出 ¥0.15–0.3 / 张,输入 ¥0.01 / 张(首张免费) - **Bocha 联网搜索**:¥0.028 / 次 #### 峰谷与折扣 DeepSeek 官方端点空闲时段 = 高峰时段 **5 折**(高峰为北京时间周一至周五 9:00–12:00、14:00–18:00);模型卡片普遍带「省 X%」折扣标签。 #### 其他信息 - **无订阅套餐**:全站无月费/年费/会员层级,纯按实际消耗 token 从余额扣减。 - **无独立定价页**:该站 `/pricing`、`/plans` 等均为空壳,**真实价格表即模型列表页 https://tokendance.space/models**(官方文档明示「所有模型的定价在模型列表页面公开透明」),无需登录即可查看逐模型单价。 - **免费额度**:内测期有「百亿 Token 补贴计划」;接入观猹统一登录体系的产品可获千元级补贴;新用户注册赠额(第三方有 10 元 / 20 元两说,官方未明示);办观猹×浦发联名卡前 200 名送 100 元 Token。 - **发票/对公**:官方称对公支付与开票能力将在正式商业化后明确,当前未定型。 — Visit website
FAQ
Is TokenDance free?
Registration comes with free trial credits, and during the beta there is a "tens-of-billions token subsidy program" — products integrated with Watcha unified login can receive subsidies worth thousands of yuan. But it is a pure pay-as-you-go credit product: you pay for what you use, with no unlimited free tier.
What are TokenDance's main features?
The core is aggregating 97 Chinese LLMs behind a unified API gateway: native support for the OpenAI/Anthropic/Gemini protocols, automatic routing by model, multi-endpoint failover for the same model, cross-provider unified token billing, plus a Playground, logs, webhooks, and OAuth-style key authorization.
How is TokenDance priced?
It uses prepaid credits deducted by actual token usage, with no subscription plans. For example DeepSeek V4.1 Flash is ¥0.6 input / ¥2.4 output per million tokens, Kimi K2.6 is ¥5.2 / ¥21.6, video is about ¥0.28–2.4/second, and web search is ¥0.028 per search. DeepSeek is half price during off-peak hours.
Which models and protocols does TokenDance support?
Testing the API covers 97 models from 16+ providers (MiniMax, Qwen, Kimi, Zhipu, DeepSeek, Doubao, and others). Protocol-wise it natively supports OpenAI, Anthropic (Claude), and Google GenAI (Gemini), with dedicated endpoints for image, video, speech, and web search; overseas models are almost absent.
How does TokenDance differ from OpenRouter?
Both are multi-model aggregation gateways. TokenDance differs in native three-protocol support (rather than OpenAI-compatible only), Watcha one-click login and ecosystem subsidies tied to the Watcha community, and low prices from Infinigence's domestic compute; its weaknesses are model count (97 vs 300–500+) and clearly less overseas model coverage.
What is the relationship between TokenDance and Watcha?
Both come from the same team: Watcha is an AI product community, and TokenDance is the model infrastructure of its ecosystem. Watcha's official recruiting page lists TokenDance as a team business, and TokenDance's homepage uses analysis.watcha.cn analytics and "Watcha one-click login".
Does TokenDance have terms of service and a privacy policy?
No. /docs/terms.md, /docs/privacy.md, and the rest all return 404, yet the login page says "by logging in you agree to the terms of service" with a link that won't open. This is its most substantive risk: there are no public terms on how data is processed, how API keys are stored, or whether request content is retained.
Is TokenDance the same as TokenDanceLab on GitHub?
No. The GitHub org TokenDanceLab (domain tokendancelab.com) is an independent same-named project with no cross-links or shared entity, and UCLA's same-named TokenDance paper on arXiv is also unrelated. Before use, make sure you are on the official domain tokendance.space.