Hunyuan Hy3
Tencent’s own flagship, 295B MoE with 256K context — strong reasoning and multimodal depth.
Pay-as-you-go availableBD LLM Token Hub — DeepSeek, GLM, Kimi, Hunyuan & more through a single OpenAI-compatible key. Launching soon, supported locally by Malaysia’s Tencent Official Agentic Cloud Partner.
Switch between frontier models by changing one model name — no new accounts, no new bills. Pay-as-you-go rates are also available on top of your subscription.
Tencent’s own flagship, 295B MoE with 256K context — strong reasoning and multimodal depth.
Pay-as-you-go availableBest-value coding workhorse with up to 90% cache-hit savings — ideal for AI coding agents.
Pay-as-you-go availableAgent-orchestration specialist with 1M context — built for complex multi-step workflows.
Pay-as-you-go availableState-of-the-art long-context model with 1M context — deep research and document-scale reading.
Pay-as-you-go availableMultimodal frontier model — text, image and richer inputs in one economical API.
Pay-as-you-go availablePlus DeepSeek-V4-Pro, GLM 5.1 / 5.3, Kimi K2.6 / K3, MiniMax M2.7 / M3, MiMo, Hy-MT2 translation and Kinfra embeddings — new models added continuously.
GPT, Claude and Gemini are priced for their brands. The same agent tasks — coding, analysis, content — run on DeepSeek, GLM, Kimi and Hunyuan at a fraction of the cost. Same OpenAI-compatible API, same tools, far smaller bill.
Input $5.00 vs $0.14 · Output $30.00 vs $0.28 per 1M tokens — the same coding-agent workload, a fraction of the bill.
Input $5.00 vs $1.11 · Output $25.00 vs $3.89 per 1M tokens — 1M context agent orchestration without the Opus premium.
Input $2.00 vs $0.132 · Output $12.00 vs $0.528 per 1M tokens — Tencent’s own 295B MoE flagship at a workhorse price.
Input $2.00 vs $0.435 · Output $10.00 vs $0.87 per 1M tokens — flagship-grade reasoning at a mid-tier fraction.
Input $10.00 vs $3.00 · Output $50.00 vs $15.00 per 1M tokens — the same 1M context frontier class, for less than a third of the premium.
Public list prices per 1M tokens, standard tier, August 2026. Your savings start from an RM39/mo BD LLM Token Hub plan — request early access for launch pricing.
Spending on a US or global frontier API (OpenAI, Anthropic, Google)? Tell us your monthly bill — we estimate the same workload on China’s flagship models — DeepSeek, GLM, Kimi & Hunyuan — via BD LLM Token Hub. Live, in your currency.
China’s flagship models — DeepSeek, Zhipu GLM, Moonshot Kimi, Tencent Hunyuan — on one OpenAI-compatible API: you only change the base URL and key. Estimates use public list prices (Aug 2026) at your output ratio; actual savings vary with workload, caching and volume. Gemini 3.7 Flash pricing is an intro rate valid to Jan 2027.
Pick a tier, get a pooled monthly quota of tokens, and switch between flagship models freely. Billed in MYR with a Malaysian invoice — no foreign-exchange surprises. Launching soon — request early access to be first in line.
Procurement-friendly plans for agencies, system integrators and enterprises that need shared access, per-key quotas and a Malaysian invoice.
50M+ tokens per month with raw-token billing, multiple keys and local onboarding.
Credit pools with advanced controls for larger deployments.
No lengthy sign-up flows. No juggling multiple dashboards. Request access now, then at launch grab your key and point your tools at it.
Tell us which plan you are interested in via the form on this page. We confirm your slot and notify you the moment we launch.
At launch, we deliver your key via ticket with the Singapore-region endpoint preconfigured — usually within the day.
Change the base URL and key in CodeBuddy Code, Cursor, Claude Code, OpenClaw, WorkBuddy & more. Done.
If your tool supports an OpenAI-compatible API, it works — just change the base URL and key. Here is what each one is, in plain English:
Powered by Tencent Cloud TokenHub — an OpenAI & Anthropic-compatible gateway, so existing clients only change base URL and key.
From customer-facing chatbots to autonomous coding agents — switch models per workload without switching vendors.
Multilingual support agents and FAQ bots on Kimi or GLM — with 24/7 reliability and MYR billing.
DeepSeek-V4-Flash and GLM-5.3 behind CodeBuddy Code, Cursor or Cline — at a fraction of direct vendor complexity.
Generate copy, scripts and campaign variants in volume — English, Bahasa Melayu or Chinese, from one API.
Long-context models like Kimi K3 power deep document Q&A and retrieval pipelines on your own knowledge base.
Everything Malaysian teams ask before switching to one LLM API key.
One subscription, one API key and one MYR bill — with access to DeepSeek, GLM, Kimi, Hunyuan and more. It is powered by Tencent Cloud TokenHub, a single OpenAI-compatible gateway, resold and supported locally by Big Domain.
Hunyuan Hy3, DeepSeek V4 Flash/Pro, GLM 5.1/5.3, Kimi K2.6/K3, MiniMax M2.7/M3, MiMo, Hy-MT2 and Kinfra embeddings — and you switch between them by changing one model name.
No. One TokenHub key covers every model in your plan. Sign up once, call any supported model through the same endpoint.
If your tool supports the OpenAI or Anthropic API (CodeBuddy Code, WorkBuddy, Cursor, Claude Code, Cline, OpenClaw and more), you only change the base URL and key — everything else stays the same.
Renew for the next month, or enable pay-as-you-go in the Tencent console for seamless overflow — you never get cut off mid-project.
Yes — Enterprise Lite and Enterprise Pro give you multiple API keys with per-key quotas and central management. Talk to Sales for a custom setup.
We are launching soon. Request early access and we will notify you first — early-access members get first API keys and launch pricing.
Complete the request form on this page and our team will confirm your slot, then keep you updated on the launch.
BD LLM Token Hub is coming soon. Tell us how you plan to use it and we will keep your spot warm — first API keys and launch pricing go to early-access members.