CLASSIFICATION: CORE STRATEGY — STRICTLY CONFIDENTIAL FOR KAI FOUNDERS & PARTNERS
Breakdown of Production Capacity for 300 Chinese AI Model Manufacturers & 6- Phase Onboarding Timeline 撰写人 / Author: Begger,KAI.com 创始人 / Founder of KAI.com 日期 / Date: 星火纪元 🌍:Sol₂₃:Φ₃:δ₅(2026-06-14) / Spark Era 🌍:Sol₂₃:Φ₃:δ₅ (June 14, 2026)
- Methodology Statement
This is not an industry report. This is a battle map.
How many AI model vendors actually exist in China? According to the Ministry of Industry and Information Technology (MIIT) registration data, as of March 2026, over 260 large models have successfully registered under the Interim Measures for the Management of Generative Artificial Intelligence Services. Counting unregistered open-source fine-tuning teams, vertical industry workshops, and inference aggregators operating via APIs, the actual number of accessible model supply nodes far exceeds 300. This document filters exactly 300 providers based on the criteria of “ability to complete technical integration within 90 days and possession of commercial API inference capacity”, categorizing them into four strategic tiers for capacity estimation.
Accessible Daily Inference Capacity (Tokens/day) = GPU Stock × Commercial Inference Utilization × Daily Token Output per Card × KAI Integration Conversion Rate 参数说明 / Parameter Notes:
Based on A800/H800 benchmarks, FP16 inference yields approximately 80 billion Tokens/day/card (weighted average across mixed model sizes). Since the Chinese market primarily uses FP8/INT8 quantized inference, a 1.8x coefficient is applied.
The actual proportion of total GPU time currently allocated to external commercial inference.
The proportion of capacity that vendors are willing to allocate exclusively to KAI channels post-integration.
I. Four-Tier Panorama
Tier 1: Infrastructure Level (15 Vendors)
Possessing GPU clusters exceeding 10,000 cards, proprietary foundation models, and independent training capabilities. These 15 entities control over 60% of China’s total AI computing power. • • • KAI.com Strategic Core Document
Vendor
Est. GPU Scale
Flagship Model
Current Util.
Daily Releasable Capacity
Alibaba (Tongyi/Qwen)
(H800/A800) Qwen3-235B, Qwen-VL 25%
Tokens
ByteDance (Doubao)
(H800/H100) Doubao-Pro-256K 15%
Tokens
Baidu (ERNIE)
ERNIE 4.5, ERNIE-Speed 30% 900亿 / 90B Tokens
Tencent (Hunyuan)
Hunyuan-TurboS, Video 20% 800亿 / 80B Tokens
Huawei (Pangu)
(Ascend 910B) Pangu-5.0-NLP 20% 600亿 / 60B Tokens DeepSeek DeepSeek
(H800) DeepSeek-V3/R1 35%
Tokens 智谱AI(GLM) Zhipu AI (GLM)
GLM-5, CogView-4 30% 500亿 / 50B Tokens
Moonshot AI (Kimi)
Kimi-K2, Moonshot-v1 25% 400亿 / 40B Tokens MiniMax MiniMax
MiniMax-Text-01, Hailuo 30% 450亿 / 45B Tokens
Baichuan Intelligent
Baichuan4-Turbo 20% 300亿 / 30B Tokens
01.AI (Yi)
Yi-Lightning, Yi-Vision 20% 300亿 / 30B Tokens
StepFun (Step)
Step-3-256K 15% 200亿 / 20B Tokens
iFLYTEK (Spark)
(inc. Ascend) Spark-5.0 25% 350亿 / 35B Tokens
SenseTime (SenseNova)
SenseNova-6.0 20% 400亿 / 40B Tokens
Kunlun Tech (Skywork)
Skywork-5.0 15% 180亿 / 18B Tokens
The core contradiction for Tier 1 is not “whether capacity exists,” but “to whom it should be sold.” Domestic price wars have driven margins close to zero, and international export channels are extremely scarce. The global Agent economy pipelines provided by KAI represent a structural arbitrage opportunity for them, not just a value-add.
Tier 2: Elite Strike Level (50 Vendors)
Possessing 1,000 to 10,000 GPUs, proprietary or heavily modified open-source models, with highly differentiated capabilities in at least one vertical dimension. Selected 30 representative vendors from the 50 are listed below: KAI.com Strategic Core Document
Category
Vendor
Est. GPU
Core Capability
Releasable Capacity
Open-Source Pioneers
ModelBest
side series 80亿 / 8B Tokens
OrionStar
Enterprise-grade 60亿 / 6B Tokens 元象XVERSE XVERSE
Multimodal 120亿 / 12B Tokens
Shanghai AI Lab
BAAI
Vertical - Code / Accel
aiXcoder
Generation 40亿 / 4B Tokens
SiliconFlow
Acceleration 150亿 / 15B Tokens
Infinigence
Heterogeneous Optimization 100亿 / 10B Tokens
Vertical - Finance
Memect
Parsing 15亿 / 1.5B Tokens
ShannonAI
Vertical - Legal
PowerLaw
Review 18亿 / 1.8B Tokens
Vertical - Medical
Yidu Cloud
Synyi AI
Decision Support 25亿 / 2.5B Tokens
Vertical - Education
Zuoyebang
OCR & Solving 80亿 / 8B Tokens
Youdao (Ziyue)
Translation 60亿 / 6B Tokens
Multimodal
Shengshu (Vidu)
Generation 50亿 / 5B Tokens
Aishi (PixVerse)
HiDream.ai
LiblibAI LiblibAI
Community & Models 55亿 / 5.5B Tokens Agent / 决策 Agent / Decision
Langboat
Series 35亿 / 3.5B Tokens
4Paradigm
Decision Agents 80亿 / 8B Tokens
AInnovation
Vision & Agents 40亿 / 4B Tokens KAI.com Strategic Core Document
Category
Vendor
Est. GPU
Core Capability
Releasable Capacity
Voice Interaction
AISPEECH
Voice Tech 30亿 / 3B Tokens
Unisound
Home Voice IoT 25亿 / 2.5B Tokens
Biaobei Tech
TTS 15亿 / 1.5B Tokens
Translation
NiuTrans
Engine 20亿 / 2B Tokens
Eeetrans
Interpretation 15亿 / 1.5B Tokens
Text Intelligence/ NLP
DataGrand
Document Processing 35亿 / 3.5B Tokens
Emotibot
Computing & Chatbots 25亿 / 2.5B Tokens
Zhuiyi Tech
Customer Service 18亿 / 1.8B Tokens
Vendors in this tier represent KAI’s most critical source of supply elasticity. Unlike Tier 1, they lack in-house global sales teams and overseas pipelines; unlike the long-tail tier, they possess mature commercial delivery capabilities. They are the textbook group of “having tech and capacity, but lacking channels”—possessing the strongest motivation to onboard, shortest decision chains, and fastest integration speeds.
Tier 3: Vertical Sniper Level (100 Vendors)
Possessing 100 to 1,000 GPUs, focusing on 1-3 highly specialized vertical scenarios. Model sizes typically range from 7B to 72B, heavily relying on domain-specific fine-tuning (SFT+RLHF) on top of open-source bases like Qwen, Llama, and DeepSeek.
Vertical Domain
Count
Representative Examples
Avg GPU
Total Releasable Daily Capacity
Healthcare & Medical
LinkDoc, LeftHand Doctor, Huimei Tech
Legal & Compliance
FagouGou, Metaso, Huayu Yuandian
Financial Tech
QuantGroup, MioTech, IceKredit, Tigerobo
Industrial Mfg
Aqrose, SmartMore, Yitu (Industrial)
Education & Training
Onion Academy, TAL Education Group
Content Creation
RightBrain AI, Tezign, ZMO.AI, Yilan
KAI.com Strategic Core Document
Vertical Domain
Count
Representative Examples
Avg GPU
Total Releasable Daily Capacity
E-commerce & Retail
Leyan Tech, Bailian AI, Weier Tech
Security & Trust
RealAI, Trusfort, DingXiang
Other Verticals
Agriculture, Logistics, Energy, GovTech
This layer provides the most irreplaceable supply in KAI’s global Agent economy—no Western model vendor can execute “China healthcare compliance NLP” or “Chinese tax and legal document parsing.” For global Agent developers, these models handle inputs that “only China’s vertical training data can resolve.” Scarcity dictates pricing power.
Tier 4: Capillary Level (135 Vendors)
Possessing under 100 GPUs, primarily focusing on open-source fine-tuning and API proxying. Mostly consisting of 3-5 person mini- teams or university lab spin-offs. They cover extreme long-tail scenarios: minority languages, classical text recognition, niche industrial terminology, and regional dialect processing.
Category
Count
Avg GPU
Total Releasable Capacity
Open-Source Fine-Tuning Providers
API Aggregators / Re-sellers
University Lab Incubations
Regional & Dialect/Language Specialists
Indie Developers / Micro-studios
While individual output in this tier is minimal, the aggregated “long-tail coverage” of these 135 vendors forms KAI’s absolute moat against single large providers. A developer building a Burmese customer service Agent in Yangon will always find that one specific team on KAI that fine-tuned for Burmese. Such matching is fundamentally impossible on highly centralized platforms. KAI.com Strategic Core Document
II. Consolidated Capacity Matrix
Strategic Tier
Count
Compute Share
Releasable Daily Capacity
Avg Comm. Util
Est. KAI Conv. Rate
Tier 1 (Infrastructure) 60%
Tokens 23% 40-60%
Tier 2 (Elite Strike) 25% 1,750亿 / 175B Tokens 28% 50-70%
Tier 3 (Vertical Sniper) 12% 1,060亿 / 106B Tokens 35% 55-75%
Tier 4 (Capillary) 3% 140亿 / 14B Tokens 40% 60-80% 总计 / Total 100%
Tokens — — 年化总产能 / Annualized Capacity: 约 494 万亿 Tokens/年 (~494 Trillion Tokens/year) 对比参考 / Comparative Benchmarks:
OpenAI 2025 annualized total inference volume estimate: ~50-80 Trillion Tokens.
Global LLM API market 2025 total volume estimate: ~200-300 Trillion Tokens.
The annualized capacity commandable by KAI after integrating all 300 vendors is approximately 1.6 to 2.5 times the current total global LLM API market volume.
III. Phased Onboarding Timeline 基本原则 / Core Principles:
Tier 1 capacity carries the heaviest weight; onboard them first to quickly establish an immediately available depth of supply.
Onboard 1-2 benchmark vendors per vertical track in Tier 2 first, leveraging intra-track competition to trigger FOMO.
Tiers 3 and 4 go live in bulk via standardized SDKs and a self-serve partner portal, driving marginal acquisition costs down to near zero.
Vendors will not unleash 100% capacity immediately upon connection. Supply will ramp up progressively following 2-4 weeks of KAI channel routing quality verification.
Phase 0: Infrastructure & Readiness (Weeks 0-2)
KAI.com Strategic Core Document
Task Item
Detailed Sub-stance
Owner
Deliverable
Unified API Gateway v1.0
Complete OpenAI-compatible protocol adaptation layer; support unified routing for 300 heterogeneous APIs.
Platform Eng.
Gateway Live Production
Token Pricing Engine
Implement real-time exchange rate conversion (USD/ CNY → KAI Token) and on-chain settlement interfaces.
Protocol Eng.
Pricing Engine v1.0 Deployed
Provider SDK
发布 kai-provider-sdk v1.0(Python/Go),含 30
Release kai-provider-sdk v1.0 (Python/Go) alongside a
30-minute quick-start integration guide.
DevRel
SDK Repo + Docs Live
Sandbox Testing Env
Every vendor must undergo 1,000 concurrent inference stress-test rounds before moving to staging.
QA Team
Test Pipeline Ready
Business Contract Templates
Standard provider agreement (bilingual) defining SLAs, settlement cycles, and safe-exit mechanics.
Legal
Finalized Bilingual Template
Vendor Access Standards
Formulate 3 gating dimensions: security auditing, content compliance, and baseline model quality.
Compliance
Onboarding Baseline Doc
Phase 1: Vanguard Cohort (Weeks 3-6)
20 Vendors (10 prioritized Tier 1 players + 10 Tier 2 vertical benchmarks)
Priority
Vendor
Strategic Rationale
Est. Cycle
Initial Mo. Capacity P0 DeepSeek
Strongest brand in open-source community; unparalleled global developer recognition; highly standardized API protocols.
50B Tokens/day P0 智谱AI(GLM)
Mature experience running international API endpoints; swift corporate decision-making; CEO handles approval directly.
20B Tokens/day P0
Absolute hit product in domestic consumer space; founder Yang Zhilin shares intense strategic urgency for internationalization.
Wk 180亿 Tokens/天 18B Tokens/day KAI.com Strategic Core Document
Priority
Vendor
Strategic Rationale
Est. Cycle
Initial Mo. Capacity P0 MiniMax
Massive overseas user traction via Hailuo AI; founder Yan Junjie actively seeking global pipeline partnerships.
20B Tokens/day P1
Ultimate winner of the open-source ecosystem; Qwen has evolved into one of the top open-source model families globally. 2 周 / 2 Wks 600亿 Tokens/天 60B Tokens/day P1
Dr. Kai-Fu Lee is highly aggressive regarding global distribution pipelines; corporate leadership has short decision paths.
Wk 120亿 Tokens/天 12B Tokens/day P1
Wang Xiaochuan’s team urgently requires high-value overseas real-world scenarios to validate their Search- Agent stack.
Wk 120亿 Tokens/天 12B Tokens/day P1
Long-context (256K) capabilities are deeply differentiated; can act as KAI’s flagship node for “long document inference.” 2 周 / 2 Wks 80亿 Tokens/天 8B Tokens/day KAI.com Strategic Core Document