CLASSIFICATION: CORE STRATEGY — STRICTLY CONFIDENTIAL FOR KAI FOUNDERS & PARTNERS

Breakdown of Production Capacity for 300 Chinese AI Model Manufacturers & 6- Phase Onboarding Timeline 撰写人 / Author: Begger,KAI.com 创始人 / Founder of KAI.com 日期 / Date: 星火纪元 🌍:Sol₂₃:Φ₃:δ₅(2026-06-14) / Spark Era 🌍:Sol₂₃:Φ₃:δ₅ (June 14, 2026)

  1. Methodology Statement

This is not an industry report. This is a battle map.

How many AI model vendors actually exist in China? According to the Ministry of Industry and Information Technology (MIIT) registration data, as of March 2026, over 260 large models have successfully registered under the Interim Measures for the Management of Generative Artificial Intelligence Services. Counting unregistered open-source fine-tuning teams, vertical industry workshops, and inference aggregators operating via APIs, the actual number of accessible model supply nodes far exceeds 300. This document filters exactly 300 providers based on the criteria of “ability to complete technical integration within 90 days and possession of commercial API inference capacity”, categorizing them into four strategic tiers for capacity estimation.

Accessible Daily Inference Capacity (Tokens/day) = GPU Stock × Commercial Inference Utilization × Daily Token Output per Card × KAI Integration Conversion Rate 参数说明 / Parameter Notes:

Based on A800/H800 benchmarks, FP16 inference yields approximately 80 billion Tokens/day/card (weighted average across mixed model sizes). Since the Chinese market primarily uses FP8/INT8 quantized inference, a 1.8x coefficient is applied.

The actual proportion of total GPU time currently allocated to external commercial inference.

The proportion of capacity that vendors are willing to allocate exclusively to KAI channels post-integration.

I. Four-Tier Panorama

Tier 1: Infrastructure Level (15 Vendors)

Possessing GPU clusters exceeding 10,000 cards, proprietary foundation models, and independent training capabilities. These 15 entities control over 60% of China’s total AI computing power. • • • KAI.com Strategic Core Document

Vendor

Est. GPU Scale

Flagship Model

Current Util.

Daily Releasable Capacity

Alibaba (Tongyi/Qwen)

(H800/A800) Qwen3-235B, Qwen-VL 25%

Tokens

ByteDance (Doubao)

(H800/H100) Doubao-Pro-256K 15%

Tokens

Baidu (ERNIE)

ERNIE 4.5, ERNIE-Speed 30% 900亿 / 90B Tokens

Tencent (Hunyuan)

Hunyuan-TurboS, Video 20% 800亿 / 80B Tokens

Huawei (Pangu)

(Ascend 910B) Pangu-5.0-NLP 20% 600亿 / 60B Tokens DeepSeek DeepSeek

(H800) DeepSeek-V3/R1 35%

Tokens 智谱AI(GLM) Zhipu AI (GLM)

GLM-5, CogView-4 30% 500亿 / 50B Tokens

Moonshot AI (Kimi)

Kimi-K2, Moonshot-v1 25% 400亿 / 40B Tokens MiniMax MiniMax

MiniMax-Text-01, Hailuo 30% 450亿 / 45B Tokens

Baichuan Intelligent

Baichuan4-Turbo 20% 300亿 / 30B Tokens

01.AI (Yi)

Yi-Lightning, Yi-Vision 20% 300亿 / 30B Tokens

StepFun (Step)

Step-3-256K 15% 200亿 / 20B Tokens

iFLYTEK (Spark)

(inc. Ascend) Spark-5.0 25% 350亿 / 35B Tokens

SenseTime (SenseNova)

SenseNova-6.0 20% 400亿 / 40B Tokens

Kunlun Tech (Skywork)

Skywork-5.0 15% 180亿 / 18B Tokens

The core contradiction for Tier 1 is not “whether capacity exists,” but “to whom it should be sold.” Domestic price wars have driven margins close to zero, and international export channels are extremely scarce. The global Agent economy pipelines provided by KAI represent a structural arbitrage opportunity for them, not just a value-add.

Tier 2: Elite Strike Level (50 Vendors)

Possessing 1,000 to 10,000 GPUs, proprietary or heavily modified open-source models, with highly differentiated capabilities in at least one vertical dimension. Selected 30 representative vendors from the 50 are listed below: KAI.com Strategic Core Document

Category

Vendor

Est. GPU

Core Capability

Releasable Capacity

Open-Source Pioneers

ModelBest

side series 80亿 / 8B Tokens

OrionStar

Enterprise-grade 60亿 / 6B Tokens 元象XVERSE XVERSE

Multimodal 120亿 / 12B Tokens

Shanghai AI Lab

BAAI

Vertical - Code / Accel

aiXcoder

Generation 40亿 / 4B Tokens

SiliconFlow

Acceleration 150亿 / 15B Tokens

Infinigence

Heterogeneous Optimization 100亿 / 10B Tokens

Vertical - Finance

Memect

Parsing 15亿 / 1.5B Tokens

ShannonAI

Vertical - Legal

PowerLaw

Review 18亿 / 1.8B Tokens

Vertical - Medical

Yidu Cloud

Synyi AI

Decision Support 25亿 / 2.5B Tokens

Vertical - Education

Zuoyebang

OCR & Solving 80亿 / 8B Tokens

Youdao (Ziyue)

Translation 60亿 / 6B Tokens

Multimodal

Shengshu (Vidu)

Generation 50亿 / 5B Tokens

Aishi (PixVerse)

HiDream.ai

LiblibAI LiblibAI

Community & Models 55亿 / 5.5B Tokens Agent / 决策 Agent / Decision

Langboat

Series 35亿 / 3.5B Tokens

4Paradigm

Decision Agents 80亿 / 8B Tokens

AInnovation

Vision & Agents 40亿 / 4B Tokens KAI.com Strategic Core Document

Category

Vendor

Est. GPU

Core Capability

Releasable Capacity

Voice Interaction

AISPEECH

Voice Tech 30亿 / 3B Tokens

Unisound

Home Voice IoT 25亿 / 2.5B Tokens

Biaobei Tech

TTS 15亿 / 1.5B Tokens

Translation

NiuTrans

Engine 20亿 / 2B Tokens

Eeetrans

Interpretation 15亿 / 1.5B Tokens

Text Intelligence/ NLP

DataGrand

Document Processing 35亿 / 3.5B Tokens

Emotibot

Computing & Chatbots 25亿 / 2.5B Tokens

Zhuiyi Tech

Customer Service 18亿 / 1.8B Tokens

Vendors in this tier represent KAI’s most critical source of supply elasticity. Unlike Tier 1, they lack in-house global sales teams and overseas pipelines; unlike the long-tail tier, they possess mature commercial delivery capabilities. They are the textbook group of “having tech and capacity, but lacking channels”—possessing the strongest motivation to onboard, shortest decision chains, and fastest integration speeds.

Tier 3: Vertical Sniper Level (100 Vendors)

Possessing 100 to 1,000 GPUs, focusing on 1-3 highly specialized vertical scenarios. Model sizes typically range from 7B to 72B, heavily relying on domain-specific fine-tuning (SFT+RLHF) on top of open-source bases like Qwen, Llama, and DeepSeek.

Vertical Domain

Count

Representative Examples

Avg GPU

Total Releasable Daily Capacity

Healthcare & Medical

LinkDoc, LeftHand Doctor, Huimei Tech

Legal & Compliance

FagouGou, Metaso, Huayu Yuandian

Financial Tech

QuantGroup, MioTech, IceKredit, Tigerobo

Industrial Mfg

Aqrose, SmartMore, Yitu (Industrial)

Education & Training

Onion Academy, TAL Education Group

Content Creation

RightBrain AI, Tezign, ZMO.AI, Yilan

KAI.com Strategic Core Document

Vertical Domain

Count

Representative Examples

Avg GPU

Total Releasable Daily Capacity

E-commerce & Retail

Leyan Tech, Bailian AI, Weier Tech

Security & Trust

RealAI, Trusfort, DingXiang

Other Verticals

Agriculture, Logistics, Energy, GovTech

This layer provides the most irreplaceable supply in KAI’s global Agent economy—no Western model vendor can execute “China healthcare compliance NLP” or “Chinese tax and legal document parsing.” For global Agent developers, these models handle inputs that “only China’s vertical training data can resolve.” Scarcity dictates pricing power.

Tier 4: Capillary Level (135 Vendors)

Possessing under 100 GPUs, primarily focusing on open-source fine-tuning and API proxying. Mostly consisting of 3-5 person mini- teams or university lab spin-offs. They cover extreme long-tail scenarios: minority languages, classical text recognition, niche industrial terminology, and regional dialect processing.

Category

Count

Avg GPU

Total Releasable Capacity

Open-Source Fine-Tuning Providers

API Aggregators / Re-sellers

University Lab Incubations

Regional & Dialect/Language Specialists

Indie Developers / Micro-studios

While individual output in this tier is minimal, the aggregated “long-tail coverage” of these 135 vendors forms KAI’s absolute moat against single large providers. A developer building a Burmese customer service Agent in Yangon will always find that one specific team on KAI that fine-tuned for Burmese. Such matching is fundamentally impossible on highly centralized platforms. KAI.com Strategic Core Document

II. Consolidated Capacity Matrix

Strategic Tier

Count

Compute Share

Releasable Daily Capacity

Avg Comm. Util

Est. KAI Conv. Rate

Tier 1 (Infrastructure) 60%

Tokens 23% 40-60%

Tier 2 (Elite Strike) 25% 1,750亿 / 175B Tokens 28% 50-70%

Tier 3 (Vertical Sniper) 12% 1,060亿 / 106B Tokens 35% 55-75%

Tier 4 (Capillary) 3% 140亿 / 14B Tokens 40% 60-80% 总计 / Total 100%

Tokens — — 年化总产能 / Annualized Capacity: 约 494 万亿 Tokens/年 (~494 Trillion Tokens/year) 对比参考 / Comparative Benchmarks:

OpenAI 2025 annualized total inference volume estimate: ~50-80 Trillion Tokens.

Global LLM API market 2025 total volume estimate: ~200-300 Trillion Tokens.

The annualized capacity commandable by KAI after integrating all 300 vendors is approximately 1.6 to 2.5 times the current total global LLM API market volume.

III. Phased Onboarding Timeline 基本原则 / Core Principles:

Tier 1 capacity carries the heaviest weight; onboard them first to quickly establish an immediately available depth of supply.

Onboard 1-2 benchmark vendors per vertical track in Tier 2 first, leveraging intra-track competition to trigger FOMO.

Tiers 3 and 4 go live in bulk via standardized SDKs and a self-serve partner portal, driving marginal acquisition costs down to near zero.

Vendors will not unleash 100% capacity immediately upon connection. Supply will ramp up progressively following 2-4 weeks of KAI channel routing quality verification.

Phase 0: Infrastructure & Readiness (Weeks 0-2)

KAI.com Strategic Core Document

Task Item

Detailed Sub-stance

Owner

Deliverable

Unified API Gateway v1.0

Complete OpenAI-compatible protocol adaptation layer; support unified routing for 300 heterogeneous APIs.

Platform Eng.

Gateway Live Production

Token Pricing Engine

Implement real-time exchange rate conversion (USD/ CNY → KAI Token) and on-chain settlement interfaces.

Protocol Eng.

Pricing Engine v1.0 Deployed

Provider SDK 发布 kai-provider-sdk v1.0(Python/Go),含 30

Release kai-provider-sdk v1.0 (Python/Go) alongside a 30-minute quick-start integration guide.

DevRel

SDK Repo + Docs Live

Sandbox Testing Env

Every vendor must undergo 1,000 concurrent inference stress-test rounds before moving to staging.

QA Team

Test Pipeline Ready

Business Contract Templates

Standard provider agreement (bilingual) defining SLAs, settlement cycles, and safe-exit mechanics.

Legal

Finalized Bilingual Template

Vendor Access Standards

Formulate 3 gating dimensions: security auditing, content compliance, and baseline model quality.

Compliance

Onboarding Baseline Doc

Phase 1: Vanguard Cohort (Weeks 3-6)

20 Vendors (10 prioritized Tier 1 players + 10 Tier 2 vertical benchmarks)

Priority

Vendor

Strategic Rationale

Est. Cycle

Initial Mo. Capacity P0 DeepSeek

Strongest brand in open-source community; unparalleled global developer recognition; highly standardized API protocols.

50B Tokens/day P0 智谱AI(GLM)

Mature experience running international API endpoints; swift corporate decision-making; CEO handles approval directly.

20B Tokens/day P0

Absolute hit product in domestic consumer space; founder Yang Zhilin shares intense strategic urgency for internationalization.

Wk 180亿 Tokens/天 18B Tokens/day KAI.com Strategic Core Document

Priority

Vendor

Strategic Rationale

Est. Cycle

Initial Mo. Capacity P0 MiniMax

Massive overseas user traction via Hailuo AI; founder Yan Junjie actively seeking global pipeline partnerships.

20B Tokens/day P1

Ultimate winner of the open-source ecosystem; Qwen has evolved into one of the top open-source model families globally. 2 周 / 2 Wks 600亿 Tokens/天 60B Tokens/day P1

Dr. Kai-Fu Lee is highly aggressive regarding global distribution pipelines; corporate leadership has short decision paths.

Wk 120亿 Tokens/天 12B Tokens/day P1

Wang Xiaochuan’s team urgently requires high-value overseas real-world scenarios to validate their Search- Agent stack.

Wk 120亿 Tokens/天 12B Tokens/day P1

Long-context (256K) capabilities are deeply differentiated; can act as KAI’s flagship node for “long document inference.” 2 周 / 2 Wks 80亿 Tokens/天 8B Tokens/day KAI.com Strategic Core Document