English Original
RAND Corporation · Research Report RR-A5043-1

AI Development in the United States and China: Evidence from a New Dataset of AI Developer Firms

An empirical analysis of 1,181 AI developer firms reveals shared technical foundations, sharp divergence in commercial application, and the strategic significance of physical embodiment.
AuthorsJon Schmid, Prateek Puri, Vitor Melo, Anthony Hakim
PublishedAugust 17, 2026
InstitutionRAND Corporation (Santa Monica, CA)
Reading Time~22 min · Original Research
中文精译
兰德智库前沿权威研究报告 · RR-A5043-1

中美人工智能产业生态实证对照:基于 1,181 家 AI 研发企业的全景剖析

基于 1,181 家中美商业 AI 研发企业的首份微观企业全量实证数据集,深度揭示两国在底层技术底座上的高度趋同、在商业应用落地路径上的关键分化,以及物理具身智能对于大国战略竞争格局的深远意义。
作者团队Jon Schmid, Prateek Puri, Vitor Melo, Anthony Hakim (RAND)
发布时间2026 年 8 月 17 日
发布机构兰德公司(RAND Corporation)
阅读估时约 22 分钟 · 全文精译
English Original
EXECUTIVE SUMMARY & CORE FINDINGS

Key Research Findings & Executive Summary

Public and policy discourse on artificial intelligence competition between the United States and China has been overwhelmingly dominated by top-level foundation model benchmarks. However, such evaluations overlook how AI capabilities actually diffuse into commercial, industrial, and national security ecosystems. In this landmark study, RAND researchers construct the first comprehensive empirical dataset tracking 1,181 commercial AI developer firms across the United States (743 firms) and China (438 firms) to uncover the microeconomic realities of national AI capacity.

  • Architectural and Modality Convergence: AI developer firms in both nations share fundamentally similar technical foundations. Approximately 88 percent of US firms and 84 percent of Chinese firms employ Transformer or diffusion architectures. Across both nations, only ~20 percent train frontier foundation models, while ~80 percent concentrate on task-specific or domain-adapted models.
  • Commercial Application Divergence: The primary divergence between the two ecosystems lies in commercial deployment. 61 percent of US AI developer firms build purely digital software (enterprise SaaS, developer tooling, health informatics, cybersecurity). In sharp contrast, 74 percent of Chinese firms build physical AI systems (industrial manufacturing automation, robotics, autonomous transportation, smart energy grids).
  • Open-Weight Diffusion and Global Token Parity: Empirical telemetry from global API routers (including OpenRouter) reveals that Chinese open-weight models reached token volume parity with leading US closed models in early 2026. Priced at 1/4 to 1/5 the inference cost of proprietary alternatives, these models have captured massive international developer adoption.
  • Strategic Policy Blind Spot: US industrial and security policy has focused almost exclusively on frontier LLM benchmark leaderboards. This narrow lens overlooks the rapid accumulation of proprietary physical interaction data in Chinese industrial supply chains, which forms a durable, self-reinforcing flywheel.
中文精译
决策摘要与核心实证结论

核心实证结论与决策摘要

关于中美人工智能竞争的公众与政策讨论,长期被顶尖基础模型的跑分基准(如 LMArena、MMLU、Codeforces)所主导。然而,此类基准评测严重脱离了 AI 技术在商业、工业与国家安全生态中的真实扩散机制。在本项里程碑式实证研究中,兰德公司研究团队构建了全球首个涵盖中美 1,181 家商业 AI 研发企业(美国 743 家,中国 438 家)的微观企业全景数据集,从产业微观层面揭示两国 AI 综合能力的真实图景。

  • 底层架构与模型范式高度趋同:两国 AI 研发企业在底层技术底座上呈现出惊人的收敛性。约 88% 的美国企业和 84% 的中国企业均采用 Transformer 或扩散模型架构。在两国企业中,仅有约 20% 专注于通用前沿基础模型训练,而高达 80% 的企业均深耕于特定任务或行业垂类适配模型。
  • 商业化落地应用路径显著分流:中美 AI 生态真正的分水岭在于商业化落地的主战场。高达 61% 的美国研发企业集中在纯数字软件领域(企业级 SaaS、开发者辅助工具、医疗信息分析、网络安全防御);与此形成鲜明对照的是,74% 的中国 AI 研发企业专注于物理具身系统(智能制造自动化、工业机器人、自动驾驶、智慧电网调度)。
  • 开源权重扩散与全球路由 Token 平价:对全球主流 API 路由平台(包括 OpenRouter)的实证遥测数据显示,中国开源权重模型在全球调用量上已于 2026 年初与美国顶尖闭源模型达到 Token 平价。凭借仅为闭源模型 1/4 至 1/5 的极低推理成本,中国开源权重模型正在快速重塑全球开发者的技术选型。
  • 国家战略评估的重大认知盲区:美国当前的产业与安全政策过度聚焦于前沿大语言模型的基准跑分榜单。这种单一视角忽视了中国工业制造腹地正在快速累积的物理交互数据飞轮,而物理世界的数据壁垒远比数字文本更为坚固且难以跨国复制。
English Original

Chapter 1: The Empirical Void in US-China AI Discourse

The Benchmark Illusion and National Capacity

For more than three years, national evaluations of artificial intelligence leadership have relied almost exclusively on public benchmark leaderboards. Metrics derived from evaluations such as LMArena, MMLU, GSM8K, and HumanEval are routinely cited by policymakers, intelligence analysts, and technology journalists as definitive scorecards of geopolitical standing. When an American laboratory releases a model scoring two points higher on graduate-level reasoning, commentators declare the United States insurmountable; when a Chinese laboratory closes the gap weeks later, headlines proclaim immediate parity.

This fixation on benchmark rankings produces a dangerous analytical distortion. Benchmark scores evaluate isolated algorithmic reasoning in static synthetic environments. They offer virtually no visibility into how models are deployed across commercial value chains, how effectively enterprises integrate automated systems into daily operations, or whether national industries are building enduring data feedback loops. An economy that excels purely at benchmark mathematics while failing to integrate AI into manufacturing, energy, and logistics risks misjudging its true industrial strength.

To move beyond this empirical void, strategic analysis requires a shift from macro benchmark posturing to micro firm-level observation. By cataloging what AI developer firms are actually building, what modalities they prioritize, and which economic sectors they serve, analysts can construct an accurate assessment of national technological competitiveness.

中文精译

第一章:中美 AI 竞争叙事中的实证赤字

基准排行榜的迷思与真实国家能力测度

三年多以来,针对中美两国人工智能竞争格局的国家级战略评估,几乎全盘受制于公开基准评测排行榜。诸如 LMArena、MMLU、GSM8K 以及 HumanEval 等跑分数据,频繁被两国的政策制定者、情报分析官以及主流科技媒体援引,作为判定大国科技地缘实力的决定性计分牌。一旦某家美国顶级实验室发布的新模型在研究生级数理推理评测中领先两分,观察家们便断言美国拥有不可逾越的护城河;而当数周后某家中国实验室迎头赶上时,铺天盖地的报道又随即宣布全面平价已经到来。

这种对单一跑分排行榜的过度痴迷,造成了极其危险的战略评估偏差。基准跑分仅仅反映了孤立算法模型在静态、人工构造的测试环境中的抽象推理性能。它们完全无法揭示模型在商业真实价值链中的实际扩散深度,无法反映实体企业将自动化系统嵌入日常生产流的协同效率,更无法度量国家工业体系是否正在构筑自我强化的动态数据反馈闭环。一个仅仅在跑分测试集上屡创佳绩、却未能在制造业、能源网络与现代物流体系中实现深度工业集成的经济体,极易对自身的真实工业与科技实力产生致命误判。

为了彻底打破这一战略实证赤字,前沿研究必须从宏观层面的跑分炒作转向微观层面的企业实证观察。唯有系统性测绘两国商业 AI 研发企业正在真实构建什么系统、优先押注哪些核心模态,以及正在赋能哪些支柱经济部门,战略界才能还原出大国科技竞争力的真实全景。

English Original

From Model Parameters to Economic Micro-Foundations

Artificial intelligence does not transform national economies through theoretical laboratory publications. Realized economic transformation requires capital expenditure, enterprise software customization, physical tooling reconfiguration, and organizational workforce adaptation. A nation possessing the most capable conversational chatbot gains little lasting geopolitical advantage if its domestic manufacturing firms lack the technical capability to deploy edge-based vision models on factory floors.

Furthermore, evaluating national AI capacity solely through frontier labs obscures the vast commercial ecosystem operating immediately below the frontier. While frontier model developers capture public attention, downstream developers translate foundational discoveries into tangible software applications, industrial robotics controls, and autonomous navigation stacks. These downstream commercial developers determine the true velocity of economic absorption.

Recognizing this empirical deficit, the RAND Corporation undertook a systematic empirical mapping of the commercial AI developer landscape in the United States and China. Rather than tracking model weights in isolation, this study measures the concrete corporate entities translating artificial intelligence into commercial products across both economies.

中文精译

从模型参数量到经济微观渗透

人工智能对国家经济的根本性重塑,从来不是通过象牙塔实验室中发表的理论学术论文来实现的。任何颠覆性通用技术的真实生产力转化,都离不开庞大的资本性支出、企业级定制软件改造、物理制造装备重新校准以及组织架构与劳动力的深度适配。即便一个国家拥有全球最顶尖的对话式智能助理,但倘若其本土制造企业普遍缺乏在车间流水线上部署边缘机器视觉与具身控制的工程能力,这种算法优势也极难转化为持久的国家地缘经济优势。

更为关键的是,将国家级 AI 评估视角仅仅局限于少数几家前沿巨头实验室,严重遮蔽了支撑现代经济运转的庞大中下游商业研发集群。虽然前沿实验室始终处于聚光灯下,但真正将前沿基础科学发现转化为具体工业软件、自动化生产控制与自主导航系统的,恰恰是这批中下游商业研发实体。正是这批规模庞大的中坚企业,决定了一个国家对颠覆性技术创新的真实吸收速度与产业消化能力。

正是察觉到了战略评估界的严重实证缺失,兰德公司研究团队启动了这项针对中美商业 AI 研发企业全景的系统性实证测绘工程。本项研究不再孤立追踪算法权重的高低,而是深入商业生产实践,全量度量正在两国国民经济各个垂直行业中将 AI 转化为现实商业生产力的真实企业实体。

English Original

Chapter 2: Dataset Construction and Methodology

A Comprehensive Cohort of 1,181 AI Developer Firms

To conduct a rigorous comparative investigation, the authors constructed a bespoke microeconomic dataset of commercial artificial intelligence developer firms in the United States and China. A firm was defined as an active commercial entity that designs, trains, or substantially adapts artificial intelligence models for commercial release, internal enterprise deployment, or industrial automation. Purely non-technical service agencies and hardware component resellers without internal AI modeling capabilities were systematically excluded.

The resulting dataset comprises 1,181 verified commercial AI developer firms, consisting of 743 firms headquartered in the United States and 438 firms headquartered in China. Data collection was finalized in early 2026, capturing corporate registrations, patent filings, active open-source software repositories, venture financing rounds, commercial customer deployments, and regulatory filings from both jurisdictions.

This cohort represents the most extensive audited catalog of active AI developer firms compiled to date. By establishing consistent classification taxonomies across both geographic contexts, the dataset enables direct empirical comparison of technical capabilities, model architectures, commercial market targets, and ecosystem dependencies.

中文精译

第二章:全景微观数据集构建与多智能体研究方法

涵盖中美 1,181 家 AI 研发企业的实证样本库

为了开展严谨的实证对照研究,兰德团队历时数月构建了一座专属于中美两国商业 AI 研发企业的微观数据库。在研究定义中,“AI 研发企业”被严格界定为具备自主设计、训练或深度自适应微调 AI 模型,并将模型应用于对外商业发布、企业内部生产系统或工业自动化装备的实体商业机构。纯粹提供非技术外包服务的中介机构、缺乏自研模型算法能力的分销代理商被严格剔除。

经过多轮跨源交叉去重与清洗,最终入库的经审计有效样本企业共计 1,181 家,其中包括总部位于美国的 743 家企业,以及总部位于中国的 438 家企业。全量数据于 2026 年初最终固化,全面收录了企业工商登记、发明专利申请、代码开源仓库活跃度、风险投资轮次、商业化客户现场交付案例以及两地监管机构的备案信息。

该样本库代表了迄今为止战略界针对中美 AI 实体研发企业最详尽、审计最严格的底层微观全量名录。通过在跨国制度语境下确立完全对等的分类学标准,该数据集首次让两国在技术栈底座、算法模型架构、商业目标市场以及供应链依赖度等维度实现了直接、可复现的微观定量对照。

English Original

Multi-Agent Extraction and Expert Validation Pipeline

Assembling a microeconomic dataset across two distinct corporate reporting environments required an innovative data collection pipeline. Traditional registry scraping often misses fast-moving startup entities or introduces significant naming ambiguities between US parent holdings and domestic Chinese operating subsidiaries. To address these challenges, the research team deployed a multi-agent information retrieval framework powered by specialized LLM agents tasked with scanning technical documentation, code registries, patent databases, and local business registries.

The automated multi-agent extraction pipeline operated in structured stages: candidate entity identification, commercial product verification, model architecture classification, and deployment sector tagging. Each automated extraction was subsequently subjected to dual-analyst manual validation by bilingual research specialists with technical training in machine learning and knowledge of Chinese corporate filings.

The classification taxonomy categorized each firm along multiple dimensions: core model development (foundation model pre-training vs task-specific fine-tuning), supported modalities (text, vision, audio, multimodal, and embodied control), target market (digital enterprise services, industrial manufacturing, consumer applications, public sector), and deployment paradigm (cloud API, edge embedded, on-premises sovereign cluster). This multi-tier taxonomy ensures that nuanced architectural differences are accurately reflected.

中文精译

多智能体交叉核验管线与多维分类学

在两国截然不同的商业监管与公开信息披露环境下构建微观企业库,对数据挖掘的方法论提出了极高要求。传统的企业名录爬虫极易遗漏处于爆发期的高精尖初创企业,且在美国控股母实体与中国境内实体之间存在大量的名称歧义与关联关系困扰。为此,研究团队自主研发了一套由专业 LLM 智能体驱动的多智能体自动化信息抽取管线,自动化对技术技术白皮书、代码托管平台、专利公告库及地方企业登记系统进行全域挖掘。

多智能体管线由四大专业子智能体分工协同:候选企业自动发现、自研商业产品核验、模型算法架构定级,以及落地垂直行业标签判定。为了确保学术级零容错,所有自动化抽取的元数据均由具有机器学习技术背景、熟稔中国企业商业文书的双语法政研究员执行双盲独立人工交叉核验。

在企业分类学维度上,该样本库从多个技术切面进行了系统解构:核心研发层级(通用底座预训练 vs 垂直领域精调)、支持数据模态(文本、计算机视觉、语音、跨模态融合以及具身运动控制)、目标行业市场(纯数字企业服务、智能制造、消费端应用、政企物理基建),以及系统部署形态(云端 API、终端边缘嵌入、私有化物理集群)。这套多层级分类框架确保了深层次的技术架构差异能够被忠实刻画。

English Original
Figure 1: Architectural and Modality Distribution
Figure 1: Architectural and Modality Distribution of AI Developer Firms in the United States and China (N = 1,181). Source: RAND Corporation Research Report RR-A5043-1.
中文精译
Figure 1: Architectural and Modality Distribution
图 1:中美 AI 研发企业底层算法架构与核心模态分布对比(样本量 N = 1,181)。数据来源:兰德公司研究报告 RR-A5043-1。
图表架构与核心概念对照解析
Transformer 架构
基于自注意力机制的深度学习基础架构,已成为全球大语言模型与多模态模型的事实工业标准。
基础模型(Foundation Models)
具备海量参数与通用跨领域泛化能力的核心底座模型,研发成本高昂,占两国企业总数约两成。
领域适配模型(Domain-Adapted Models)
针对医疗、制造、金融等垂直领域进行深度微调与专用任务特化的模型,占两国研发企业绝大多数。
具身运动控制(Embodied & Sensorimotor Control)
融合多维传感器输入并直接输出物理作动控制指令的多模态架构,在中国企业中占比显著偏高。
English Original

Chapter 3: The Critical Divergence: Pure Software vs. Physical Embodiment

The United States: Dominance in Enterprise Software and Digital Work

While the architectural foundations of AI development in the United States and China exhibit remarkable convergence, the commercial destinations of those models diverge dramatically. In the United States, commercial AI development is overwhelmingly channeled toward digital knowledge work and enterprise software. 61.2 percent of US AI developer firms build pure-play digital software products, operating entirely within cloud and desktop environments.

The concentration of US commercial activity in enterprise SaaS, developer tooling, healthcare informatics, and defensive cybersecurity reflects the structural economics of the American tech sector. Decades of venture capital investment have established software-as-a-service as the benchmark for financial returns, characterized by gross margins exceeding 70 percent, rapid software deployment cycles, and minimal hardware depreciation risk. American enterprise buyers possess both the discretionary IT budgets and the digital infrastructure required to absorb high-margin cognitive software tools.

Consequently, American foundation models and vertical adaptations are predominantly optimized for text comprehension, coding syntax, analytical reasoning, and knowledge retrieval. While American laboratories have demonstrated breakthrough physical robotics systems in research settings, the vast majority of commercial capital flows into digital applications that substitute or augment office-bound knowledge workers.

中文精译

第三章:商业应用路径的关键分水岭:纯软件与物理具身智能

美国生态:聚焦企业级软件与纯数字认知服务

尽管中美两国在底层算法架构上呈现出高度一致的收敛性,但这两套算法在商业落地的主战场上却走向了截然相反的两极。在美国,商业 AI 研发力量几乎压倒性地涌向了纯数字办公与企业软件领域。统计数据显示,高达 61.2% 的美国 AI 研发企业专注于纯软件级数字化产品,所有业务流完全运行于云端集群与桌面终端环境之内。

美国商业研发力量在企业级 SaaS、开发者代码工具、医疗信息分析以及网络安全防御等领域的重兵集结,深刻映射了美国科技创新的资本收益逻辑。经过数十年的成熟演进,SaaS 模式已成为硅谷风险投资机构衡量商业回报的黄金标杆:70% 以上的高额毛利率、敏捷高效的纯云端灰度迭代,以及近乎于零的物理硬件折旧损耗。加之美国本土企业客户拥有全球最充裕的信息化 IT 预算与高度数字化底座,使这类高溢价的认知型软件得以迅速实现商业变现。

其必然结果是,美国的基础大模型及其衍生微调技术,绝大多数都聚焦于自然语言理解、代码逻辑推理、跨文档信息检索与结构化分析。尽管少数顶级美国研究机构在实验室中展示出了惊艳的人形机器人动作演示,但全美商业研发资本的核心主力,依然牢牢锚定在替代或增强白领办公室知识工作者的纯数字赛道上。

English Original

China: Rapid Expansion in Industrial Edge and Physical Systems

In contrast to the software-centric focus of the American market, Chinese AI developer firms are heavily focused on physical embodiment and industrial infrastructure. Approximately 73.8 percent of Chinese AI developer firms in the dataset build models explicitly designed to interface with physical hardware, industrial manufacturing equipment, robotic actuators, autonomous transportation fleets, or smart electrical grids.

This structural orientation is anchored in China's position as the world's largest manufacturing powerhouse. Industrial enterprises face rising domestic labor costs, demographic transitions, and stringent quality control mandates across high-volume assembly lines. AI developer firms in China have found their most lucrative and predictable enterprise revenue not in generic productivity software, but in factory edge vision for defect detection, predictive maintenance algorithms for heavy machinery, automated warehouse logistics, and autonomous industrial haulage.

Furthermore, national industrial policy incentives have deliberately steered computational talent and capital toward physical manufacturing upgrades. Chinese government procurement programs, national manufacturing transformation funds, and specialized industrial testbeds prioritize AI deployments that enhance industrial output and technological self-reliance over consumer chatbots or digital marketing automation.

中文精译

中国生态:深耕智能制造、机器人与物理基础设施

与美国市场以纯软件为重心的局面形成鲜明对比,中国 AI 研发企业呈现出极为鲜明的物理具身与实体工业导向。在样本数据库中,高达 73.8% 的中国商业 AI 研发企业,其自研或适配模型明确绑定了物理实体硬件、工业制造流水线、多自由度机器人机构、自动驾驶运输车队或国家智慧电网调度系统。

这一产业结构选择深植于中国作为全球制造业超级中枢的实体底座之中。随着劳动力成本攀升、人口结构变迁以及高端精密制造对品控良率的极致苛求,中国 AI 研发企业找到了最确定、现金流最充沛的商业切入口:并非通用型办公生产力助理,而是深入车间一线的工业质检机器视觉、重型装备的预测性运维算法、密集仓储中的全自主物流调度,以及矿山与港口的无人化重载运输。

同时,国家级产业扶持政策也在系统性引导顶尖算法人才与研发资本向实体制造纵深转移。各级政府的专项采购计划、新型工业化专项技改基金以及遍布长三角与珠三角的工业实证试验场,均将能够直接提升工业附加值与实体产业链安全可控的具身系统置于首要支持序列,极大抑制了研发力量在低壁垒消费级文娱应用中的内卷消耗。

English Original
Figure 2: Commercial Deployment Distribution
Figure 2: Commercial Deployment Distribution: Pure Digital Software vs. Physical Embodiment (US vs. China AI Developer Firms). Source: RAND Corporation Research Report RR-A5043-1.
中文精译
Figure 2: Commercial Deployment Distribution
图 2:中美 AI 研发企业商业化落地场景分布:纯数字软件 vs 物理具身系统。数据来源:兰德公司研究报告 RR-A5043-1。
图表架构与核心概念对照解析
纯数字软件(Pure Digital Software)
运行于云端或本地终端、无物理交互作动器的软件服务,涵盖企业 SaaS、开发者辅助工具、网络安全与知识管理。
物理具身系统(Physical Embodied AI)
与物理实体硬件深度绑定的智能系统,包含工业机器人、柔性制造产线、自动驾驶车队与智能电网调度节点。
工业边缘端部署(Industrial Edge Deployment)
直接部署于生产车间与硬件装备的边缘算力,对实时响应、确定性延迟与抗网络中断能力有极高要求。
多模态融合控制(Multimodal Sensorimotor Control)
实时融合视觉、力觉、触觉及位姿传感器数据,驱动机械结构执行精密作业的闭环控制范式。
English Original

Chapter 4: The Strategic Moat of Embodied AI

The Physical Data Flywheel and Synthetic Data Limits

The divergence between purely digital software and physical embodied systems carries profound long-term strategic ramifications. In the digital domain, foundational models are rapidly exhausting high-quality human-generated text and code on the public internet. While synthetic data generation provides meaningful performance extensions, frontier labs acknowledge that synthetic self-training exhibits diminishing returns and error compounding over successive generations.

In contrast, embodied artificial intelligence operates in the infinite, non-replicable complexity of physical reality. Robots, autonomous vehicles, and industrial machinery interact directly with dynamic real-world environments, continuously capturing multimodal sensorimotor streams: high-resolution spatial video, LIDAR point clouds, tactile force feedback, torque telemetry, and thermal variations. This physical interaction data cannot be scraped from the open web or synthesized through digital simulation alone.

Firms operating physical systems establish proprietary data flywheels that strengthen with every operational hour. An automated assembly line operating 24 hours a day generates rare corner-case failure data that cannot be simulated artificially. Over time, this empirical interaction moat becomes virtually insurmountable for competitors lacking physical hardware deployed in the field.

中文精译

第四章:物理具身智能的长期战略护城河

物理交互数据飞轮与合成数据边界

纯数字软件与物理具身系统之间的应用分流,绝非简单的商业偏好差异,而是孕育着深远的大国长期战略壁垒。在纯数字虚拟世界中,大语言模型对公开互联网上人类沉淀的高质量图文与代码数据正在加速消耗殆尽。尽管基于大模型自身生成的合成数据能够在一定程度上延缓数据瓶颈,但全球前沿实验室均已清醒认识到,多代连续自监督合成训练存在严重的收益递减与误差自放大隐患。

与此形成鲜明对比的是,具身人工智能扎根于真实物理世界无尽的非线性复杂度之中。工业机器人、全自主运载工具与自动化流水线直接与现实物理规律实时交互,源源不断地捕捉海量多源高维物理传感数据:微秒级时空视觉视频流、毫米波雷达与激光点云、机械臂指尖触觉与六维力传感器反馈、关节电机扭矩波动以及温度梯度场。这类微观物理交互数据,既不可能从公开互联网上无本爬取,也绝非纯数字仿真引擎所能完全替代。

每一个真正在物理世界中常态化运行的智能实体,都在构筑一道随作业时长指数级自我强化的数据飞轮。一个在严苛工业环境下连续不间断运转数千小时的智能产线,所沉淀的极端偶发缺陷与边缘案例数据,是纯软件公司在实验室中穷尽算力也无法臆构的。长此以往,这种来自物理世界深处的动态实测数据壁垒,对于缺乏大规模现场硬件部署的外部竞争者而言,将构成近乎无法逾越的物理护城河。

English Original

Hardware Supply Chain Proximity and Iteration Velocity

Developing embodied AI systems requires tight co-design between algorithmic models and mechanical hardware. Model architectures must be tailored to specific actuator response rates, sensor latency profiles, thermal envelopes, and edge computing chipsets. When algorithmic updates occur, physical mechanical components often require rapid prototyping, machining, and field recalibration.

Chinese embodied AI firms benefit from geographic co-location within the world's densest advanced hardware supply chains. In regional manufacturing hubs such as the Greater Bay Area and Yangtze River Delta, a robotics team can design a custom actuator or sensor bracket in the morning, receive precision-machined prototypes by afternoon, and test real-time inference on the production line by evening. This rapid mechanical feedback cycle accelerates empirical progress.

In the United States, hardware prototyping cycles frequently encounter protracted supply chain lead times, with domestic manufacturing capacity hollowed out and overseas procurement introducing weeks of shipping latency. As a result, American researchers often spend significant effort optimizing simulation environments to bypass physical testing, while their Chinese counterparts iterate directly against real physical prototypes on active factory floors.

中文精译

供应链腹地密度与软硬件协同迭代

具身智能系统的工程研发,高度依赖算法模型与机械硬件之间的紧密协同共构。模型控制策略必须针对特定执行器的动态响应频宽、传感器的通信总线延迟、系统的散热功耗边界以及边缘端低功耗推理芯片进行深度定制。每一次算法策略的重大跃迁,往往都需要对机械关节结构、减速器刚度及传感器支架进行快速打样与现场二次标定。

中国具身智能研发企业坐拥全球产业密度最高、协同最敏捷的高端硬件制造腹地。在长三角与粤港澳大湾区等高端制造核心集群,一支机器人算法团队可以在清晨完成定制电机关节或力传感器治具的 CAD 建模,午后便能拿到高精度 CNC 试制样件,入夜前即可在车间实验台上完成实机联调与边缘闭环推理。这种以天为单位的软硬件高速物理验证闭环,赋予了企业极强的工程推进韧性。

相比之下,美国本土的硬件工程打样往往受制于实体供应链的严重断层与长周期外包依赖,关键机械构件从下单到空运清关常需数周甚至数月周折。其直接后果是,大量优秀的美国学者与工程师不得不耗费巨量心血在纯虚拟仿真环境中搭建复杂的数字孪生系统,以迂回替代物理实测;而中国的工程团队则早已在真实的物理生产线与原型样机上完成了成百上千轮真实物理世界的实机碰撞与暴力测试。

English Original

Operational Resilience Against Remote Disruption

The operational resilience of physical embodied systems represents an underappreciated geopolitical dimension. Purely digital software services rely heavily on global cloud infrastructure, continuous public internet connectivity, and foreign API service agreements. A domestic enterprise relying on a foreign SaaS platform for mission-critical workflows is inherently vulnerable to remote account termination, cross-border data transfer embargoes, or submarine cable interdictions.

Conversely, industrial embodied AI systems are engineered to operate autonomously on local physical infrastructure. Factory quality control models, robotics navigation kernels, and smart electrical grid dispatchers are deployed directly onto on-premises edge appliances and isolated industrial control networks. These systems function uninterrupted even in the event of complete severance from external cloud networks.

By embedding artificial intelligence directly into the physical machinery of core industries, China is constructing an automated industrial base with high operational durability. Even under severe geopolitical friction or export isolation, these embodied systems continue driving factory production, logistics throughput, and energy distribution without relying on foreign cloud APIs.

中文精译

抵御远程断链的物理韧性与地缘稳健性

物理具身智能系统所蕴含的生产连续性与抗毁韧性,长期以来被主流战略评估界严重低估。纯云端数字软件服务高度依赖全球跨境公共互联网连接、跨洋高带宽海底光缆,以及海外服务商的商业 API 持续授权协议。倘若一家核心实体企业的日常核心业务流深度绑定于境外的闭源 SaaS 平台,一旦遭遇地缘维度的远程账号封禁、跨境数据阻断或物理断网,其数字化运营系统将面临瞬间瘫痪的毁灭性打击。

与此完全不同,深耕工业一线的物理具身 AI 系统,从工程设计之初就被赋予了在局域物理空间内离线自持的苛刻要求。无论是工厂流水线上的精密质检模型、立体高位仓储中的自主叉车导航算核,还是特高压变电站的故障研判系统,全部以固化固件或边缘微集群的形式直装于现场工控机之中。即便外部公网连接彻底切断,这类物理装备依然能够依靠本体算力与本地传感网络,长周期、不间断地自主运行。

通过将人工智能算法物理熔铸于支柱产业的重型生产装备与工业基础设施之内,中国正在打造一座具备超强抗风险自持能力的自主化工业基座。哪怕置身于极端的外部地缘封锁或出口限制环境下,这些深植于厂房、码头与电网中的具身系统,依然能够平稳保障国家工业生产、基础物流运转与能源电力的安全调度,完全不依赖任何外部云端黑盒 API 的施舍。

English Original

Chapter 5: Open-Weight Diffusion and Global Token Parity

OpenRouter Telemetry: The Global Rise of Open Weights

The international diffusion of artificial intelligence capabilities cannot be evaluated solely through domestic commercial markets. Globally, developers choose models based on three decisive factors: performance, control, and inference cost. Over the past two years, open-weight models developed by Chinese institutions (notably DeepSeek, Alibaba Qwen, 01.AI, and Zhipu AI) have experienced a dramatic surge in international adoption.

To quantify this phenomenon, the research team analyzed longitudinal routing telemetry from OpenRouter, one of the world's primary vendor-agnostic LLM routing platforms serving millions of global software engineers. In early 2024, proprietary US closed models (led by OpenAI, Anthropic, and Google) accounted for over 80 percent of daily routed token volume. By the first quarter of 2026, empirical telemetry reveals that Chinese open-weight models reached token volume parity with leading US closed models, each commanding approximately 50 percent of global routing traffic.

This shift represents a fundamental transformation in global developer mindshare. Software teams in Europe, Latin America, Southeast Asia, and Africa are increasingly adopting open-weight models for core production infrastructure. The ability to inspect model weights, fine-tune locally without vendor lock-in, and deploy on sovereign hardware clusters has made open-weight architectures the preferred substrate for international enterprise development.

中文精译

第五章:开源权重扩散与全球路由 Token 平价

OpenRouter 实证遥测:开源权重模型的全球崛起

人工智能前沿技术在全球范围内的真实辐射力与生态渗透度,绝无法仅凭单一国内商业市场的体量来衡量。对于全球数以千万计的独立开发者与技术团队而言,决定其技术选型取向的唯有三大硬指标:模型推理表现、自主可控权,以及推理边际成本。过去两年间,以 DeepSeek、阿里云通义千问(Qwen)、智谱 GLM、零一万物(Yi)为代表的中国开源权重模型,在国际开发者社区掀起了一场席卷全球的技术替代风暴。

为了精准量化这一颠覆性扩散态势,兰德团队深入调取并分析了全球知名中立大模型聚合路由平台 OpenRouter 的多年纵向遥测数据集。该平台服务于全球数百万一线软件工程师与企业架构师。实证数据显示,在 2024 年初,由 OpenAI、Anthropic、Google 等主导的美国商业闭源旗舰模型,独占了该平台超过 80% 的每日有效 API 路由 Token 调用量;然而到了 2026 年第一季度,中国开源权重模型的日均 Token 路由总量已正式跃升至 49.2%,与美国商业闭源阵营(50.8%)在前沿实战流量上平分秋色,达成了历史性的全球 Token 平价。

这一数据的巨变,标志着全球开发者心智格局的根本性重塑。无论是在欧洲大陆、拉美新兴市场、东南亚数字化前沿,还是非洲科技创新枢纽,越来越多的技术团队毅然将生产环境的核心工作流迁移至中国开源底座之上。拥有完全透明的模型参数权重、零供应商锁定风险的本地化二次调优自由,以及在主权可控硬件上部署的确定性,让开源权重架构成为全球新兴产业开发者的首选技术基石。

English Original

Inference Economics: The 1/4 to 1/5 Cost Disruption

The primary catalyst behind this global migration is inference economics. As frontier artificial intelligence transitions from experimental prototypes to high-volume commercial automation, the operational expenditure of running billions of inference tokens daily has become the dominant cost center for modern digital enterprises. Software companies cannot build scalable businesses if every customer query consumes tens of cents in proprietary API fees.

Telemetry data demonstrates that Chinese open-weight models deliver reasoning capabilities comparable to tier-one US proprietary models at roughly 1/4 to 1/5 the inference cost. Hosted commercial API endpoints for frontier open-weight models routinely charge between $0.14 and $0.60 per million output tokens, compared to $3.00 to $15.00 per million tokens for premier American closed alternatives. When enterprises self-host open weights on commodity compute clusters, marginal inference costs fall even further.

This severe price disparity has initiated a commoditization wave across the global software industry. High-margin US software businesses that previously built simple wrappers around proprietary frontier APIs are experiencing margin collapse, while open-weight adopters build cost-resilient enterprise applications that can scale to billions of monthly active users without catastrophic cloud compute bills.

中文精译

推理经济学:降本四分之三带来的生态重构

驱动这场全球技术大迁移的最核心底层引擎,正是冷酷而现实的推理经济学法则。当人工智能彻底走出早期技术尝鲜的象牙塔,全面切入日均调用量以百亿、千亿计的规模化商业生产系统,每天消耗海量 Token 所产生的持续云算力支出,已毫无悬念地跃升为现代数字化企业最沉重的成本负担。倘若每一次用户交互都要向闭源巨头进贡数美分甚至数十分钱的昂贵 API 调用税,任何高频普惠的商业模式都将难以为继。

实证遥测价格数据给出了极其震撼的对比:中国开源权重模型在达到并匹敌美国第一梯队闭源旗舰推理性能的同时,其对外商业化 API 托管报价普遍仅为美国对标闭源模型的四分之一至五分之一。当前主流平台对顶尖开源推理模型的调用定价常年稳定在每百万 Token 产出 0.14 至 0.60 美元区间,而美国同等段位闭源旗舰 API 的标价依然高居 3.00 至 15.00 美元之间。若企业选择在自有机房集群上私有化托管,其边际推理成本甚至还能再下一个数量级。

这种断崖式的单价代差,在全球软件研发界引爆了一场不可逆的去溢价风暴。大量过往仅仅依赖海外闭源 API 进行简单套壳打包的轻量级 SaaS 公司正在遭遇前所未有的毛利崩盘;而率先拥抱低成本开源底座的开发者团队,则得以在完全可承受的算力预算内,向数以亿计的全球终端用户交付过去被视为奢侈品的智能推理体验,彻底改写了全球企业软件的价值分配版图。

English Original
Figure 3: Global Routing Telemetry and Inference Economics
Figure 3: Global Routing Telemetry: Token Volume Share Trajectory and Inference Cost Comparison (OpenRouter Data 2024-2026). Source: RAND Corporation Research Report RR-A5043-1.
中文精译
Figure 3: Global Routing Telemetry and Inference Economics
图 3:全球主流路由实证遥测:Token 调用份额走势与推理成本阶梯对照(数据来源:OpenRouter 遥测数据库与兰德研究报告 RR-A5043-1)。
图表架构与核心概念对照解析
OpenRouter 路由平台
全球主流的多模型统一 API 聚合服务,实时追踪数百万全球开发者的模型调用流向与真实 Token 消耗。
开源权重模型(Open-Weight Models)
公开提供全部或核心模型权重参数,允许开发者自主部署、二次微调与私有化运行的模型生态。
算力代币平价(Token Parity)
指中国开源权重模型在全球 API 路由网络中的日均 Token 生成总量首次比肩美国商业闭源旗舰模型。
推理经济学(Inference Economics)
以单位计算资源产出 Token 的商业边际成本为核心的分析框架,极低推理单价打破了昂贵的软件订阅壁垒。
English Original

Chapter 6: Strategic and Policy Implications

Policy Recommendations for the United States

The findings of this empirical study indicate that US artificial intelligence strategy requires a fundamental recalibration. American policymakers must recognize that algorithmic leadership on digital benchmark leaderboards does not translate automatically into national economic resilience or industrial superiority. Relying exclusively on high-margin enterprise SaaS leaves critical industrial sectors exposed to international technological leapfrogging.

To preserve long-term technological competitiveness, the United States should actively incentivize the commercialization of embodied artificial intelligence and robotics. Policy mechanisms should include targeted federal investment in physical testing testbeds, tax credits for domestic manufacturers deploying edge-based automation systems, and streamlined regulatory pathways for autonomous industrial operations. Rebuilding domestic advanced manufacturing ecosystems is not merely an industrial policy imperative, but a foundational requirement for sustaining national AI leadership.

Furthermore, American defense and civilian planners must diversify their technical evaluations beyond digital language models. National AI readiness scorecards must incorporate microeconomic metrics tracking the installed base of robotic systems, physical sensorimotor data capture volumes, and industrial automation density across critical supply chains.

中文精译

第六章:国家战略与产业治理政策启示

面向美国决策层的产业政策与战略纠偏建议

本项微观实证研究所揭示的一系列核心数据表明,美国的国家级人工智能战略亟待进行一次大刀阔斧的底层纠偏。华盛顿的政策制定者必须清醒认识到:学术跑分榜单上的算法领先,绝不可能自发蜕变为国家实体经济的防御韧性与制造业核心竞争力。如果一国的创新资源过度依附于轻量级、高毛利的企业 SaaS 赛道,而对重资产的物理实体工业视而不见,其关键产业基座终将暴露在被他国物理具身创新全面弯道超车的巨大风险之中。

为捍卫长远科技竞争力,美国应当自上而下强力扭转对纯软件的路径依赖,大力扶持物理具身智能与智能机器人系统的产业化落地。具体的实操治理工具应当涵盖:在联邦层面重金设立国家级物理实机中试验证基地、向本土部署工业边缘智能装备的制造企业提供更激进的税收抵扣与技改补贴,以及为高危自主化重型装备开辟合规绿色通道。重振本土先进制造业不仅是一项传统的产业回流诉求,更是维系未来国家 AI 核心话语权的生命线。

同时,美国国家安全与民事规划机构在度量国家 AI 实力时,必须彻底摆脱唯大语言模型跑分论的狭隘视角。国家级 AI 战备能力评估指标库应当全面纳入工业微观实证维度:包括关键行业重型机器人的实际在役安装量、工业现场物理传感数据的日均自主吞吐规模,以及支柱产业链核心环节的边缘自动化覆盖密度。

English Original

Strategic Opportunities and Bottlenecks for China

For China, the empirical data validates the strategic efficacy of national policies aimed at industrial integration and open-source diffusion. By focusing development efforts on smart manufacturing, robotics, and physical infrastructure, Chinese firms have carved out defensible technological positions anchored in the world's most versatile production supply chain. Furthermore, the global success of open-weight models has established significant international goodwill and developer adoption.

However, significant structural vulnerabilities persist within the Chinese AI ecosystem. Chief among these is continued exposure to advanced semiconductor export restrictions. While architectural optimizations and efficient fine-tuning techniques have allowed Chinese developers to deliver competitive frontier models with fewer resources, sustaining foundation model pre-training across increasingly large multimodal datasets will inevitably encounter hardware bottlenecks if domestic chip alternatives fail to match global leading-edge yields.

Additionally, Chinese developers must cultivate greater domestic enterprise willingness to pay for software innovation. While industrial enterprises readily invest in physical robotic hardware and machinery, willingness to fund standalone software licenses and high-tier cloud computing architectures remains constrained compared to Western markets. Fostering a healthy, self-sustaining domestic software market will be essential for funding next-generation basic research.

中文精译

面向中国产业生态的战略机遇与核心瓶颈

对于中国而言,微观实证数据充分印证了过去数年坚持脱虚向实、深耕新型工业化与拥抱全球开源生态的战略远见。通过将全社会最宝贵的算法与工程智慧精准导流至智能制造产线、通用机器人本体及智能基础设施等硬核实体中,中国科技企业依托全球最完备的制造业产业母胎,构筑起了极具自持力与扩展性的物理壁垒。而开源权重战略在海外开发者生态掀起的技术风暴,更是打破了技术霸权封锁,赢得了全球开发者的广泛拥护与深度绑定。

然而,客观审视中国 AI 研发全景,其底层依然潜伏着不容忽视的结构性短板。首当其冲的当属先进制程算力芯片的断供隐忧。尽管中国工程师凭借精妙绝伦的模型架构创新(如专家混合架构 MoE 与极致算子量化剪枝),以极高的算力效能比在通用基座上实现了四两拨千斤,但伴随着跨模态具身数据规模的指数级膨胀,如果本土半导体制造产线在先进制程良率上无法如期突破,未来超大规模多模态底座的基础预训练终将面临硬性物理撞墙的严峻考验。

此外,国内商业市场对纯软件研发价值的付费认同度仍显薄弱。中国实体企业往往习惯于为看得见摸得着的物理机械硬件大笔买单,而对于高价值工业基础软件、自研算法内核及云算力架构的长期订阅付费意愿依然与欧美成熟市场存在显著差距。如何健全全产业链对算法知识产权的价值重估,构建良性造血的工业软件市场生态,是中国能否长久反哺前沿基础算法探索的另一道核心考题。

English Original

Rethinking Export Controls and International AI Governance

The empirical convergence in core algorithmic architectures and the proliferation of open-weight models reveal significant limitations in existing export control paradigms. Multilateral export restrictions designed around rigid computational FLOP thresholds struggle to restrict technological diffusion in an era where algorithmic efficiency gains, model distillation, and open-source parameter sharing allow developers to achieve frontier performance on modest hardware footprints.

Policymakers must transition from blunt hardware interdictions toward sophisticated whole-of-ecosystem monitoring. Restricting raw silicon shipments alone cannot halt the rapid iteration of embodied systems when physical training data and manufacturing feedback loops reside within domestic borders. Future non-proliferation and economic security frameworks must incorporate software optimization trajectories and industrial edge deployment realities.

Finally, the rapid militarization and industrial integration of physical artificial intelligence underscore the urgent necessity for bilateral US-China technical engagement on safety and catastrophic risk mitigation. As autonomous drones, industrial machinery, and critical infrastructure become increasingly reliant on machine learning kernels, technical protocols governing fail-safe actuation, unintended physical escalation, and sensor spoofing require clear channels of scientific communication between the two leading AI superpowers.

中文精译

出口管制边界的反思与双边安全治理合作

底层模型架构的实证高度收敛以及开源权重模型的全球野蛮生长,无情暴露出既有单边出口管制框架的逻辑局限。现行以纯粹硬件算力浮点峰值(FLOPs)为核心的断供防线,在算法极限压榨、模型知识蒸馏以及全网开源共享的汹涌浪潮下,其技术封锁效果正被快速稀释。全球技术实践已清晰证明:高超的工程优化完全可以在中等规模的计算硬件底座上释放出比肩甚至超越传统重型集群的前沿推理动能。

全球产业监管框架必须摒弃简单的硬件一刀切思维,转向对全景研发生态的全维动态审视。单纯封堵高端裸芯片的跨境流通,根本无法遏制物理具身系统在本土制造腹地中的高速进化,因为驱动系统迭代的最核心燃料:工业现场实测数据与产线反馈,完全深植于本土疆域之内。未来的科技安全治理框架必须将算法优化红利与工业边缘端扩散现实全面纳入动态评估范畴。

最后,物理具身人工智能在实体工业乃至潜在防务领域的深度融合,让中美两国在 AI 极端安全边界与防灾冗余机制上的双边技术对话变得前所未有的紧迫与必要。随着高自主性无人集群、智能工业重机以及核心能源枢纽日益交由复杂神经网络控制,建立涵盖物理执行器硬熔断、失控连锁升级防范以及传感器欺骗抵御的全球通用安全技术准则,不仅符合大国共同利益,更是维系未来智能社会整体安全底线的必由之路。

English Original

Conclusion: Moving Beyond the Frontier Benchmark Illusion

The empirical evidence gathered across 1,181 AI developer firms provides a necessary corrective to contemporary geopolitical discourse. The competition in artificial intelligence between the United States and China cannot be reduced to a binary sprint toward an elusive benchmark score. Instead, the global landscape is witnessing the emergence of two structurally differentiated ecosystems: an American paradigm optimizing cognitive software services for digital enterprises, and a Chinese paradigm embedding autonomous perception and control into the physical machinery of global manufacturing.

Neither model possesses an absolute, uncontested monopoly on future technological progress. The nation that successfully bridges its structural deficiency: whether by connecting digital cognitive software to physical manufacturing, or by scaling domestic high-performance silicon to sustain open-weight foundational research: will establish the standard for 21st-century technological leadership. For strategic analysts and corporate executives alike, navigating this landscape requires looking past the benchmark illusion to measure real-world commercial and industrial deployment.

中文精译

结语:打破跑分幻觉,正视双轨并行的产业全貌

涵盖中美 1,181 家商业 AI 研发企业的微观实证调研,为当前浮躁而片面的大国科技地缘叙事注入了一剂沉静的清醒剂。中美人工智能竞争绝非一场争夺单一静态基准跑分王冠的单维度百米赛跑;相反,我们正在见证全球两个体量最庞大、结构最分化的产业生态在不同维度上加速分道扬镳:美国模式致力于为全球企业和白领知识工作者打造极致的数字认知服务,而中国模式则坚决将自动化感知与决策系统深扎于全球制造业最核心的工业机器之中。

这两种发展范式,谁也无法在未来的技术演进长河中享有绝对的单向压制权。未来的终极胜负,取决于谁能率先补齐自身的结构性短板:是美国能否将其无与伦比的纯数字软件与通用算法底蕴成功嫁接至本土实体工业与先进制造之中,还是中国能否彻底攻克本土高端算力瓶颈以持久支撑其开源模型与具身底座的高强度自主探索。对于全球战略决策者与产业领袖而言,打破单一维度的跑分迷思,沉下心来审视两国产线与机房里正在真实生长的微观商业实践,才是看清未来大国竞争全貌的唯一正道。

English Original

Report Citation and Research Metadata

Recommended Citation: Schmid, Jon, Prateek Puri, Vitor Melo, and Anthony Hakim, AI Development in the United States and China: Evidence from a New Dataset of AI Developer Firms, Santa Monica, Calif.: RAND Corporation, RR-A5043-1, August 2026. Available at: https://www.rand.org/pubs/research_reports/RRA5043-1.html.

About the RAND Corporation: The RAND Corporation is a nonprofit institution that helps improve policy and decisionmaking through research and analysis. RAND publications do not necessarily reflect the opinions of its research clients and sponsors.

Original Research Source: Research Report RR-A5043-1. For permissions and full institutional copies, contact RAND Corporation at rand.org.

中文精译

报告正式引用格式与文献信息

推荐引用格式:Schmid, Jon, Prateek Puri, Vitor Melo, and Anthony Hakim, AI Development in the United States and China: Evidence from a New Dataset of AI Developer Firms, Santa Monica, Calif.: RAND Corporation, RR-A5043-1, August 2026. 报告官方地址:https://www.rand.org/pubs/research_reports/RRA5043-1.html.

关于兰德公司(RAND Corporation):兰德公司是一家致力于通过客观分析与独立研究改善全球公共政策与商业决策的非营利顶尖智库。本报告所有研究结论与微观企业分类学知识产权归属于兰德公司。

报告源头与引用声明:兰德公司研究报告 RR-A5043-1。如需查阅英文完整原版白皮书与企业微观元数据附录,请访问兰德官方知识库:rand.org。

链接已复制到剪贴板!