<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>China Open Models on Korea Invest Insights</title><link>https://koreainvestinsights.com/tags/china-open-models/</link><description>Recent content in China Open Models on Korea Invest Insights</description><generator>Hugo -- gohugo.io</generator><language>en</language><copyright>koreainvestinsights.com · @korea_invest_insights</copyright><lastBuildDate>Fri, 17 Jul 2026 19:16:08 +0900</lastBuildDate><atom:link href="https://koreainvestinsights.com/tags/china-open-models/feed.xml" rel="self" type="application/rss+xml"/><item><title>What China's Open-Model Convergence Actually Changes: Value-Chain Redistribution, Not Demand Collapse</title><link>https://koreainvestinsights.com/post/china-open-model-convergence-value-chain-redistribution-2026-07-17/</link><pubDate>Fri, 17 Jul 2026 22:00:00 +0900</pubDate><guid>https://koreainvestinsights.com/post/china-open-model-convergence-value-chain-redistribution-2026-07-17/</guid><description>
 &lt;blockquote&gt;
 &lt;p&gt;Context
This piece is a follow-up to &lt;a class="link" href="https://koreainvestinsights.com/post/kimi-k3-linear-api-pricing-semiconductor-big-tech-impact-2026-07-17/" &gt;Kimi K3 Resets the AI Price Curve&lt;/a&gt;. Where that piece verified the pricing and architecture of &lt;strong&gt;a single model&lt;/strong&gt;, this one expands the lens to how &lt;strong&gt;the entire Chinese open-model ecosystem&lt;/strong&gt; redistributes the semiconductor and Big Tech value chain, and how US policy and US-China tension reshape that reading. It pairs well with &lt;a class="link" href="https://koreainvestinsights.com/post/semiconductor-bull-bear-four-clocks-capital-intensity-cycle-2026-07-17/" &gt;The Real Debate in Semiconductors&lt;/a&gt;, &lt;a class="link" href="https://koreainvestinsights.com/post/memory-fair-value-fcfe-terminal-samsung-hynix-micron-2026-07-17/" &gt;Are Semiconductors Cyclical, and What Is Fair Value?&lt;/a&gt;, and &lt;a class="link" href="https://koreainvestinsights.com/post/cxmt-ipo-memory-price-risk-hbm-client-dram-2026-06-21/" &gt;CXMT IPO And Memory Price Risk&lt;/a&gt;. Related hubs are the &lt;a class="link" href="https://koreainvestinsights.com/page/korea-semiconductor-hbm-kospi-hub/" &gt;AI HBM Hub&lt;/a&gt; and the &lt;a class="link" href="https://koreainvestinsights.com/page/exclusive-analysis-hub/" &gt;Exclusive Analysis Hub&lt;/a&gt;.&lt;/p&gt;

 &lt;/blockquote&gt;
&lt;h2 id="tldr"&gt;TL;DR
&lt;/h2&gt;&lt;ul&gt;
&lt;li&gt;Chinese open models converging in performance looks more like &lt;strong&gt;value-chain redistribution than an AI demand collapse&lt;/strong&gt;. The monopoly value of model APIs and leading-edge GPUs falls, while cheaper inference lifts usage and can raise the value of memory, storage, networking, power, and cloud distribution.&lt;/li&gt;
&lt;li&gt;Relative preference splits this way. Within Korean memory, &lt;strong&gt;Samsung Electronics &amp;gt; SK Hynix&lt;/strong&gt;; within US semiconductors, &lt;strong&gt;Micron and SanDisk &amp;gt; NVIDIA&lt;/strong&gt;; within Big Tech, &lt;strong&gt;Meta and Amazon &amp;gt; Google and Microsoft &amp;gt; pure model vendors&lt;/strong&gt;. [Inference: relative judgment]&lt;/li&gt;
&lt;li&gt;Chinese open models have proven &lt;strong&gt;falling compute and memory cost per unit of intelligence&lt;/strong&gt;. They have not proven &lt;strong&gt;a decline in total silicon spend&lt;/strong&gt;. TSMC, if anything, raised its 2026 capex from $52 billion-$56 billion to $60 billion-$64 billion.&lt;/li&gt;
&lt;li&gt;In HBM, CXMT trails the leading three by &lt;strong&gt;1.5 to 2 product generations&lt;/strong&gt; and by a 2- to 3-year commercialization gap. The 2028 threat is therefore a &lt;strong&gt;conditional option&lt;/strong&gt;, not a present-day supply shock.&lt;/li&gt;
&lt;li&gt;The &lt;strong&gt;technical diffusion&lt;/strong&gt; of Chinese models and the &lt;strong&gt;revenue diffusion&lt;/strong&gt; of Chinese API vendors are two different things. The most realistic path is rehosting Chinese models on AWS, Azure, or enterprise VPCs and selling them through Western security and contracting frameworks. [Analysis scope]&lt;/li&gt;
&lt;/ul&gt;
&lt;hr&gt;
&lt;div class="thesis-callout"&gt;
 &lt;div class="thesis-callout__label"&gt;Key Framing&lt;/div&gt;
 &lt;div class="thesis-callout__body"&gt;
 Chinese models can shake the pricing of US models, but they do not immediately collapse demand for AWS, Azure, US chips, or Korean memory. The direct casualty is the margin of closed-model APIs. What survives longest is usage-based infrastructure: the cloud distribution and security layer, along with memory, networking, and power.
 &lt;/div&gt;
&lt;/div&gt;
&lt;hr&gt;
&lt;h2 id="1-getting-the-facts-straight-first"&gt;1. Getting the Facts Straight First
&lt;/h2&gt;&lt;p&gt;The ecosystem direction is real, but &lt;strong&gt;not every new model has disclosed its training hardware&lt;/strong&gt;. That distinction matters.&lt;/p&gt;
&lt;p&gt;DeepSeek-V3 was trained on &lt;strong&gt;2,048 NVIDIA H800 GPUs&lt;/strong&gt; using 2.788 million GPU-hours. [Fact: DeepSeek V3 technical report] The H800 is export-restricted, but it is not a cheap consumer GPU. Kimi K3, by contrast, has disclosed 2.8 trillion parameters, a 1M-token context window, and 2.5x higher scaling efficiency versus K2, but &lt;strong&gt;the full technical report and training hardware are scheduled for release on July 27&lt;/strong&gt;. [Fact: Kimi official announcement]&lt;/p&gt;
&lt;p&gt;So the claim that &amp;ldquo;Kimi K3 was also trained on cheap NVIDIA GPUs&amp;rdquo; cannot yet be confirmed. [Blocked]&lt;/p&gt;
&lt;h3 id="the-efficiency-gains-are-real"&gt;The Efficiency Gains Are Real
&lt;/h3&gt;&lt;p&gt;DeepSeek V4-Pro &lt;strong&gt;activates only 49 billion&lt;/strong&gt; of its 1.6 trillion parameters. At the 1M-token range, per-token compute is &lt;strong&gt;27%&lt;/strong&gt; and KV cache is &lt;strong&gt;10%&lt;/strong&gt; of V3.2&amp;rsquo;s levels. [Fact: DeepSeek V4 model card] That is direct evidence that GPU and HBM use per token can fall sharply while performance holds.&lt;/p&gt;
&lt;h3 id="but-so-is-the-other-side"&gt;But So Is the Other Side
&lt;/h3&gt;&lt;p&gt;Huawei&amp;rsquo;s CloudMatrix384 pools 384 Ascend 910C chips to serve DeepSeek-R1. By JPMAM&amp;rsquo;s tally, CloudMatrix uses 49TB of HBM and 599kW of power, versus 21TB and 145kW for the comparison system, GB300 NVL72. [Fact: CloudMatrix paper, JPMAM comparison]&lt;/p&gt;
&lt;p&gt;It is not an apples-to-apples performance comparison, but the direction is clear: China compensates for a weaker single chip with &lt;strong&gt;more chips, more memory, and more power&lt;/strong&gt;. A cheaper individual chip does not necessarily mean less total silicon, networking, or power. [Inference: structural reading]&lt;/p&gt;
&lt;hr&gt;
&lt;h2 id="2-the-core-equation-what-to-actually-watch"&gt;2. The Core Equation: What to Actually Watch
&lt;/h2&gt;&lt;p&gt;The answer to this debate is not in benchmark scores. It is in two equations.&lt;/p&gt;
&lt;div class="highlight"&gt;&lt;pre tabindex="0" style="color:#f8f8f2;background-color:#272822;-moz-tab-size:4;-o-tab-size:4;tab-size:4;-webkit-text-size-adjust:none;"&gt;&lt;code class="language-text" data-lang="text"&gt;&lt;span style="display:flex;"&gt;&lt;span&gt;Total compute demand = total tokens × compute per token
&lt;/span&gt;&lt;/span&gt;&lt;span style="display:flex;"&gt;&lt;span&gt;
&lt;/span&gt;&lt;/span&gt;&lt;span style="display:flex;"&gt;&lt;span&gt;Total memory demand = total tokens × memory use per token
&lt;/span&gt;&lt;/span&gt;&lt;span style="display:flex;"&gt;&lt;span&gt; + model, KV, and retrieval data storage
&lt;/span&gt;&lt;/span&gt;&lt;/code&gt;&lt;/pre&gt;&lt;/div&gt;&lt;p&gt;If open models cut token prices 70% and usage rises 5x, &lt;strong&gt;total demand increases&lt;/strong&gt; even after the efficiency gain. Conversely, if usage merely doubles while per-token HBM falls 70%, &lt;strong&gt;HBM demand declines&lt;/strong&gt;.&lt;/p&gt;
&lt;p&gt;So the indicator to watch going forward compresses into one question: &lt;strong&gt;by how much does the token growth rate outpace the decline rate in memory per token?&lt;/strong&gt;&lt;/p&gt;
&lt;h3 id="the-verdict-so-far"&gt;The Verdict So Far
&lt;/h3&gt;&lt;p&gt;Total spend is determined by the following equation.&lt;/p&gt;
&lt;div class="highlight"&gt;&lt;pre tabindex="0" style="color:#f8f8f2;background-color:#272822;-moz-tab-size:4;-o-tab-size:4;tab-size:4;-webkit-text-size-adjust:none;"&gt;&lt;code class="language-text" data-lang="text"&gt;&lt;span style="display:flex;"&gt;&lt;span&gt;Total silicon spend = number of training runs × cost per run
&lt;/span&gt;&lt;/span&gt;&lt;span style="display:flex;"&gt;&lt;span&gt; + inference tokens × cost per token
&lt;/span&gt;&lt;/span&gt;&lt;/code&gt;&lt;/pre&gt;&lt;/div&gt;&lt;p&gt;On its 2Q26 call, TSMC raised its 2026 capex from $52 billion-$56 billion to &lt;strong&gt;$60 billion-$64 billion&lt;/strong&gt;. Microsoft, too, raised inference throughput 40%, yet large-customer token use still rose &lt;strong&gt;30% quarter-over-quarter&lt;/strong&gt;, and it held roughly $190 billion in 2026 capex. [Fact: TSMC 2Q26 call, Microsoft FY26 Q3 call]&lt;/p&gt;
&lt;p&gt;&lt;strong&gt;So far, usage growth is beating efficiency gains.&lt;/strong&gt; [Inference: data synthesis]&lt;/p&gt;
&lt;hr&gt;
&lt;h2 id="3-semiconductor-impact-it-diverges-by-stock"&gt;3. Semiconductor Impact: It Diverges by Stock
&lt;/h2&gt;&lt;table&gt;
 &lt;thead&gt;
 &lt;tr&gt;
 &lt;th&gt;Stock / Group&lt;/th&gt;
 &lt;th&gt;Stock-Price Impact&lt;/th&gt;
 &lt;th&gt;Key Interpretation&lt;/th&gt;
 &lt;/tr&gt;
 &lt;/thead&gt;
 &lt;tbody&gt;
 &lt;tr&gt;
 &lt;td&gt;Samsung Electronics&lt;/td&gt;
 &lt;td&gt;Positive&lt;/td&gt;
 &lt;td&gt;Broadest beneficiary, capturing not just HBM but server DRAM, NAND/eSSD, and general-purpose memory&lt;/td&gt;
 &lt;/tr&gt;
 &lt;tr&gt;
 &lt;td&gt;SK Hynix&lt;/td&gt;
 &lt;td&gt;Mixed-to-positive&lt;/td&gt;
 &lt;td&gt;Token growth helps, but MoE and KV compression cut HBM per token, pressuring the scarcity multiple&lt;/td&gt;
 &lt;/tr&gt;
 &lt;tr&gt;
 &lt;td&gt;Micron&lt;/td&gt;
 &lt;td&gt;Positive&lt;/td&gt;
 &lt;td&gt;Combined DRAM, HBM, and NAND exposure across the US supply chain; broad-memory upside similar to Samsung&lt;/td&gt;
 &lt;/tr&gt;
 &lt;tr&gt;
 &lt;td&gt;SanDisk&lt;/td&gt;
 &lt;td&gt;Positive&lt;/td&gt;
 &lt;td&gt;Local deployment, model weights, and growing RAG/cache use flow into eSSD and NAND demand&lt;/td&gt;
 &lt;/tr&gt;
 &lt;tr&gt;
 &lt;td&gt;NVIDIA&lt;/td&gt;
 &lt;td&gt;Negative near-term, mixed long-term&lt;/td&gt;
 &lt;td&gt;China share and top-tier GPU monopoly value decline; offset by rising total tokens and H200 shipments&lt;/td&gt;
 &lt;/tr&gt;
 &lt;tr&gt;
 &lt;td&gt;AMD&lt;/td&gt;
 &lt;td&gt;Relatively positive&lt;/td&gt;
 &lt;td&gt;Open models make it easier to migrate across heterogeneous hardware; ROCm and actual serving share are the key variables&lt;/td&gt;
 &lt;/tr&gt;
 &lt;tr&gt;
 &lt;td&gt;Broadcom / Marvell / Arista&lt;/td&gt;
 &lt;td&gt;Positive medium-term&lt;/td&gt;
 &lt;td&gt;Rising demand for Western custom ASICs, Ethernet, optical, and SerDes&lt;/td&gt;
 &lt;/tr&gt;
 &lt;tr&gt;
 &lt;td&gt;TSMC&lt;/td&gt;
 &lt;td&gt;Neutral-to-positive&lt;/td&gt;
 &lt;td&gt;Training-chip demand from NVIDIA, AMD, and ASICs holds, but Chinese inference shifts to Ascend and SMIC&lt;/td&gt;
 &lt;/tr&gt;
 &lt;/tbody&gt;
&lt;/table&gt;
&lt;h3 id="why-samsung-has-the-relative-edge-over-hynix"&gt;Why Samsung Has the Relative Edge Over Hynix
&lt;/h3&gt;&lt;p&gt;Among the same three memory makers, &lt;strong&gt;Samsung Electronics and SK Hynix diverge in direction&lt;/strong&gt;. Cheap inference and open-model diffusion do not just lift HBM; they raise the total volume of server DRAM, NAND, and general-purpose memory. Samsung Electronics, with its broader exposure, captures more of that diffusion.&lt;/p&gt;
&lt;p&gt;MoE and KV compression, conversely, reduce &lt;strong&gt;HBM per token&lt;/strong&gt;. Hynix, with its concentrated HBM exposure, receives both the benefit of rising tokens and the burden of falling HBM per token at the same time. It is a structure where the &lt;strong&gt;scarcity multiple&lt;/strong&gt;, not absolute demand, comes under pressure first. [Inference: exposure structure analysis]&lt;/p&gt;
&lt;p&gt;Samsung Electronics does carry an offsetting risk, however: it is the first to be exposed to CXMT&amp;rsquo;s ramp-up of general-purpose DRAM output.&lt;/p&gt;
&lt;h3 id="nvidia-and-h200"&gt;NVIDIA and H200
&lt;/h3&gt;&lt;p&gt;The US recently began shipping limited volumes of H200 to China, but the quantity is still small. It is positive for NVIDIA&amp;rsquo;s near-term revenue, but not enough to reverse the shift toward Huawei-centered domestic infrastructure. [Fact: July 2026 Reuters reporting] [Inference: impact judgment]&lt;/p&gt;
&lt;hr&gt;
&lt;h2 id="4-big-tech-impact"&gt;4. Big Tech Impact
&lt;/h2&gt;&lt;table&gt;
 &lt;thead&gt;
 &lt;tr&gt;
 &lt;th&gt;Company&lt;/th&gt;
 &lt;th&gt;Verdict&lt;/th&gt;
 &lt;th&gt;Reason&lt;/th&gt;
 &lt;/tr&gt;
 &lt;/thead&gt;
 &lt;tbody&gt;
 &lt;tr&gt;
 &lt;td&gt;Meta&lt;/td&gt;
 &lt;td&gt;Most positive&lt;/td&gt;
 &lt;td&gt;Open-ecosystem strategy is validated, and AI returns are captured through advertising and recommendation rather than APIs&lt;/td&gt;
 &lt;/tr&gt;
 &lt;tr&gt;
 &lt;td&gt;Amazon&lt;/td&gt;
 &lt;td&gt;Positive&lt;/td&gt;
 &lt;td&gt;Sells AWS inference, storage, and networking regardless of which model wins&lt;/td&gt;
 &lt;/tr&gt;
 &lt;tr&gt;
 &lt;td&gt;Google&lt;/td&gt;
 &lt;td&gt;Mixed-to-positive&lt;/td&gt;
 &lt;td&gt;TPU and cloud benefit, but Gemini API price premium comes under pressure&lt;/td&gt;
 &lt;/tr&gt;
 &lt;tr&gt;
 &lt;td&gt;Microsoft&lt;/td&gt;
 &lt;td&gt;Mixed&lt;/td&gt;
 &lt;td&gt;Azure usage rises, but OpenAI model rent and capex payback are pressured&lt;/td&gt;
 &lt;/tr&gt;
 &lt;tr&gt;
 &lt;td&gt;Apple&lt;/td&gt;
 &lt;td&gt;Positive&lt;/td&gt;
 &lt;td&gt;Cheaper small and open models lower on-device and Private Cloud costs&lt;/td&gt;
 &lt;/tr&gt;
 &lt;tr&gt;
 &lt;td&gt;OpenAI / Anthropic&lt;/td&gt;
 &lt;td&gt;Negative&lt;/td&gt;
 &lt;td&gt;Shrinking performance gap and API price premium pressure high valuations and fundraising&lt;/td&gt;
 &lt;/tr&gt;
 &lt;/tbody&gt;
&lt;/table&gt;
&lt;p&gt;&lt;strong&gt;Contrary to a common misconception, Meta is not the biggest casualty.&lt;/strong&gt; It makes money from advertising and recommendation rather than model sales, and it uses open models to cut costs, so it stands to relatively benefit. The burden falls instead on capex and depreciation. [Inference: business model analysis]&lt;/p&gt;
&lt;h3 id="the-lesson-from-the-2025-deepseek-shock"&gt;The Lesson from the 2025 DeepSeek Shock
&lt;/h3&gt;&lt;p&gt;During the 2025 DeepSeek shock, NVIDIA lost &lt;strong&gt;17% and roughly $593 billion&lt;/strong&gt; in market capitalization in a single day. The initial reaction was a broad sell-off across GPUs, power, and data-center infrastructure, but it partly rebounded soon after. [Fact: 2025 market reporting]&lt;/p&gt;
&lt;p&gt;This time, too, &lt;strong&gt;near-term stock prices are likely to react to efficiency fears, while medium-term earnings react to rising usage&lt;/strong&gt;. [Inference: historical pattern]&lt;/p&gt;
&lt;hr&gt;
&lt;h2 id="5-verdicts-on-the-chinese-open-model-thesis"&gt;5. Verdicts on the Chinese Open-Model Thesis
&lt;/h2&gt;&lt;p&gt;Judging the claims coming from both the bull and bear sides, one by one, looks like this.&lt;/p&gt;
&lt;table&gt;
 &lt;thead&gt;
 &lt;tr&gt;
 &lt;th&gt;Claim&lt;/th&gt;
 &lt;th&gt;Verdict&lt;/th&gt;
 &lt;/tr&gt;
 &lt;/thead&gt;
 &lt;tbody&gt;
 &lt;tr&gt;
 &lt;td&gt;Frontier performance requires top-tier GPUs&lt;/td&gt;
 &lt;td&gt;Weakened, but not disproven&lt;/td&gt;
 &lt;/tr&gt;
 &lt;tr&gt;
 &lt;td&gt;AI intelligence is a scarce resource&lt;/td&gt;
 &lt;td&gt;Largely collapsed at the level of raw model and token pricing&lt;/td&gt;
 &lt;/tr&gt;
 &lt;tr&gt;
 &lt;td&gt;Capital scale is the moat&lt;/td&gt;
 &lt;td&gt;The pretraining moat weakens; the moat shifts to deployment, data, power, and distribution&lt;/td&gt;
 &lt;/tr&gt;
 &lt;tr&gt;
 &lt;td&gt;Open source has zero marginal cost&lt;/td&gt;
 &lt;td&gt;Only the weight price approaches zero; inference cost is ongoing&lt;/td&gt;
 &lt;/tr&gt;
 &lt;tr&gt;
 &lt;td&gt;A $1 billion training run has overtaken $100 billion of investment&lt;/td&gt;
 &lt;td&gt;An inaccurate claim comparing two different cost categories&lt;/td&gt;
 &lt;/tr&gt;
 &lt;/tbody&gt;
&lt;/table&gt;
&lt;p&gt;The last item is especially often misused. Training cost and total infrastructure investment are not comparable categories.&lt;/p&gt;
&lt;hr&gt;
&lt;h2 id="6-how-far-has-cxmts-hbm-actually-gotten"&gt;6. How Far Has CXMT&amp;rsquo;s HBM Actually Gotten
&lt;/h2&gt;&lt;p&gt;This is the part of the China-threat discussion that gets exaggerated most often. Breaking it down gate by gate reveals the reality.&lt;/p&gt;
&lt;table&gt;
 &lt;thead&gt;
 &lt;tr&gt;
 &lt;th&gt;Gate&lt;/th&gt;
 &lt;th&gt;Current Verdict&lt;/th&gt;
 &lt;th&gt;Basis for Judgment&lt;/th&gt;
 &lt;/tr&gt;
 &lt;/thead&gt;
 &lt;tbody&gt;
 &lt;tr&gt;
 &lt;td&gt;DRAM cell process&lt;/td&gt;
 &lt;td&gt;Commercialized, improving fast&lt;/td&gt;
 &lt;td&gt;Selling DDR5 and LPDDR5X; Apple is also testing DRAM for China-market products&lt;/td&gt;
 &lt;/tr&gt;
 &lt;tr&gt;
 &lt;td&gt;TSV stacking&lt;/td&gt;
 &lt;td&gt;Small-volume HBM2, early HBM3&lt;/td&gt;
 &lt;td&gt;HBM2 production has been reported, but volume and yield are undisclosed&lt;/td&gt;
 &lt;/tr&gt;
 &lt;tr&gt;
 &lt;td&gt;Packaging and base die&lt;/td&gt;
 &lt;td&gt;Under construction&lt;/td&gt;
 &lt;td&gt;The domestic ecosystem is forming, but mass-production yield, thermal, and reliability data are absent&lt;/td&gt;
 &lt;/tr&gt;
 &lt;tr&gt;
 &lt;td&gt;Customer qualification&lt;/td&gt;
 &lt;td&gt;HBM unconfirmed&lt;/td&gt;
 &lt;td&gt;The Tencent contract and Apple testing are evidence of general DRAM capability, not HBM&lt;/td&gt;
 &lt;/tr&gt;
 &lt;tr&gt;
 &lt;td&gt;Distance from the leaders&lt;/td&gt;
 &lt;td&gt;1.5-2 generations&lt;/td&gt;
 &lt;td&gt;The leading three are ramping HBM4; CXMT has not even verified HBM3 mass production&lt;/td&gt;
 &lt;/tr&gt;
 &lt;/tbody&gt;
&lt;/table&gt;
&lt;p&gt;As of July 17, 2026, CXMT&amp;rsquo;s official published product lineup includes only DDR5, LPDDR5/5X, DDR4, and LPDDR4X, and &lt;strong&gt;no HBM&lt;/strong&gt;. TrendForce likewise classifies CXMT&amp;rsquo;s HBM3 as still in early verification, and assesses that technical barriers and domestic-equipment requirements are delaying mass production. [Fact: CXMT official materials, TrendForce]&lt;/p&gt;
&lt;p&gt;Capacity estimates also diverge. A plateau near 240,000 wafers per month conflicts with a year-end forecast of 350,000, and because the 350,000 figure comes from a private model rather than company guidance, it is hard to treat as a confirmed number. [Blocked]&lt;/p&gt;
&lt;h3 id="how-the-legacy-dram-hypothesis-needs-to-be-revised"&gt;How the Legacy-DRAM Hypothesis Needs to Be Revised
&lt;/h3&gt;&lt;ul&gt;
&lt;li&gt;&lt;strong&gt;2026-2027&lt;/strong&gt;: CXMT&amp;rsquo;s HBM investment &lt;strong&gt;delays&lt;/strong&gt; the easing of Chinese legacy-DRAM supply.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;2028&lt;/strong&gt;: If HBM3/3E secures Chinese accelerator customer qualification and meaningful volume, it could begin displacing the older-generation HBM market within China first.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;2029 and beyond&lt;/strong&gt;: Only once HBM4, base die, and packaging yield catch up does it directly pressure the global leading market and Hynix&amp;rsquo;s margins.&lt;/li&gt;
&lt;/ul&gt;
&lt;p&gt;In other words, rather than &amp;ldquo;CXMT is making legacy supply tighter right now,&amp;rdquo; the more valid framing is &lt;strong&gt;&amp;ldquo;CXMT is not freeing up as much legacy supply as expected.&amp;quot;&lt;/strong&gt; [Inference: stage-by-stage judgment]&lt;/p&gt;
&lt;hr&gt;
&lt;h2 id="7-how-us-policy-reshapes-this-reading"&gt;7. How US Policy Reshapes This Reading
&lt;/h2&gt;&lt;p&gt;The US is treating AI less as a commercial technology and more as &lt;strong&gt;core infrastructure for the allied bloc&lt;/strong&gt;. That is why a purely technical analysis cannot supply the full answer.&lt;/p&gt;
&lt;p&gt;The US AI Action Plan and Executive Order 14320 state explicitly that the US intends to export a &lt;strong&gt;full-stack American AI system&lt;/strong&gt;, bundling hardware, cloud, networking, models, and applications, to allied nations, while reducing technological dependence on adversary nations. Even if a Chinese model is superior on performance and price, the US stack retains a policy advantage in allied-nation government procurement and critical industries. [Fact: White House AI Action Plan, EO 14320]&lt;/p&gt;
&lt;h3 id="yet-this-is-not-a-full-decoupling"&gt;Yet This Is Not a Full Decoupling
&lt;/h3&gt;&lt;p&gt;In January 2026, the US BIS moved to review H200 and MI325X exports to China case by case, under approved-customer and security conditions. It is a compromise that keeps China partly inside the US chip ecosystem while retaining control. [Fact: BIS 2026-01-13]&lt;/p&gt;
&lt;p&gt;Nor is the use of Chinese models by US private companies fully banned. The &lt;code&gt;No DeepSeek on Government Devices Act&lt;/code&gt; is still only at the bill-introduction stage. Australia&amp;rsquo;s government, however, has ordered DeepSeek removed from government systems, and Italy&amp;rsquo;s privacy authority has restricted processing of user data. [Fact: official actions by country]&lt;/p&gt;
&lt;p&gt;&lt;strong&gt;De facto bans from procurement and security departments are likely to take effect before legislation does.&lt;/strong&gt; [Inference: policy sequencing]&lt;/p&gt;
&lt;hr&gt;
&lt;h2 id="8-can-western-enterprises-actually-use-chinese-model-apis"&gt;8. Can Western Enterprises Actually Use Chinese Model APIs?
&lt;/h2&gt;&lt;p&gt;Diffusion potential differs completely by pathway. This table is the answer to the question.&lt;/p&gt;
&lt;table&gt;
 &lt;thead&gt;
 &lt;tr&gt;
 &lt;th&gt;Adoption Method&lt;/th&gt;
 &lt;th&gt;Diffusion Potential&lt;/th&gt;
 &lt;th&gt;Judgment&lt;/th&gt;
 &lt;/tr&gt;
 &lt;/thead&gt;
 &lt;tbody&gt;
 &lt;tr&gt;
 &lt;td&gt;Direct calls to mainland China-hosted APIs&lt;/td&gt;
 &lt;td&gt;Low&lt;/td&gt;
 &lt;td&gt;Data, jurisdiction, and procurement risk&lt;/td&gt;
 &lt;/tr&gt;
 &lt;tr&gt;
 &lt;td&gt;Qwen&amp;rsquo;s US, EU, Japan, and Singapore APIs&lt;/td&gt;
 &lt;td&gt;Medium&lt;/td&gt;
 &lt;td&gt;Data localization is possible, but Chinese-vendor risk remains&lt;/td&gt;
 &lt;/tr&gt;
 &lt;tr&gt;
 &lt;td&gt;Chinese models rehosted by AWS or Azure&lt;/td&gt;
 &lt;td&gt;Medium-high to high&lt;/td&gt;
 &lt;td&gt;Western cloud providers hold the contracting, security, and data control&lt;/td&gt;
 &lt;/tr&gt;
 &lt;tr&gt;
 &lt;td&gt;Enterprise VPC or on-premises open weights&lt;/td&gt;
 &lt;td&gt;High&lt;/td&gt;
 &lt;td&gt;No need to send data to Chinese servers&lt;/td&gt;
 &lt;/tr&gt;
 &lt;tr&gt;
 &lt;td&gt;Government, defense, and critical infrastructure&lt;/td&gt;
 &lt;td&gt;Very low&lt;/td&gt;
 &lt;td&gt;Procurement restrictions can extend even to model lineage&lt;/td&gt;
 &lt;/tr&gt;
 &lt;/tbody&gt;
&lt;/table&gt;
&lt;p&gt;The most realistic pathway looks like this.&lt;/p&gt;
&lt;div class="highlight"&gt;&lt;pre tabindex="0" style="color:#f8f8f2;background-color:#272822;-moz-tab-size:4;-o-tab-size:4;tab-size:4;-webkit-text-size-adjust:none;"&gt;&lt;code class="language-text" data-lang="text"&gt;&lt;span style="display:flex;"&gt;&lt;span&gt;Chinese model developed
&lt;/span&gt;&lt;/span&gt;&lt;span style="display:flex;"&gt;&lt;span&gt;→ Rehosted on AWS, Azure, or enterprise VPC
&lt;/span&gt;&lt;/span&gt;&lt;span style="display:flex;"&gt;&lt;span&gt;→ Sold through Western security, contracting, and audit frameworks
&lt;/span&gt;&lt;/span&gt;&lt;/code&gt;&lt;/pre&gt;&lt;/div&gt;&lt;p&gt;In other words, &lt;strong&gt;the technical diffusion of Chinese models and the revenue diffusion of Chinese API vendors are separate matters&lt;/strong&gt;. [Inference: pathway analysis]&lt;/p&gt;
&lt;h3 id="the-limits-of-direct-apis"&gt;The Limits of Direct APIs
&lt;/h3&gt;&lt;p&gt;DeepSeek&amp;rsquo;s Privacy Policy states that it &lt;strong&gt;collects, processes, and stores personal data, including input data, directly in China&lt;/strong&gt;, and may use it to improve its service and models. [Fact: DeepSeek Privacy Policy] That is a disqualifying condition for companies handling source code, customer information, healthcare or financial data, or export-controlled technology.&lt;/p&gt;
&lt;p&gt;Qwen is somewhat different. Alibaba Cloud Model Studio offers US, German, Japanese, and Singapore regions, can restrict data and inference to a specific region, and states explicitly that it does not use customer data for model training. [Fact: Alibaba Cloud region and security policy] Direct API adoption could therefore grow in non-regulated industries and in Southeast Asia, the Middle East, and Latin America.&lt;/p&gt;
&lt;p&gt;Data localization and SOC 2, however, do not eliminate &lt;strong&gt;the risk of service disruption, sanctions, or procurement exposure arising from US-China tension&lt;/strong&gt;. [Inference: residual risk]&lt;/p&gt;
&lt;hr&gt;
&lt;h2 id="9-revising-the-thesis-what-changes-and-what-stays"&gt;9. Revising the Thesis: What Changes and What Stays
&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;First, the case for a US hyperscaler collapse weakens.&lt;/strong&gt; AWS and Azure directly host DeepSeek, providing data isolation, SLAs, and security assessments. US clouds can absorb the cost innovation of Chinese models and turn it into revenue. [Fact: AWS and Azure official announcements]&lt;/p&gt;
&lt;p&gt;&lt;strong&gt;Second, pricing pressure hits closed-model APIs first.&lt;/strong&gt; It burdens the token prices and margins of OpenAI, Anthropic, and Google, but inference volume and AWS/Azure usage can still rise. It is relatively favorable for Meta&amp;rsquo;s open-model strategy.&lt;/p&gt;
&lt;p&gt;&lt;strong&gt;Third, the US-China split supports total infrastructure investment.&lt;/strong&gt; Both blocs build out &lt;strong&gt;duplicate&lt;/strong&gt; accelerators, memory, networking, and power grids. That lowers the likelihood that efficiency gains translate directly into a decline in global silicon spend.&lt;/p&gt;
&lt;p&gt;&lt;strong&gt;Fourth, SK Hynix&amp;rsquo;s HBM premium can persist longer within the allied bloc.&lt;/strong&gt; CXMT&amp;rsquo;s HBM is more likely to penetrate Chinese accelerators and Chinese data centers first. Adoption by Western CSPs requires policy and supply-chain certification on top of technical qualification. That said, export controls accelerate China&amp;rsquo;s push for self-sufficiency, which is a risk to Hynix&amp;rsquo;s China-market share from 2028 onward.&lt;/p&gt;
&lt;p&gt;&lt;strong&gt;Fifth, Samsung Electronics is a relative hedge.&lt;/strong&gt; Cheap inference and open-model diffusion can raise the total volume of server DRAM, NAND, and general-purpose memory, not just HBM. On the other side is the risk of being first exposed to CXMT&amp;rsquo;s ramp-up of general-purpose DRAM output.&lt;/p&gt;
&lt;hr&gt;
&lt;h2 id="10-the-order-of-casualties-and-beneficiaries"&gt;10. The Order of Casualties and Beneficiaries
&lt;/h2&gt;&lt;ol&gt;
&lt;li&gt;&lt;strong&gt;Greatest casualty&lt;/strong&gt;: the excess profit of closed-model APIs, and highly leveraged GPU-leasing operators&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Intermediate risk&lt;/strong&gt;: heavily front-invested, customer-concentrated operators such as Oracle&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Meta&lt;/strong&gt;: not the biggest casualty. It uses open models to cut costs, so relative benefit is possible. The burden falls on capex and depreciation&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;NVIDIA&lt;/strong&gt;: the multiple and product mix come under pressure before near-term EPS does. Displacement inside China is a risk, but the CUDA, networking, power-efficiency, and development-timeline moats remain&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Memory&lt;/strong&gt;: the base case is not a decline in absolute HBM demand, but &lt;strong&gt;a narrowing HBM scarcity premium plus tier expansion into DDR/CXL/eSSD&lt;/strong&gt;&lt;/li&gt;
&lt;/ol&gt;
&lt;hr&gt;
&lt;h2 id="11-scenario-update"&gt;11. Scenario Update
&lt;/h2&gt;&lt;p&gt;It makes sense to provisionally revise the probabilities used in &lt;a class="link" href="https://koreainvestinsights.com/post/memory-fair-value-fcfe-terminal-samsung-hynix-micron-2026-07-17/" &gt;Are Semiconductors Cyclical, and What Is Fair Value?&lt;/a&gt;&lt;/p&gt;
&lt;table&gt;
 &lt;thead&gt;
 &lt;tr&gt;
 &lt;th&gt;Scenario&lt;/th&gt;
 &lt;th style="text-align: right"&gt;Prior&lt;/th&gt;
 &lt;th style="text-align: right"&gt;Revised&lt;/th&gt;
 &lt;/tr&gt;
 &lt;/thead&gt;
 &lt;tbody&gt;
 &lt;tr&gt;
 &lt;td&gt;Excess demand persists&lt;/td&gt;
 &lt;td style="text-align: right"&gt;30%&lt;/td&gt;
 &lt;td style="text-align: right"&gt;30%&lt;/td&gt;
 &lt;/tr&gt;
 &lt;tr&gt;
 &lt;td&gt;Asset re-concentration&lt;/td&gt;
 &lt;td style="text-align: right"&gt;40%&lt;/td&gt;
 &lt;td style="text-align: right"&gt;40%&lt;/td&gt;
 &lt;/tr&gt;
 &lt;tr&gt;
 &lt;td&gt;Supply/efficiency normalization&lt;/td&gt;
 &lt;td style="text-align: right"&gt;20%&lt;/td&gt;
 &lt;td style="text-align: right"&gt;&lt;strong&gt;25%&lt;/strong&gt;&lt;/td&gt;
 &lt;/tr&gt;
 &lt;tr&gt;
 &lt;td&gt;System demand short-circuit&lt;/td&gt;
 &lt;td style="text-align: right"&gt;10%&lt;/td&gt;
 &lt;td style="text-align: right"&gt;&lt;strong&gt;5%&lt;/strong&gt;&lt;/td&gt;
 &lt;/tr&gt;
 &lt;/tbody&gt;
&lt;/table&gt;
&lt;p&gt;Chinese model performance &lt;strong&gt;lowers the probability that AI demand itself disappears&lt;/strong&gt; (the short-circuit scenario, from 10% to 5%). In exchange, efficiency gains and Chinese-made hardware lower the scarcity premium of NVIDIA and HBM (the normalization scenario, from 20% to 25%).&lt;/p&gt;
&lt;p&gt;The timeline scenarios also hold: usage beats efficiency through 2027 at 70%, mix and pricing normalize from 2028 onward at 25%, and total capex contracts at 5%. However, &lt;strong&gt;the driver of the 70% scenario shifts from a single global ecosystem to dual investment across the US bloc and the China bloc.&lt;/strong&gt; [Inference: scenario recalibration]&lt;/p&gt;
&lt;hr&gt;
&lt;h2 id="12-what-would-change-this-judgment"&gt;12. What Would Change This Judgment
&lt;/h2&gt;&lt;ul&gt;
&lt;li&gt;&lt;strong&gt;The Kimi K3 technical report (July 27)&lt;/strong&gt;: once training hardware is disclosed, the truth of the &amp;ldquo;frontier training on cheap GPUs&amp;rdquo; claim will be settled&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Token growth rate versus the decline rate in memory per token&lt;/strong&gt;: the gap between these two values is the real answer to this debate&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;CXMT&amp;rsquo;s HBM3E mass-production qualification and actual packaging volume&lt;/strong&gt;: if confirmed, both Hynix&amp;rsquo;s 2028 EPS and its multiple need to come down together&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Slowing HBM pricing and content growth&lt;/strong&gt;: the first signal of a narrowing scarcity premium&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Expanding rehosting of Chinese models by Western CSPs&lt;/strong&gt;: evidence that cloud-distribution value is rising&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Actual enforcement of US procurement and security regulation&lt;/strong&gt;: the scope of the de facto ban that operates ahead of legislation&lt;/li&gt;
&lt;/ul&gt;
&lt;p&gt;This is not the stage to conflate &lt;strong&gt;terminal risk with current-period earnings&lt;/strong&gt;. Hynix&amp;rsquo;s 2028-and-beyond outlook can be adjusted once CXMT&amp;rsquo;s HBM3E mass-production qualification is confirmed. [Inference: sequencing of judgment]&lt;/p&gt;
&lt;hr&gt;
&lt;h2 id="closing"&gt;Closing
&lt;/h2&gt;&lt;p&gt;Returning to the question, the answer is this.&lt;/p&gt;
&lt;p&gt;It is a fact that Chinese open models are converging on the US frontier, and it is a fact that cost per unit of intelligence has fallen sharply. But that does not mean total silicon spend is declining. TSMC&amp;rsquo;s capex increase and Microsoft&amp;rsquo;s 30% rise in tokens are the answer so far.&lt;/p&gt;
&lt;p&gt;US policy substantially reshapes this reading. The US-China split produces &lt;strong&gt;duplicate investment&lt;/strong&gt; across both blocs, supporting total infrastructure demand, and within the allied bloc, it protects the position of the US stack and Korean memory by policy.&lt;/p&gt;
&lt;p&gt;Western enterprise adoption of Chinese model APIs must be viewed by &lt;strong&gt;separating the model from the vendor&lt;/strong&gt;. The technology spreads as open weight, but the party selling it is more likely to be AWS, Azure, and enterprise VPCs than the Chinese API vendors themselves.&lt;/p&gt;
&lt;p&gt;So the conclusion is redistribution. What collapses is the excess profit of closed-model APIs. What remains is usage-based infrastructure.&lt;/p&gt;
&lt;hr&gt;
&lt;p&gt;&lt;em&gt;This post synthesizes public papers (DeepSeek V3/V4, Huawei CloudMatrix384), official company announcements (Kimi, the TSMC 2Q26 call, the Microsoft FY26 Q3 call, CXMT, Alibaba Cloud, DeepSeek&amp;rsquo;s Privacy Policy), US government materials (the AI Action Plan, EO 14320, BIS), regulatory actions by various countries, and market research (TrendForce, JPMAM). Kimi K3&amp;rsquo;s training hardware remains unconfirmed until its technical report is released on July 27, and CXMT&amp;rsquo;s capacity estimates and the scenario probabilities are author estimates as of the time of writing, not company guidance. The stocks mentioned are examples used to illustrate value-chain structure and are not a recommendation to buy or sell any specific security. Investment decisions and responsibility for them rest with the individual investor.&lt;/em&gt;&lt;/p&gt;
&lt;hr&gt;
&lt;h3 id="related-posts"&gt;Related Posts
&lt;/h3&gt;&lt;ul&gt;
&lt;li&gt;&lt;a class="link" href="https://koreainvestinsights.com/post/kimi-k3-linear-api-pricing-semiconductor-big-tech-impact-2026-07-17/" &gt;Kimi K3 Resets the AI Price Curve: From Kimi Linear to HBM and Big Tech Strategy&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a class="link" href="https://koreainvestinsights.com/post/memory-fair-value-fcfe-terminal-samsung-hynix-micron-2026-07-17/" &gt;Are Semiconductors Cyclical, and What Is Fair Value? Pricing Samsung, SK Hynix and Micron with FCFE and Normalized Earnings&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a class="link" href="https://koreainvestinsights.com/post/semiconductor-bull-bear-four-clocks-capital-intensity-cycle-2026-07-17/" &gt;The Real Debate in Semiconductors: Four Physical Clocks and One Stock-Price Clock&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a class="link" href="https://koreainvestinsights.com/post/cxmt-ipo-memory-price-risk-hbm-client-dram-2026-06-21/" &gt;CXMT IPO And Memory Price Risk: HBM Is Not The First Place To Break&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a class="link" href="https://koreainvestinsights.com/post/hbm-2030-supply-demand-267eb-demand-model-crosscheck-2026-07-13/" &gt;HBM 2030 Supply-Demand Deep Research: Dissecting the 26.7EB Demand Model Against the Capacity Schedule&lt;/a&gt;&lt;/li&gt;
&lt;/ul&gt;</description></item></channel></rss>