locsic.com
Thinking
Long-form notes for reading the direction of technical change.
Thinking is a professional column: fewer quick posts, more durable essays. It keeps observations, source trails, assumptions, and evolving viewpoints visible over time.
Essays
41 essays shownJune 2026
41 essaysThe Glass Bridge: How Corning Uses Glass to Remove CPO's Last Production Hurdle
9μm fiber core vs 0.5μm PIC waveguide — CPO's biggest production bottleneck is a glass optical waveguide. From first principles through complete transceiver path analysis,…
When Agents Start Designing Chips: CHIA and the Restructuring of Chip Design Workflows
From Berkeley's CHIA framework to Princeton's AI-RFIC breakthrough — AI is breaching both digital and analog chip design simultaneously. Architects don't write RTL; architects…
When Tokens Stop Being Costs and Start Being Capital: The Economics of Tokenmaxxing 2.0
Empirical validation of compounding correctness is turning inference from opex into capex — changing the competitive dynamics of AI infrastructure
Two Cracks in the Export Control Wall: Apple Courts CXMT, GLM-5.2 Rivals Mythos
Apple lobbies to buy chips from blacklisted Chinese memory maker CXMT; WSJ reports China's open GLM-5.2 matches banned US model Mythos at security bug detection. Independent…
Inside Jalapeño: What Happens When an AI Company Builds Its Own Heart
# Inside Jalapeño: What Happens When an AI Company Builds Its Own Heart On June 24, 2026, OpenAI and Broadcom jointly released Jalapeño, OpenAI's first in-hous
LineShine Addendum: New Details Confirmed by The Next Platform's Deep Dive
**Sources:**
From ISC 2026 to AI4S: HPC Is Shifting from a "Peak-FLOPS Machine" into a "Scientific Validation Engine"
**Declaration:** This article is written based on publicly available information as of June 25, 2026 (Beijing Time), including the ISC 2026 agenda, talk abstracts, the June…
LineShine Tops TOP500: 2 EFLOPS Pure-CPU (6/29 Update)
ISC 2026 confirms LineShine #1 in both TOP500 and HPCG. Top 500 stagnation. 64GB HBM unified. Chips and Cheese analysis.
The Toll Booth in the Throat: Why the AI Compiler Layer Creates Enormous Value but Captures Almost None of It
Starting from Qualcomm’s $4B Modular acquisition, an analysis of the AI compiler layer: technical bottlenecks, value capture trap, and five terminal judgments.
How Agents Go Off Track
When an AI Agent picks the wrong tool at step 7, the remaining 13 steps are pure waste — and traditional monitoring can't see it
Inside the Inference Engine
From Prefill to CUDA Kernel — a Millisecond-Level Breakdown of an Inference Request
Anatomy of an Inference Bill
85% of your AI bill is infrastructure tax; only 15% creates value
The Three Blind Spots of AI Observability
When your monitoring says "all green" while your AI systems quietly burn money, drift off track, and spiral out of control
The $14 Billion Bet: HPE Discover 2026 Strategic Overview
A decade of divestiture, then all-in on networking and AI factories. HPE's first test after the $14B Juniper acquisition: three core judgments, ten structural challenges. Right…
Network as the AI Control Plane: HPE's Networking Gamble
QFX six-tier coverage from training to inference, GreenLake Intelligence four-entry consolidation. But HPE doesn't design its own switch chips. The integrator's margin is always…
Decoding the HPE AI Factory: Compute, Storage, Software, and the Cray Integration Experiment
DL 394 Gen 12, Alletra MPX 10000 MCP-native storage, GreenLake full-stack software, Cray technology repurposed. Six equipment gaps. 2027 will tell.
When AI Agents Become Workloads: HPE's Agent Infrastructure Blueprint
Zero-code registration, three-tier identity, NVIDIA sandbox, MCP-native storage. HPE is the first traditional vendor to build full-stack infrastructure for AI agents.
From Allbirds' AI Infra Pivot: How Product Power and Marketing Sustain Lasting Companies
Allbirds sold its shoes and renamed itself Smartbird—the cost of a company whose entire value sits on narrative. Marketing is the amplifier; product power is the chassis. When…
No Independent Future for Tool Companies Without Their Own Models? SpaceX's $60B Cursor Acquisition and the Endgame of AI Coding
SpaceX acquired Cursor's parent Anysphere for $60B in stock. Simultaneously, Cursor revealed a 1.5T-parameter from-scratch model at its Compile conference. This is more than the…
MaaS Inference Tech Stack: How Six Levers Cut Cost by 96%
Technical dissection of DeepSeek's 96% per-token cost reduction. Six levers, inference engine comparison, and frontier directions.
The Token Distribution Era: MaaS Service Models and Business Anatomy
China's MaaS break-even line is 5-7 yuan/million tokens, with mainstream pricing below cost. How do four service models coexist? Three competitive paths each have distinct…
The Tyranny of Memory: How KV Cache Is Reshaping Every Layer of AI Inference
In the 1M-context era, KV Cache is the defining bottleneck of inference cost.
When SSD Becomes Memory: How AI Inference Is Rewriting the Storage Hierarchy
Three facts, sitting side by side in the first half of 2026, create a tension too sharp to ignore.
Compute Power Is Not Fighting Power: What the SpaceX Colossus Lease Tells Us About AI Infrastructure Reality
220,000 GPUs built and leased out within a year. The SpaceX Colossus lease reveals the vast gap between owning compute and using it effectively.
One Report Wiped Out Optical Stocks: Is CPO Actually Dead?
On June 9, 2026, SemiAnalysis sent a research note to institutional clients titled "Powered Down, Lights Off." By the end of the trading session, AAOI had dropped 14%, COHR 11%,…
RAMageddon: The Memory Famine and Storage Supercycle in AI Data Centers
AI inference will rewrite the storage hierarchy. HBM margins exceed GPUs, NAND has the most elasticity, China's memory gap. FMS 2026 has validated the storage hierarchy rewrite…
When Agents Need a Desk: The Execution Environment War Behind OpenAI's Acquisition of Ona
OpenAI didn't buy Ona for the tool — they bought judgment. A deep dive into Agent execution environment tiers, Big Tech strategies, and the independent platform landscape.
Text Diffusion vs. Autoregressive: The Paradigm War
DiffusionGemma at 1,107 tok/s, Mercury in commercial deployment, Dream 7B matching same-scale AR. Text diffusion went from papers to products in two years, but reasoning quality…
The Agent Payment Protocol War
In 18 months, agent payments went from zero to six competing protocols. Visa and Mastercard are placing separate bets on consumer and machine rails. This analysis breaks down the…
When AI Learns to Lie: A Behavioral Profile of Claude Fable 5
Anthropic released the model it called too dangerous four months ago. Same weights, plus a safety classifier. But the real story isn't the benchmarks—it's the five behavioral…
Making K8s Understand Super-Nodes: openFuyao and the Lingqu Cloud-Layer Breakout
The Lingqu cloud layer wraps hardware capabilities into K8s-native interfaces via openFuyao. InferNex inference cluster orchestration is the most commercially valuable component.…
Heart of the Super-Node: How the Lingqu Service Layer Weaves 8,192 Cards Together
The Lingqu service layer answers how 8,192 cards cooperate — UBS Engine control plane, MemFabric unified memory weaving, HCCL collective communication, NPU Direct storage bypass,…
Making Linux Understand Super-Nodes: Technical Anatomy of the Lingqu Kernel Layer
Lingqu (UnifiedBus) super-nodes require systemic changes to the Linux kernel: a new bus type, cross-node address translation, unified memory management, and URMA communication…
Breaking the Transceiver Bottleneck: How Optical Shuffle Reshapes AI Cluster Economics
Panduit engineer Castro proposes Optical Shuffle at IEEE 802.3, cutting 32K-GPU cluster transceivers by 33% and spine switches by 75%. Orthogonal to AWS RNG, three-layer stacking…
A New Direction for Data Center Networking: What RNG Opens Up
In April 2026, AWS switched all new non-GPU datacenters to a flat topology called RNG. 69% fewer routers, up to 33% better throughput. Not an experiment — the production default.…
The Great Token Retreat: When AI Bills Get More Expensive Than People
Uber burned through its annual AI budget in four months. An unnamed enterprise racked up a $500 million monthly bill on Anthropic. Klarna replaced 700 humans with AI, then…
NVIDIA Rubin Respins: Is AMD GPU Competitiveness for Real?
Fubon Research reveals Rubin was respun due to MI450 pressure. Full analysis of AMD AI infra stack — hardware architecture, software ecosystem, customer deployments, ROCm status.
Build 2026: Microsoft's Agent OS Gambit
Windows is shifting from "an OS that runs apps" to "a platform that runs agents"—this isn't just a change in technical direction; it's a trillion-dollar company redefining its…
From CLOS to ZCube: Network Topology Evolution for AI Computing Clusters
From Charles Clos's non-blocking telephone switching network in 1953 to ByteDance's SIGCOMM 2025 Best Paper ZCube — topology design has evolved from expert intuition to automated…
MRC: When the NIC Becomes the Brain of the Network
OpenAI, together with NVIDIA, AMD, Broadcom, Arista, and Cisco, overturned five long-standing data center networking conventions simultaneously with the MRC protocol. By pushing…
The PC, Reinvented: Computex 2026 and NVIDIA's Infrastructure Ambitions
At GTC Taipei, Jensen Huang unveiled RTX Spark—33 years of technology distilled into a single chip, officially marking the PC's entry into the agent era. Full-volume production…