Industry Signals · Technology Trends · Personal Judgment

Thinking

Long-form notes for reading the direction of technical change.

Thinking is a professional column: fewer quick posts, more durable essays. It keeps observations, source trails, assumptions, and evolving viewpoints visible over time.

Essays

41 essays shown

June 2026

41 essays
Thinking36 min read

The Glass Bridge: How Corning Uses Glass to Remove CPO's Last Production Hurdle

9μm fiber core vs 0.5μm PIC waveguide — CPO's biggest production bottleneck is a glass optical waveguide. From first principles through complete transceiver path analysis,…

AI
Read essay
Thinking13 min read

When Agents Start Designing Chips: CHIA and the Restructuring of Chip Design Workflows

From Berkeley's CHIA framework to Princeton's AI-RFIC breakthrough — AI is breaching both digital and analog chip design simultaneously. Architects don't write RTL; architects…

AI
Read essay
Thinking18 min read

When Tokens Stop Being Costs and Start Being Capital: The Economics of Tokenmaxxing 2.0

Empirical validation of compounding correctness is turning inference from opex into capex — changing the competitive dynamics of AI infrastructure

AI
Read essay
Thinking28 min read

Two Cracks in the Export Control Wall: Apple Courts CXMT, GLM-5.2 Rivals Mythos

Apple lobbies to buy chips from blacklisted Chinese memory maker CXMT; WSJ reports China's open GLM-5.2 matches banned US model Mythos at security bug detection. Independent…

AI
Read essay
Thinking90 min read

Inside Jalapeño: What Happens When an AI Company Builds Its Own Heart

# Inside Jalapeño: What Happens When an AI Company Builds Its Own Heart On June 24, 2026, OpenAI and Broadcom jointly released Jalapeño, OpenAI's first in-hous

AI
Read essay
Thinking28 min read

LineShine Addendum: New Details Confirmed by The Next Platform's Deep Dive

**Sources:**

LineShine
Read essay
Thinking58 min read

From ISC 2026 to AI4S: HPC Is Shifting from a "Peak-FLOPS Machine" into a "Scientific Validation Engine"

**Declaration:** This article is written based on publicly available information as of June 25, 2026 (Beijing Time), including the ISC 2026 agenda, talk abstracts, the June…

AI4S
Read essay
Thinking90 min read

LineShine Tops TOP500: 2 EFLOPS Pure-CPU (6/29 Update)

ISC 2026 confirms LineShine #1 in both TOP500 and HPCG. Top 500 stagnation. 64GB HBM unified. Chips and Cheese analysis.

HPC
Read essay
Thinking24 min read

The Toll Booth in the Throat: Why the AI Compiler Layer Creates Enormous Value but Captures Almost None of It

Starting from Qualcomm’s $4B Modular acquisition, an analysis of the AI compiler layer: technical bottlenecks, value capture trap, and five terminal judgments.

AI
Read essay
Thinking19 min read

How Agents Go Off Track

When an AI Agent picks the wrong tool at step 7, the remaining 13 steps are pure waste — and traditional monitoring can't see it

AI
Read essay
Thinking19 min read

Inside the Inference Engine

From Prefill to CUDA Kernel — a Millisecond-Level Breakdown of an Inference Request

AI
Read essay
Thinking31 min read

Anatomy of an Inference Bill

85% of your AI bill is infrastructure tax; only 15% creates value

AI
Read essay
Thinking18 min read

The Three Blind Spots of AI Observability

When your monitoring says "all green" while your AI systems quietly burn money, drift off track, and spiral out of control

AI
Read essay
Thinking21 min read

The $14 Billion Bet: HPE Discover 2026 Strategic Overview

A decade of divestiture, then all-in on networking and AI factories. HPE's first test after the $14B Juniper acquisition: three core judgments, ten structural challenges. Right…

HPE
Read essay
Thinking22 min read

Network as the AI Control Plane: HPE's Networking Gamble

QFX six-tier coverage from training to inference, GreenLake Intelligence four-entry consolidation. But HPE doesn't design its own switch chips. The integrator's margin is always…

HPE
Read essay
Thinking21 min read

Decoding the HPE AI Factory: Compute, Storage, Software, and the Cray Integration Experiment

DL 394 Gen 12, Alletra MPX 10000 MCP-native storage, GreenLake full-stack software, Cray technology repurposed. Six equipment gaps. 2027 will tell.

HPE
Read essay
Thinking16 min read

When AI Agents Become Workloads: HPE's Agent Infrastructure Blueprint

Zero-code registration, three-tier identity, NVIDIA sandbox, MCP-native storage. HPE is the first traditional vendor to build full-stack infrastructure for AI agents.

HPE
Read essay
Thinking17 min read

From Allbirds' AI Infra Pivot: How Product Power and Marketing Sustain Lasting Companies

Allbirds sold its shoes and renamed itself Smartbird—the cost of a company whose entire value sits on narrative. Marketing is the amplifier; product power is the chassis. When…

DTC
Read essay
Thinking29 min read

No Independent Future for Tool Companies Without Their Own Models? SpaceX's $60B Cursor Acquisition and the Endgame of AI Coding

SpaceX acquired Cursor's parent Anysphere for $60B in stock. Simultaneously, Cursor revealed a 1.5T-parameter from-scratch model at its Compile conference. This is more than the…

AI
Read essay
Thinking24 min read

MaaS Inference Tech Stack: How Six Levers Cut Cost by 96%

Technical dissection of DeepSeek's 96% per-token cost reduction. Six levers, inference engine comparison, and frontier directions.

MaaS
Read essay
Thinking27 min read

The Token Distribution Era: MaaS Service Models and Business Anatomy

China's MaaS break-even line is 5-7 yuan/million tokens, with mainstream pricing below cost. How do four service models coexist? Three competitive paths each have distinct…

MaaS
Read essay
Thinking50 min read

The Tyranny of Memory: How KV Cache Is Reshaping Every Layer of AI Inference

In the 1M-context era, KV Cache is the defining bottleneck of inference cost.

AI
Read essay
Thinking35 min read

When SSD Becomes Memory: How AI Inference Is Rewriting the Storage Hierarchy

Three facts, sitting side by side in the first half of 2026, create a tension too sharp to ignore.

AI
Read essay
Thinking16 min read

Compute Power Is Not Fighting Power: What the SpaceX Colossus Lease Tells Us About AI Infrastructure Reality

220,000 GPUs built and leased out within a year. The SpaceX Colossus lease reveals the vast gap between owning compute and using it effectively.

AI基础设施
Read essay
Thinking11 min read

One Report Wiped Out Optical Stocks: Is CPO Actually Dead?

On June 9, 2026, SemiAnalysis sent a research note to institutional clients titled "Powered Down, Lights Off." By the end of the trading session, AAOI had dropped 14%, COHR 11%,…

CPO
Read essay
Thinking4 min read

RAMageddon: The Memory Famine and Storage Supercycle in AI Data Centers

AI inference will rewrite the storage hierarchy. HBM margins exceed GPUs, NAND has the most elasticity, China's memory gap. FMS 2026 has validated the storage hierarchy rewrite…

存储
Read essay
Thinking31 min read

When Agents Need a Desk: The Execution Environment War Behind OpenAI's Acquisition of Ona

OpenAI didn't buy Ona for the tool — they bought judgment. A deep dive into Agent execution environment tiers, Big Tech strategies, and the independent platform landscape.

AI
Read essay
Thinking13 min read

Text Diffusion vs. Autoregressive: The Paradigm War

DiffusionGemma at 1,107 tok/s, Mercury in commercial deployment, Dream 7B matching same-scale AR. Text diffusion went from papers to products in two years, but reasoning quality…

AI
Read essay
Thinking28 min read

The Agent Payment Protocol War

In 18 months, agent payments went from zero to six competing protocols. Visa and Mastercard are placing separate bets on consumer and machine rails. This analysis breaks down the…

AI
Read essay
Thinking12 min read

When AI Learns to Lie: A Behavioral Profile of Claude Fable 5

Anthropic released the model it called too dangerous four months ago. Same weights, plus a safety classifier. But the real story isn't the benchmarks—it's the five behavioral…

AI
Read essay
Thinking13 min read

Making K8s Understand Super-Nodes: openFuyao and the Lingqu Cloud-Layer Breakout

The Lingqu cloud layer wraps hardware capabilities into K8s-native interfaces via openFuyao. InferNex inference cluster orchestration is the most commercially valuable component.…

AI基础设施
Read essay
Thinking11 min read

Heart of the Super-Node: How the Lingqu Service Layer Weaves 8,192 Cards Together

The Lingqu service layer answers how 8,192 cards cooperate — UBS Engine control plane, MemFabric unified memory weaving, HCCL collective communication, NPU Direct storage bypass,…

AI基础设施
Read essay
Thinking34 min read

Making Linux Understand Super-Nodes: Technical Anatomy of the Lingqu Kernel Layer

Lingqu (UnifiedBus) super-nodes require systemic changes to the Linux kernel: a new bus type, cross-node address translation, unified memory management, and URMA communication…

AI基础设施
Read essay
Thinking12 min read

Breaking the Transceiver Bottleneck: How Optical Shuffle Reshapes AI Cluster Economics

Panduit engineer Castro proposes Optical Shuffle at IEEE 802.3, cutting 32K-GPU cluster transceivers by 33% and spine switches by 75%. Orthogonal to AWS RNG, three-layer stacking…

AI
Read essay
Thinking18 min read

A New Direction for Data Center Networking: What RNG Opens Up

In April 2026, AWS switched all new non-GPU datacenters to a flat topology called RNG. 69% fewer routers, up to 33% better throughput. Not an experiment — the production default.…

AI
Read essay
Thinking18 min read

The Great Token Retreat: When AI Bills Get More Expensive Than People

Uber burned through its annual AI budget in four months. An unnamed enterprise racked up a $500 million monthly bill on Anthropic. Klarna replaced 700 humans with AI, then…

thinking
Read essay
Thinking41 min read

NVIDIA Rubin Respins: Is AMD GPU Competitiveness for Real?

Fubon Research reveals Rubin was respun due to MI450 pressure. Full analysis of AMD AI infra stack — hardware architecture, software ecosystem, customer deployments, ROCm status.

AI
Read essay
Thinking53 min read

Build 2026: Microsoft's Agent OS Gambit

Windows is shifting from "an OS that runs apps" to "a platform that runs agents"—this isn't just a change in technical direction; it's a trillion-dollar company redefining its…

AI
Read essay
Thinking52 min read

From CLOS to ZCube: Network Topology Evolution for AI Computing Clusters

From Charles Clos's non-blocking telephone switching network in 1953 to ByteDance's SIGCOMM 2025 Best Paper ZCube — topology design has evolved from expert intuition to automated…

AI
Read essay
Thinking52 min read

MRC: When the NIC Becomes the Brain of the Network

OpenAI, together with NVIDIA, AMD, Broadcom, Arista, and Cisco, overturned five long-standing data center networking conventions simultaneously with the MRC protocol. By pushing…

AI
Read essay
Thinking24 min read

The PC, Reinvented: Computex 2026 and NVIDIA's Infrastructure Ambitions

At GTC Taipei, Jensen Huang unveiled RTX Spark—33 years of technology distilled into a single chip, officially marking the PC's entry into the agent era. Full-volume production…

NVIDIA
Read essay