Architects' Tech Alliance
Author

Architects' Tech Alliance

Sharing project experiences, insights into cutting-edge architectures, focusing on cloud computing, microservices, big data, hyper-convergence, storage, data protection, artificial intelligence, industry practices and solutions.

2.3k
Articles
0
Likes
12.2k
Views
0
Comments
Recent Articles

Latest from Architects' Tech Alliance

100 recent articles max
Architects' Tech Alliance
Architects' Tech Alliance
Sep 8, 2026 · Industry Insights

Huawei's Tao Law: Logic Folding Delivers 55% Density Gain Without Process Shrink in Kirin 9050 Pro

Huawei's Kirin 9050 Pro implements the new Tao Law, replacing geometric scaling with logic folding to achieve a 55% transistor density increase on the same process node, while cutting NPU power 66%, GPU 58%, and CPU core power 41% by using saved timing margin to lower voltage instead of raising frequency.

3D ICHuaweiKirin 9050 Pro
0 likes · 7 min read
Huawei's Tao Law: Logic Folding Delivers 55% Density Gain Without Process Shrink in Kirin 9050 Pro
Architects' Tech Alliance
Architects' Tech Alliance
Sep 6, 2026 · Artificial Intelligence

xDeepServe on CloudMatrix384: Full-Stack Design for Large-Scale MoE Model Serving

This article details Huawei's xDeepServe system for deploying massive Mixture-of-Experts models on the CloudMatrix384 supernode, covering the XCCL communication library, FlowServe decentralized serving engine, Transformerless execution architecture with prefill-decode and MoE-Attention decoupling, and hierarchical fault tolerance, achieving 2400 tokens/s per chip at 50ms TPOT.

Ascend 910CCloudMatrix384EPLB
0 likes · 18 min read
xDeepServe on CloudMatrix384: Full-Stack Design for Large-Scale MoE Model Serving
Architects' Tech Alliance
Architects' Tech Alliance
Sep 5, 2026 · Artificial Intelligence

CloudMatrix384: Co-Designing Supernodes for Trillion-Parameter LLM Inference

Huawei's CloudMatrix384 integrates 384 Ascend 910 NPUs with a unified bus network and decoupled PDC service architecture to achieve 4.45 tokens/s/TFLOPS prefill efficiency on DeepSeek-R1 671B, outperforming H100 baselines through fused MoE communication, INT8 quantization, and heterogeneous pipelining.

Ascend 910CloudMatrix384Expert Parallelism
0 likes · 20 min read
CloudMatrix384: Co-Designing Supernodes for Trillion-Parameter LLM Inference
Architects' Tech Alliance
Architects' Tech Alliance
Sep 1, 2026 · Artificial Intelligence

Understanding SuperNode, SuperPoD, and SuperCluster: Classification and Technology Evolution

The article analyzes the definition, technical traits, and three classifications of supernodes (SuperNode, SuperPoD, SuperCluster), explains how they boost large‑model training and inference efficiency, and examines the evolution of domestic supernode architectures, interconnects, and ecosystem design.

AI trainingGPU interconnectSuperCluster
0 likes · 9 min read
Understanding SuperNode, SuperPoD, and SuperCluster: Classification and Technology Evolution
Architects' Tech Alliance
Architects' Tech Alliance
Sep 1, 2026 · Industry Insights

$1.4 Trillion Valuation: China’s Unicorn List Highlights Beijing’s Four Semiconductor Leaders

The GEI China Unicorn Enterprise Research Report 2026 shows that China’s unicorns reached a total valuation of over $1.4 trillion in 2025, with more than 80 % in hard‑tech fields and Beijing’s four semiconductor unicorns—Qingwei Intelligent, Kunlun Chip, SMIC Beijing and Yiswei Computing—accounting for over 1300 billion RMB and leading the industry’s shift from internet to advanced chip technologies.

AIChinaRISC-V
0 likes · 4 min read
$1.4 Trillion Valuation: China’s Unicorn List Highlights Beijing’s Four Semiconductor Leaders
Architects' Tech Alliance
Architects' Tech Alliance
Sep 1, 2026 · Industry Insights

HaiGuang Secures China Patent Gold for Trusted Computing Platform

China’s National Intellectual Property Administration awarded HaiGuang Information’s patent on a virtual trusted environment loading and execution method the Gold Prize, highlighting the company’s 2026 R&D spend of 2.652 billion yuan, its 1,388 active patents, and the security features that raise industry entry barriers and drive commercial adoption across cloud, finance, and industrial sectors.

Cloud ComputingHaiGuangHardware Security
0 likes · 6 min read
HaiGuang Secures China Patent Gold for Trusted Computing Platform
Architects' Tech Alliance
Architects' Tech Alliance
Aug 30, 2026 · Industry Insights

High Bandwidth Flash (HBF): A Complete AI Storage Hierarchy Analysis

The article examines how AI models expose storage bottlenecks, introduces High Bandwidth Flash (HBF) as a 3D NAND‑based, HBM‑like stacked memory that combines subarray parallelism and a dedicated logic die, evaluates its technical principles, industry positioning, benchmark results showing 2.69× performance‑per‑watt gains, supply‑chain landscape, emerging standards, and the challenges that must be solved before commercial deployment.

AI storageCBA architectureHBM
0 likes · 10 min read
High Bandwidth Flash (HBF): A Complete AI Storage Hierarchy Analysis